Jiashuo Yu
Jiashuo Yu on Hugging Face Daily Papers: 10 papers, 5 in the top 3 of their day, 487 upvotes.
- ExpVid: A Benchmark for Experiment Video Understanding & Reasoning 3 upvotes, #32 of 2025-10-15
- Intern-S1: A Scientific Multimodal Foundation Model 242 upvotes, #1 of 2025-08-22
- VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos 31 upvotes, #6 of 2025-06-13
- VBench++: Comprehensive and Versatile Benchmark Suite for Video Generative Models 28 upvotes, #2 of 2024-11-21
- OmniCorpus: A Unified Multimodal Corpus of 10 Billion-Level Images Interleaved with Text 28 upvotes, #6 of 2024-06-17
- InternVideo2: Scaling Video Foundation Models for Multimodal Video Understanding 14 upvotes, #3 of 2024-03-25
- SEINE: Short-to-Long Video Diffusion Model for Generative Transition and Prediction 10 upvotes, #9 of 2023-11-01
- LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models 43 upvotes, #2 of 2023-09-27
- InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation 26 upvotes, #3 of 2023-07-14
- InternChat: Solving Vision-Centric Tasks by Interacting with Chatbots Beyond Language 5 upvotes, #5 of 2023-05-10
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.