Jiashuo Yu

Jiashuo Yu on Hugging Face Daily Papers: 10 papers, 5 in the top 3 of their day, 487 upvotes.

  1. ExpVid: A Benchmark for Experiment Video Understanding & Reasoning 3 upvotes, #32 of 2025-10-15
  2. Intern-S1: A Scientific Multimodal Foundation Model 242 upvotes, #1 of 2025-08-22
  3. VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos 31 upvotes, #6 of 2025-06-13
  4. VBench++: Comprehensive and Versatile Benchmark Suite for Video Generative Models 28 upvotes, #2 of 2024-11-21
  5. OmniCorpus: A Unified Multimodal Corpus of 10 Billion-Level Images Interleaved with Text 28 upvotes, #6 of 2024-06-17
  6. InternVideo2: Scaling Video Foundation Models for Multimodal Video Understanding 14 upvotes, #3 of 2024-03-25
  7. SEINE: Short-to-Long Video Diffusion Model for Generative Transition and Prediction 10 upvotes, #9 of 2023-11-01
  8. LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models 43 upvotes, #2 of 2023-09-27
  9. InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation 26 upvotes, #3 of 2023-07-14
  10. InternChat: Solving Vision-Centric Tasks by Interacting with Chatbots Beyond Language 5 upvotes, #5 of 2023-05-10

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.