Limin Wang

Limin Wang on Hugging Face Daily Papers: 11 papers, 6 in the top 3 of their day, 320 upvotes.

  1. UniAVGen: Unified Audio and Video Generation with Asymmetric Cross-Modal Interactions 50 upvotes, #2 of 2025-11-06
  2. Eagle 2.5: Boosting Long-Context Post-Training for Frontier Vision-Language Models 65 upvotes, #2 of 2025-04-22
  3. DDT: Decoupled Diffusion Transformer 69 upvotes, #1 of 2025-04-10
  4. LeviTor: 3D Trajectory Oriented Image-to-Video Synthesis 14 upvotes, #8 of 2024-12-20
  5. InternVideo2: Scaling Video Foundation Models for Multimodal Video Understanding 14 upvotes, #3 of 2024-03-25
  6. Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding 10 upvotes, #9 of 2024-03-15
  7. VideoMamba: State Space Model for Efficient Video Understanding 20 upvotes, #5 of 2024-03-12
  8. Scaffold-GS: Structured 3D Gaussians for View-Adaptive Rendering 11 upvotes, #13 of 2023-12-04
  9. InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation 26 upvotes, #3 of 2023-07-14
  10. VideoChat: Chat-Centric Video Understanding 3 upvotes, #3 of 2023-05-11
  11. InternChat: Solving Vision-Centric Tasks by Interacting with Chatbots Beyond Language 5 upvotes, #5 of 2023-05-10

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.