Limin Wang
Limin Wang on Hugging Face Daily Papers: 11 papers, 6 in the top 3 of their day, 320 upvotes.
- UniAVGen: Unified Audio and Video Generation with Asymmetric Cross-Modal Interactions 50 upvotes, #2 of 2025-11-06
- Eagle 2.5: Boosting Long-Context Post-Training for Frontier Vision-Language Models 65 upvotes, #2 of 2025-04-22
- DDT: Decoupled Diffusion Transformer 69 upvotes, #1 of 2025-04-10
- LeviTor: 3D Trajectory Oriented Image-to-Video Synthesis 14 upvotes, #8 of 2024-12-20
- InternVideo2: Scaling Video Foundation Models for Multimodal Video Understanding 14 upvotes, #3 of 2024-03-25
- Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding 10 upvotes, #9 of 2024-03-15
- VideoMamba: State Space Model for Efficient Video Understanding 20 upvotes, #5 of 2024-03-12
- Scaffold-GS: Structured 3D Gaussians for View-Adaptive Rendering 11 upvotes, #13 of 2023-12-04
- InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation 26 upvotes, #3 of 2023-07-14
- VideoChat: Chat-Centric Video Understanding 3 upvotes, #3 of 2023-05-11
- InternChat: Solving Vision-Centric Tasks by Interacting with Chatbots Beyond Language 5 upvotes, #5 of 2023-05-10
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.