Cong Wei
Cong Wei on Hugging Face Daily Papers: 13 papers, 7 in the top 3 of their day, 1,145 upvotes.
- PixelUMM: Encoder-Free Unified Image and Video Understanding and Generation 31 upvotes, #28 of 2026-10-01
- VGI-BENCH: Probing Visual Intelligence in Video Generation Models 177 upvotes, #2 of 2026-08-27
- Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models 107 upvotes, #1 of 2026-07-15
- Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation 86 upvotes, #2 of 2026-07-15
- Beyond Semantic Similarity: Rethinking Retrieval for Agentic Search via Direct Corpus Interaction 105 upvotes, #2 of 2026-05-08
- RationalRewards: Reasoning Rewards Scale Visual Generation Both Training and Test Time 100 upvotes, #2 of 2026-04-16
- Context Forcing: Consistent Autoregressive Video Generation with Long Context 35 upvotes, #6 of 2026-02-06
- UniVideo: Unified Understanding, Generation, and Editing for Videos 64 upvotes, #5 of 2025-10-10
- MoCha: Towards Movie-Grade Talking Character Synthesis 103 upvotes, #1 of 2025-04-01
- VISTA: Enhancing Long-Duration and High-Resolution Video Understanding by Video Spatiotemporal Augmentation 24 upvotes, #5 of 2024-12-03
- OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision 42 upvotes, #2 of 2024-11-12
- AnyV2V: A Plug-and-Play Framework For Any Video-to-Video Editing Tasks 17 upvotes, #4 of 2024-03-22
- ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation 27 upvotes, #6 of 2024-02-08
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.