Yanhong Zeng
Yanhong Zeng on Hugging Face Daily Papers: 13 papers, 2 in the top 3 of their day, 365 upvotes.
- MagicQuillV2: Precise and Interactive Image Editing with Layered Visual Cues 11 upvotes, #21 of 2025-12-03
- HoloCine: Holistic Generation of Cinematic Multi-Shot Long Video Narratives 38 upvotes, #4 of 2025-10-24
- Scaling Instruction-Based Video Editing with a High-Quality Synthetic Dataset 49 upvotes, #4 of 2025-10-20
- CharacterShot: Controllable and Consistent 4D Character Animation 37 upvotes, #5 of 2025-08-13
- WORLDMEM: Long-term Consistent World Simulation with Memory 30 upvotes, #6 of 2025-04-18
- LEGO-Puzzles: How Good Are MLLMs at Multi-Step Spatial Reasoning? 31 upvotes, #6 of 2025-03-27
- DiffSensei: Bridging Multi-Modal LLMs and Diffusion Models for Customized Manga Generation 43 upvotes, #3 of 2024-12-11
- HumanVid: Demystifying Training Data for Camera-controllable Human Image Animation 21 upvotes, #3 of 2024-07-25
- Live2Diff: Live Stream Translation via Uni-directional Attention in Video Diffusion Models 8 upvotes, #15 of 2024-07-12
- FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds 10 upvotes, #11 of 2024-07-03
- Auto Cherry-Picker: Learning from High-quality Generative Data Driven by Language 9 upvotes, #16 of 2024-07-02
- MotionBooth: Motion-Aware Customized Text-to-Video Generation 17 upvotes, #8 of 2024-06-26
- PIA: Your Personalized Image Animator via Plug-and-Play Modules in Text-to-Image Models 19 upvotes, #8 of 2023-12-22
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.