yfdeng
yfdeng on Hugging Face Daily Papers: 15 papers, 2 in the top 3 of their day, 431 upvotes.
- PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation 50 upvotes, #2 of 2026-06-29
- HumanScale: Egocentric Human Video Can Outperform Real-Robot Data for Embodied Pretraining 13 upvotes, #15 of 2026-06-19
- StableVLA: Towards Robust Vision-Language-Action Models without Extra Data 15 upvotes, #18 of 2026-05-19
- HumanNet: Scaling Human-centric Video Learning to One Million Hours 51 upvotes, #7 of 2026-05-11
- Enhancing Spatial Understanding in Image Generation via Reward Modeling 50 upvotes, #3 of 2026-03-02
- Rethinking Video Generation Model for the Embodied World 42 upvotes, #4 of 2026-01-22
- Focal Guidance: Unlocking Controllability from Semantic-Weak Layers in Video Diffusion Models 4 upvotes, #20 of 2026-01-15
- MHLA: Restoring Expressivity of Linear Attention via Token-Level Multi-Head 46 upvotes, #4 of 2026-01-13
- MAGREF: Masked Guidance for Any-Reference Video Generation 9 upvotes, #33 of 2025-05-30
- OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation 52 upvotes, #8 of 2025-05-28
- VARGPT-v1.1: Improve Visual Autoregressive Large Unified Model via Iterative Instruction Tuning and Reinforcement Learning 18 upvotes, #4 of 2025-04-07
- MagicComp: Training-free Dual-Phase Refinement for Compositional Video Generation 8 upvotes, #21 of 2025-03-25
- CINEMA: Coherent Multi-Subject Video Generation via MLLM-Based Guidance 10 upvotes, #24 of 2025-03-14
- VideoTetris: Towards Compositional Text-to-Video Generation 18 upvotes, #6 of 2024-06-07
- I2V-Adapter: A General Image-to-Video Adapter for Video Diffusion Models 14 upvotes, #9 of 2023-12-29
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.