yfdeng

yfdeng on Hugging Face Daily Papers: 15 papers, 2 in the top 3 of their day, 431 upvotes.

  1. PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation 50 upvotes, #2 of 2026-06-29
  2. HumanScale: Egocentric Human Video Can Outperform Real-Robot Data for Embodied Pretraining 13 upvotes, #15 of 2026-06-19
  3. StableVLA: Towards Robust Vision-Language-Action Models without Extra Data 15 upvotes, #18 of 2026-05-19
  4. HumanNet: Scaling Human-centric Video Learning to One Million Hours 51 upvotes, #7 of 2026-05-11
  5. Enhancing Spatial Understanding in Image Generation via Reward Modeling 50 upvotes, #3 of 2026-03-02
  6. Rethinking Video Generation Model for the Embodied World 42 upvotes, #4 of 2026-01-22
  7. Focal Guidance: Unlocking Controllability from Semantic-Weak Layers in Video Diffusion Models 4 upvotes, #20 of 2026-01-15
  8. MHLA: Restoring Expressivity of Linear Attention via Token-Level Multi-Head 46 upvotes, #4 of 2026-01-13
  9. MAGREF: Masked Guidance for Any-Reference Video Generation 9 upvotes, #33 of 2025-05-30
  10. OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation 52 upvotes, #8 of 2025-05-28
  11. VARGPT-v1.1: Improve Visual Autoregressive Large Unified Model via Iterative Instruction Tuning and Reinforcement Learning 18 upvotes, #4 of 2025-04-07
  12. MagicComp: Training-free Dual-Phase Refinement for Compositional Video Generation 8 upvotes, #21 of 2025-03-25
  13. CINEMA: Coherent Multi-Subject Video Generation via MLLM-Based Guidance 10 upvotes, #24 of 2025-03-14
  14. VideoTetris: Towards Compositional Text-to-Video Generation 18 upvotes, #6 of 2024-06-07
  15. I2V-Adapter: A General Image-to-Video Adapter for Video Diffusion Models 14 upvotes, #9 of 2023-12-29

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.