Donghao Zhou

Donghao Zhou on Hugging Face Daily Papers: 14 papers, 1 in the top 3 of their day, 531 upvotes.

  1. ThinkV2V: Unleashing the Reasoning Capability of MLLMs for Instruction-Guided Video Editing 31 upvotes, #26 of 2026-10-01
  2. PackLab: A Comprehensive Framework for Developing, Training, and Evaluating MLLMs in Robotic Bin Packing 14 upvotes, #15 of 2026-09-24
  3. AgenticGen: Reward-Guided Agentic Video Generation for Advertising 9 upvotes, #23 of 2026-09-10
  4. OpenWAM: An Open, Modular Exploration Towards Systematic World-Action Model Pretraining 77 upvotes, #5 of 2026-09-09
  5. The Missing Temporal Link: Temporal Context Routing for Script-Driven Audio-Video Generation 35 upvotes, #14 of 2026-09-04
  6. Orchestra-o1: Omnimodal Agent Orchestration 45 upvotes, #7 of 2026-06-15
  7. SkillEvolBench: Benchmarking the Evolution from Episodic Experience to Procedural Skills 20 upvotes, #16 of 2026-05-26
  8. OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation 69 upvotes, #5 of 2026-04-14
  9. HiFi-Inpaint: Towards High-Fidelity Reference-Based Inpainting for Generating Detail-Preserving Human-Product Images 28 upvotes, #6 of 2026-03-06
  10. DSDR: Dual-Scale Diversity Regularization for Exploration in LLM Reasoning 13 upvotes, #10 of 2026-02-24
  11. MMDeepResearch-Bench: A Benchmark for Multimodal Deep Research Agents 49 upvotes, #2 of 2026-01-22
  12. StereoPilot: Learning Unified and Efficient Stereo Conversion via Generative Priors 37 upvotes, #6 of 2025-12-19
  13. An Empirical Study of GPT-4o Image Generation Capabilities 59 upvotes, #4 of 2025-04-09
  14. MagicTailor: Component-Controllable Personalization in Text-to-Image Diffusion Models 33 upvotes, #4 of 2024-10-21

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.