Donghao Zhou
Donghao Zhou on Hugging Face Daily Papers: 14 papers, 1 in the top 3 of their day, 531 upvotes.
- ThinkV2V: Unleashing the Reasoning Capability of MLLMs for Instruction-Guided Video Editing 31 upvotes, #26 of 2026-10-01
- PackLab: A Comprehensive Framework for Developing, Training, and Evaluating MLLMs in Robotic Bin Packing 14 upvotes, #15 of 2026-09-24
- AgenticGen: Reward-Guided Agentic Video Generation for Advertising 9 upvotes, #23 of 2026-09-10
- OpenWAM: An Open, Modular Exploration Towards Systematic World-Action Model Pretraining 77 upvotes, #5 of 2026-09-09
- The Missing Temporal Link: Temporal Context Routing for Script-Driven Audio-Video Generation 35 upvotes, #14 of 2026-09-04
- Orchestra-o1: Omnimodal Agent Orchestration 45 upvotes, #7 of 2026-06-15
- SkillEvolBench: Benchmarking the Evolution from Episodic Experience to Procedural Skills 20 upvotes, #16 of 2026-05-26
- OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation 69 upvotes, #5 of 2026-04-14
- HiFi-Inpaint: Towards High-Fidelity Reference-Based Inpainting for Generating Detail-Preserving Human-Product Images 28 upvotes, #6 of 2026-03-06
- DSDR: Dual-Scale Diversity Regularization for Exploration in LLM Reasoning 13 upvotes, #10 of 2026-02-24
- MMDeepResearch-Bench: A Benchmark for Multimodal Deep Research Agents 49 upvotes, #2 of 2026-01-22
- StereoPilot: Learning Unified and Efficient Stereo Conversion via Generative Priors 37 upvotes, #6 of 2025-12-19
- An Empirical Study of GPT-4o Image Generation Capabilities 59 upvotes, #4 of 2025-04-09
- MagicTailor: Component-Controllable Personalization in Text-to-Image Diffusion Models 33 upvotes, #4 of 2024-10-21
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.