Ziqi Huang

Ziqi Huang on Hugging Face Daily Papers: 25 papers, 10 in the top 3 of their day, 1,314 upvotes.

  1. Learning Native Reflection in Unified Models with Interleaved Reinforcement Learning 49 upvotes, #13 of 2026-09-29
  2. VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning 266 upvotes, #1 of 2026-08-27
  3. HarnessEval-W: Agentifying the Evaluation of Visual Worlds 123 upvotes, #3 of 2026-08-18
  4. Apple-π: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence 43 upvotes, #9 of 2026-07-21
  5. Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond 181 upvotes, #1 of 2026-04-27
  6. Prompt Relay: Inference-Time Temporal Control for Multi-Event Video Generation 15 upvotes, #18 of 2026-04-14
  7. Demystifing Video Reasoning 356 upvotes, #1 of 2026-03-18
  8. UniT: Unified Multimodal Chain-of-Thought Test-time Scaling 19 upvotes, #8 of 2026-02-18
  9. BabyVision: Visual Reasoning Beyond Language 184 upvotes, #2 of 2026-01-13
  10. Exploring MLLM-Diffusion Information Transfer with MetaCanvas 12 upvotes, #7 of 2025-12-15
  11. Simulating the Visual World with Artificial Intelligence: A Roadmap 28 upvotes, #6 of 2025-11-17
  12. The Quest for Generalizable Motion Generation: Data, Model, and Evaluation 27 upvotes, #10 of 2025-10-31
  13. RealDPO: Real or Not Real, that is the Preference 6 upvotes, #31 of 2025-10-17
  14. Uni-MMMU: A Massive Multi-discipline Multimodal Unified Benchmark 9 upvotes, #21 of 2025-10-16
  15. VChain: Chain-of-Visual-Thought for Reasoning in Video Generation 34 upvotes, #5 of 2025-10-07
  16. CineScale: Free Lunch in High-Resolution Cinematic Visual Generation 19 upvotes, #10 of 2025-08-27
  17. Cut2Next: Generating Next Shot via In-Context Tuning 12 upvotes, #16 of 2025-08-13
  18. ShotBench: Expert-Level Cinematic Understanding in Vision-Language Models 22 upvotes, #5 of 2025-06-30
  19. VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness 30 upvotes, #5 of 2025-03-28
  20. RepVideo: Rethinking Cross-Layer Representation for Video Generation 15 upvotes, #4 of 2025-01-16
  21. Evaluation Agent: Efficient and Promptable Evaluation Framework for Visual Generative Models 35 upvotes, #2 of 2024-12-17
  22. VBench++: Comprehensive and Versatile Benchmark Suite for Video Generative Models 28 upvotes, #2 of 2024-11-21
  23. FreeInit: Bridging Initialization Gap in Video Diffusion Models 26 upvotes, #1 of 2023-12-13
  24. LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models 43 upvotes, #2 of 2023-09-27
  25. FreeU: Free Lunch in Diffusion U-Net 66 upvotes, #2 of 2023-09-21

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.