Zhu
Zhu on Hugging Face Daily Papers: 11 papers, 1 in the top 3 of their day, 493 upvotes.
- HiRAE: Hierarchical Representation Autoencoding with Residual Budgets 25 upvotes, #38 of 2026-09-30
- Think Before You Score: Thinking Reward Model for Visual Generation 101 upvotes, #11 of 2026-09-30
- RefCaptioner: Multi-Reference Image-Grounded Video Captioning 30 upvotes, #14 of 2026-07-31
- KeyFrame-Compass: Towards Comprehensive Evaluation of Keyframe-Conditioned Video Generation 42 upvotes, #6 of 2026-07-17
- LongAV-Compass: Towards Unified Evaluation of Minute-Scale Audio-Visual Generation Across T2AV, I2AV, and V2AV 38 upvotes, #8 of 2026-05-27
- Artifact-Bench: Evaluating MLLMs on Detecting and Assessing the Artifacts of AI-Generated Videos 22 upvotes, #12 of 2026-05-20
- Edit-Compass & EditReward-Compass: A Unified Benchmark for Image Editing and Reward Modeling 32 upvotes, #9 of 2026-05-14
- Beyond the Last Layer: Multi-Layer Representation Fusion for Visual Tokenization 33 upvotes, #10 of 2026-05-13
- LongCat-Next: Lexicalizing Modalities as Discrete Tokens 137 upvotes, #3 of 2026-04-01
- VTC-Bench: Evaluating Agentic Multimodal Models via Compositional Visual Tool Chaining 21 upvotes, #14 of 2026-03-20
- RealUnify: Do Unified Models Truly Benefit from Unification? A Comprehensive Benchmark 44 upvotes, #6 of 2025-09-30
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.