Zhongang Cai
Zhongang Cai on Hugging Face Daily Papers: 21 papers, 6 in the top 3 of their day, 1,429 upvotes.
- SenseNova-U1.5: Towards Native Unified Visual Intelligence 262 upvotes, #2 of 2026-09-11
- VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning 266 upvotes, #1 of 2026-08-27
- Apple-π: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence 43 upvotes, #9 of 2026-07-21
- Vision as Unified Multimodal Generation 46 upvotes, #6 of 2026-07-08
- From Pixels to Words -- Towards Native One-Vision Models at Scale 72 upvotes, #4 of 2026-05-28
- SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture 184 upvotes, #1 of 2026-05-13
- Bridging Semantic and Kinematic Conditions with Diffusion-based Discrete Motion Tokenizer 41 upvotes, #7 of 2026-03-20
- Demystifing Video Reasoning 356 upvotes, #1 of 2026-03-18
- A Very Big Video Reasoning Suite 503 upvotes, #1 of 2026-02-24
- Scaling Spatial Intelligence with Multimodal Foundation Models 41 upvotes, #6 of 2025-11-21
- The Quest for Generalizable Motion Generation: Data, Model, and Evaluation 27 upvotes, #10 of 2025-10-31
- Has GPT-5 Achieved Spatial Intelligence? An Empirical Study 31 upvotes, #8 of 2025-08-19
- DPoser-X: Diffusion Model as Robust 3D Whole-body Human Pose Prior 6 upvotes, #23 of 2025-08-07
- Controllable Human-centric Keyframe Interpolation with Generative Prior 2 upvotes, #43 of 2025-06-04
- EgoLife: Towards Egocentric Life Assistant 35 upvotes, #4 of 2025-03-07
- WHAC: World-grounded Humans and Cameras 2 upvotes, #26 of 2025-02-24
- SOLAMI: Social Vision-Language-Action Modeling for Immersive Interaction with 3D Autonomous Characters 21 upvotes, #6 of 2024-12-03
- Disco4D: Disentangled 4D Human Generation and Animation from a Single Image 8 upvotes, #9 of 2024-09-27
- MeshAnything: Artist-Created Mesh Generation with Autoregressive Transformers 21 upvotes, #6 of 2024-06-18
- Story-to-Motion: Synthesizing Infinite and Controllable Character Animation from Long Text 29 upvotes, #3 of 2023-11-14
- DNA-Rendering: A Diverse Neural Actor Repository for High-Fidelity Human-centric Rendering 6 upvotes, #9 of 2023-07-20
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.