Zheng Zhu
Zheng Zhu on Hugging Face Daily Papers: 11 papers, 1 in the top 3 of their day, 204 upvotes.
- SwiftVLA: Unlocking Spatiotemporal Dynamics for Lightweight VLA Models at Minimal Overhead 5 upvotes, #28 of 2025-12-03
- GigaWorld-0: World Models as Data Engine to Empower Embodied AI 30 upvotes, #9 of 2025-11-26
- GigaBrain-0: A World Model-Powered Vision-Language-Action Model 42 upvotes, #6 of 2025-10-23
- DriveGen3D: Boosting Feed-Forward Driving Scene Generation with Efficient Video Diffusion 1 upvotes, #28 of 2025-10-20
- R2RGEN: Real-to-Real 3D Data Generation for Spatially Generalized Manipulation 4 upvotes, #40 of 2025-10-10
- VLA-R1: Enhancing Reasoning in Vision-Language-Action Models 7 upvotes, #28 of 2025-10-03
- VolSplat: Rethinking Feed-Forward 3D Gaussian Splatting with Voxel-Aligned Prediction 23 upvotes, #8 of 2025-09-24
- HumanDreamer-X: Photorealistic Single-image Human Avatars Reconstruction via Gaussian Restoration 10 upvotes, #10 of 2025-04-07
- EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation 24 upvotes, #2 of 2024-11-14
- WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens 17 upvotes, #6 of 2024-01-19
- On the Road with GPT-4V(ision): Early Explorations of Visual-Language Model on Autonomous Driving 10 upvotes, #6 of 2023-11-10
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.