Zheng Zhu

Zheng Zhu on Hugging Face Daily Papers: 11 papers, 1 in the top 3 of their day, 204 upvotes.

  1. SwiftVLA: Unlocking Spatiotemporal Dynamics for Lightweight VLA Models at Minimal Overhead 5 upvotes, #28 of 2025-12-03
  2. GigaWorld-0: World Models as Data Engine to Empower Embodied AI 30 upvotes, #9 of 2025-11-26
  3. GigaBrain-0: A World Model-Powered Vision-Language-Action Model 42 upvotes, #6 of 2025-10-23
  4. DriveGen3D: Boosting Feed-Forward Driving Scene Generation with Efficient Video Diffusion 1 upvotes, #28 of 2025-10-20
  5. R2RGEN: Real-to-Real 3D Data Generation for Spatially Generalized Manipulation 4 upvotes, #40 of 2025-10-10
  6. VLA-R1: Enhancing Reasoning in Vision-Language-Action Models 7 upvotes, #28 of 2025-10-03
  7. VolSplat: Rethinking Feed-Forward 3D Gaussian Splatting with Voxel-Aligned Prediction 23 upvotes, #8 of 2025-09-24
  8. HumanDreamer-X: Photorealistic Single-image Human Avatars Reconstruction via Gaussian Restoration 10 upvotes, #10 of 2025-04-07
  9. EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation 24 upvotes, #2 of 2024-11-14
  10. WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens 17 upvotes, #6 of 2024-01-19
  11. On the Road with GPT-4V(ision): Early Explorations of Visual-Language Model on Autonomous Driving 10 upvotes, #6 of 2023-11-10

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.