JeffWang
JeffWang on Hugging Face Daily Papers: 20 papers, 1 in the top 3 of their day, 411 upvotes.
- GigaWorld-Policy-0.5: A Faster and Stronger WAM Empowered by AutoResearch 28 upvotes, #10 of 2026-07-16
- GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation 37 upvotes, #8 of 2026-07-07
- iMaC: Translating Actions into Motion and Contact Images for Embodied World Models 13 upvotes, #19 of 2026-06-15
- ReconPhys: Reconstruct Appearance and Physical Attributes from Single Video 9 upvotes, #20 of 2026-04-16
- ViVa: A Video-Generative Value Model for Robot Reinforcement Learning 17 upvotes, #21 of 2026-04-10
- DriveDreamer-Policy: A Geometry-Grounded World-Action Model for Unified Generation and Planning 6 upvotes, #16 of 2026-04-06
- GigaWorld-Policy: An Efficient Action-Centered World--Action Model 24 upvotes, #9 of 2026-03-19
- π-StepNFT: Wider Space Needs Finer Steps in Online RL for Flow-based VLAs 9 upvotes, #10 of 2026-03-09
- GigaBrain-0.5M*: a VLA That Learns From World Model-Based Reinforcement Learning 55 upvotes, #5 of 2026-02-13
- SwiftVLA: Unlocking Spatiotemporal Dynamics for Lightweight VLA Models at Minimal Overhead 5 upvotes, #28 of 2025-12-03
- GigaWorld-0: World Models as Data Engine to Empower Embodied AI 30 upvotes, #9 of 2025-11-26
- GigaBrain-0: A World Model-Powered Vision-Language-Action Model 42 upvotes, #6 of 2025-10-23
- DriveGen3D: Boosting Feed-Forward Driving Scene Generation with Efficient Video Diffusion 1 upvotes, #28 of 2025-10-20
- VLA-R1: Enhancing Reasoning in Vision-Language-Action Models 7 upvotes, #28 of 2025-10-03
- A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment 13 upvotes, #11 of 2025-04-24
- HumanDreamer-X: Photorealistic Single-image Human Avatars Reconstruction via Gaussian Restoration 10 upvotes, #10 of 2025-04-07
- Timestep Embedding Tells: It's Time to Cache for Video Diffusion Model 15 upvotes, #9 of 2024-12-02
- EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation 24 upvotes, #2 of 2024-11-14
- WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens 17 upvotes, #6 of 2024-01-19
- On the Road with GPT-4V(ision): Early Explorations of Visual-Language Model on Autonomous Driving 10 upvotes, #6 of 2023-11-10
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.