Zhang
Zhang on Hugging Face Daily Papers: 10 papers, 0 in the top 3 of their day, 244 upvotes.
- ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing? 19 upvotes, #11 of 2026-06-19
- LIMMT: Less is More for Motion Tracking 16 upvotes, #16 of 2026-06-08
- Humanoid-GPT: Scaling Data and Structure for Zero-Shot Motion Tracking 38 upvotes, #5 of 2026-06-03
- Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining 2 upvotes, #25 of 2026-04-28
- VLA-JEPA: Enhancing Vision-Language-Action Model with Latent World Model 17 upvotes, #18 of 2026-02-11
- Reasoning in Space via Grounding in the World 14 upvotes, #15 of 2025-10-16
- Hybrid-grained Feature Aggregation with Coarse-to-fine Language Guidance for Self-supervised Monocular Depth Estimation 1 upvotes, #43 of 2025-10-13
- DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge 37 upvotes, #5 of 2025-07-08
- OmniSpatial: Towards Comprehensive Spatial Reasoning Benchmark for Vision Language Models 37 upvotes, #7 of 2025-06-04
- SoFar: Language-Grounded Orientation Bridges Spatial Reasoning and Object Manipulation 29 upvotes, #10 of 2025-02-19
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.