Zekun Qi

Zekun Qi on Hugging Face Daily Papers: 12 papers, 2 in the top 3 of their day, 361 upvotes.

  1. ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing? 19 upvotes, #11 of 2026-06-19
  2. LIMMT: Less is More for Motion Tracking 16 upvotes, #16 of 2026-06-08
  3. Humanoid-GPT: Scaling Data and Structure for Zero-Shot Motion Tracking 38 upvotes, #5 of 2026-06-03
  4. Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining 2 upvotes, #25 of 2026-04-28
  5. VLA-JEPA: Enhancing Vision-Language-Action Model with Latent World Model 17 upvotes, #18 of 2026-02-11
  6. Reasoning in Space via Grounding in the World 14 upvotes, #15 of 2025-10-16
  7. Hybrid-grained Feature Aggregation with Coarse-to-fine Language Guidance for Self-supervised Monocular Depth Estimation 1 upvotes, #43 of 2025-10-13
  8. DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge 37 upvotes, #5 of 2025-07-08
  9. OmniSpatial: Towards Comprehensive Spatial Reasoning Benchmark for Vision Language Models 37 upvotes, #7 of 2025-06-04
  10. SoFar: Language-Grounded Orientation Bridges Spatial Reasoning and Object Manipulation 29 upvotes, #10 of 2025-02-19
  11. DreamBench++: A Human-Aligned Benchmark for Personalized Image Generation 53 upvotes, #1 of 2024-06-25
  12. DreamLLM: Synergistic Multimodal Comprehension and Creation 60 upvotes, #3 of 2023-09-21

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.