Martin Ziqiao Ma

Martin Ziqiao Ma on Hugging Face Daily Papers: 9 papers, 2 in the top 3 of their day, 296 upvotes.

  1. Next-Embedding Prediction Makes Strong Vision Learners 78 upvotes, #3 of 2025-12-19
  2. SimWorld: An Open-ended Realistic Simulator for Autonomous Agents in Physical and Social Worlds 33 upvotes, #10 of 2025-12-03
  3. ROVER: Benchmarking Reciprocal Cross-Modal Reasoning for Omnimodal Generation 31 upvotes, #8 of 2025-11-04
  4. AimBot: A Simple Auxiliary Visual Cue to Enhance Spatial Awareness of Visuomotor Policies 9 upvotes, #19 of 2025-08-13
  5. Do Vision-Language Models Have Internal World Models? Towards an Atomic Evaluation 27 upvotes, #4 of 2025-06-30
  6. 4D-LRM: Large Space-Time Reconstruction Model From and To Any View at Any Time 5 upvotes, #27 of 2025-06-24
  7. Can Vision Language Models Infer Human Gaze Direction? A Controlled Study 4 upvotes, #19 of 2025-06-12
  8. Humanity's Last Exam 50 upvotes, #1 of 2025-01-27
  9. Multi-Object Hallucination in Vision-Language Models 7 upvotes, #12 of 2024-07-09

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.