Zixian Ma

Zixian Ma on Hugging Face Daily Papers: 8 papers, 1 in the top 3 of their day, 174 upvotes.

  1. VFIG: Vectorizing Complex Figures in SVG with Vision-Language Models 18 upvotes, #10 of 2026-03-27
  2. Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding 26 upvotes, #15 of 2026-01-16
  3. SAGE: Training Smart Any-Horizon Agents for Long Video Reasoning with Reinforcement Learning 16 upvotes, #12 of 2025-12-18
  4. Reinforced Visual Perception with Tools 29 upvotes, #6 of 2025-09-09
  5. Explain Before You Answer: A Survey on Compositional Visual Reasoning 3 upvotes, #18 of 2025-08-26
  6. Unfolding Spatial Cognition: Evaluating Multimodal Models on Visual Simulations 18 upvotes, #14 of 2025-06-06
  7. NaturalBench: Evaluating Vision-Language Models on Natural Adversarial Samples 35 upvotes, #3 of 2024-10-21
  8. Task Me Anything 7 upvotes, #20 of 2024-06-18

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.