Zixian Ma
Zixian Ma on Hugging Face Daily Papers: 8 papers, 1 in the top 3 of their day, 174 upvotes.
- VFIG: Vectorizing Complex Figures in SVG with Vision-Language Models 18 upvotes, #10 of 2026-03-27
- Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding 26 upvotes, #15 of 2026-01-16
- SAGE: Training Smart Any-Horizon Agents for Long Video Reasoning with Reinforcement Learning 16 upvotes, #12 of 2025-12-18
- Reinforced Visual Perception with Tools 29 upvotes, #6 of 2025-09-09
- Explain Before You Answer: A Survey on Compositional Visual Reasoning 3 upvotes, #18 of 2025-08-26
- Unfolding Spatial Cognition: Evaluating Multimodal Models on Visual Simulations 18 upvotes, #14 of 2025-06-06
- NaturalBench: Evaluating Vision-Language Models on Natural Adversarial Samples 35 upvotes, #3 of 2024-10-21
- Task Me Anything 7 upvotes, #20 of 2024-06-18
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.