Xinyu Fang

Xinyu Fang on Hugging Face Daily Papers: 11 papers, 1 in the top 3 of their day, 342 upvotes.

  1. ARM-Thinker: Reinforcing Multimodal Generative Reward Models with Agentic Tool Use and Visual Reasoning 45 upvotes, #5 of 2025-12-05
  2. ATLAS: A High-Difficulty, Multidisciplinary Benchmark for Frontier Scientific Reasoning 14 upvotes, #12 of 2025-11-19
  3. IWR-Bench: Can LVLMs reconstruct interactive webpage from a user interaction video? 3 upvotes, #64 of 2025-09-30
  4. Creation-MMBench: Assessing Context-Aware Creative Intelligence in MLLM 41 upvotes, #4 of 2025-03-19
  5. OmniAlign-V: Towards Enhanced Alignment of MLLMs with Human Preference 67 upvotes, #1 of 2025-02-26
  6. Redundancy Principles for MLLMs Benchmarks 25 upvotes, #4 of 2025-01-27
  7. MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs 18 upvotes, #4 of 2024-11-27
  8. ProSA: Assessing and Understanding the Prompt Sensitivity of LLMs 13 upvotes, #8 of 2024-10-17
  9. VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models 11 upvotes, #6 of 2024-07-17
  10. MMBench-Video: A Long-Form Multi-Shot Benchmark for Holistic Video Understanding 27 upvotes, #7 of 2024-06-21
  11. Prism: A Framework for Decoupling and Assessing the Capabilities of VLMs 33 upvotes, #5 of 2024-06-21

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.