Haoqin Tu

Haoqin Tu on Hugging Face Daily Papers: 10 papers, 2 in the top 3 of their day, 275 upvotes.

  1. VisualClaw: A Real-Time, Personalized Agent for the Physical World 28 upvotes, #9 of 2026-06-16
  2. AHELM: A Holistic Evaluation of Audio-Language Models 9 upvotes, #11 of 2025-09-01
  3. OpenVision: A Fully-Open, Cost-Effective Family of Advanced Vision Encoders for Multimodal Learning 20 upvotes, #7 of 2025-05-08
  4. SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models 26 upvotes, #5 of 2025-04-17
  5. ViLBench: A Suite for Vision-Language Process Reward Modeling 7 upvotes, #14 of 2025-03-27
  6. VHELM: A Holistic Evaluation of Vision Language Models 2 upvotes, #47 of 2024-10-10
  7. A Preliminary Study of o1 in Medicine: Are We Closer to an AI Doctor? 33 upvotes, #2 of 2024-09-24
  8. MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation? 49 upvotes, #1 of 2024-07-09
  9. What If We Recaption Billions of Web Images with LLaMA-3? 35 upvotes, #4 of 2024-06-13
  10. Eagle and Finch: RWKV with Matrix-Valued States and Dynamic Recurrence 23 upvotes, #4 of 2024-04-10

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.