Haoqin Tu
Haoqin Tu on Hugging Face Daily Papers: 10 papers, 2 in the top 3 of their day, 275 upvotes.
- VisualClaw: A Real-Time, Personalized Agent for the Physical World 28 upvotes, #9 of 2026-06-16
- AHELM: A Holistic Evaluation of Audio-Language Models 9 upvotes, #11 of 2025-09-01
- OpenVision: A Fully-Open, Cost-Effective Family of Advanced Vision Encoders for Multimodal Learning 20 upvotes, #7 of 2025-05-08
- SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models 26 upvotes, #5 of 2025-04-17
- ViLBench: A Suite for Vision-Language Process Reward Modeling 7 upvotes, #14 of 2025-03-27
- VHELM: A Holistic Evaluation of Vision Language Models 2 upvotes, #47 of 2024-10-10
- A Preliminary Study of o1 in Medicine: Are We Closer to an AI Doctor? 33 upvotes, #2 of 2024-09-24
- MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation? 49 upvotes, #1 of 2024-07-09
- What If We Recaption Billions of Web Images with LLaMA-3? 35 upvotes, #4 of 2024-06-13
- Eagle and Finch: RWKV with Matrix-Valued States and Dynamic Recurrence 23 upvotes, #4 of 2024-04-10
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.