Xiaoyu Tan
Xiaoyu Tan on Hugging Face Daily Papers: 8 papers, 2 in the top 3 of their day, 428 upvotes.
- Youtu-VL: Unleashing Visual Potential via Unified Vision-Language Supervision 41 upvotes, #4 of 2026-01-28
- Youtu-Agent: Scaling Agent Productivity with Automated Generation and Hybrid Policy Optimization 108 upvotes, #2 of 2026-01-05
- Youtu-LLM: Unlocking the Native Agentic Potential for Lightweight Large Language Models 119 upvotes, #2 of 2026-01-01
- SmartSnap: Proactive Evidence Seeking for Self-Verifying Agents 38 upvotes, #7 of 2025-12-30
- Training-Free Group Relative Policy Optimization 40 upvotes, #11 of 2025-10-10
- Learn the Ropes, Then Trust the Wins: Self-imitation with Progressive Exploration for Agentic Reinforcement Learning 10 upvotes, #26 of 2025-09-29
- The Choice of Divergence: A Neglected Key to Mitigating Diversity Collapse in Reinforcement Learning with Verifiable Reward 3 upvotes, #19 of 2025-09-12
- One Example Shown, Many Concepts Known! Counterexample-Driven Conceptual Reasoning in Mathematical LLMs 7 upvotes, #20 of 2025-02-18
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.