Yikun Ban
Yikun Ban on Hugging Face Daily Papers: 6 papers, 4 in the top 3 of their day, 547 upvotes.
- Heterogeneous Agent Collaborative Reinforcement Learning 170 upvotes, #1 of 2026-03-05
- Does Your Reasoning Model Implicitly Know When to Stop Thinking? 253 upvotes, #1 of 2026-02-23
- Weak-Driven Learning: How Weak Agents make Strong Agents Stronger 254 upvotes, #1 of 2026-02-10
- Real-Time Aligned Reward Model beyond Semantics 4 upvotes, #28 of 2026-02-02
- Your Group-Relative Advantage Is Biased 144 upvotes, #1 of 2026-01-19
- Transformer Copilot: Learning from The Mistake Log in LLM Fine-tuning 6 upvotes, #31 of 2025-05-26
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.