Yifei Zhou
Yifei Zhou on Hugging Face Daily Papers: 3 papers, 0 in the top 3 of their day, 78 upvotes.
- Learning Adaptive Parallel Reasoning with Language Models 42 upvotes, #5 of 2025-04-23
- SWEET-RL: Training Multi-Turn LLM Agents on Collaborative Reasoning Tasks 8 upvotes, #16 of 2025-03-20
- DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning 17 upvotes, #11 of 2024-06-21
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.