Yifei Zhou

Yifei Zhou on Hugging Face Daily Papers: 3 papers, 0 in the top 3 of their day, 78 upvotes.

  1. Learning Adaptive Parallel Reasoning with Language Models 42 upvotes, #5 of 2025-04-23
  2. SWEET-RL: Training Multi-Turn LLM Agents on Collaborative Reasoning Tasks 8 upvotes, #16 of 2025-03-20
  3. DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning 17 upvotes, #11 of 2024-06-21

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.