youyou

youyou on Hugging Face Daily Papers: 6 papers, 3 in the top 3 of their day, 380 upvotes.

  1. EvoTrainer: Co-Evolving LLM Policies and Training Harnesses for Autonomous Agentic Reinforcement Learning 11 upvotes, #22 of 2026-06-11
  2. ESPO: Early-Stopping Proximal Policy Optimization 19 upvotes, #17 of 2026-06-02
  3. QwenLong-L1.5: Post-Training Recipe for Long-Context Reasoning and Memory Management 98 upvotes, #2 of 2025-12-16
  4. Tongyi DeepResearch Technical Report 89 upvotes, #2 of 2025-10-29
  5. QwenLong-CPRS: Towards infty-LLMs with Dynamic Context Optimization 40 upvotes, #8 of 2025-05-26
  6. QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning 83 upvotes, #2 of 2025-05-26

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.