youyou
youyou on Hugging Face Daily Papers: 6 papers, 3 in the top 3 of their day, 380 upvotes.
- EvoTrainer: Co-Evolving LLM Policies and Training Harnesses for Autonomous Agentic Reinforcement Learning 11 upvotes, #22 of 2026-06-11
- ESPO: Early-Stopping Proximal Policy Optimization 19 upvotes, #17 of 2026-06-02
- QwenLong-L1.5: Post-Training Recipe for Long-Context Reasoning and Memory Management 98 upvotes, #2 of 2025-12-16
- Tongyi DeepResearch Technical Report 89 upvotes, #2 of 2025-10-29
- QwenLong-CPRS: Towards infty-LLMs with Dynamic Context Optimization 40 upvotes, #8 of 2025-05-26
- QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning 83 upvotes, #2 of 2025-05-26
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.