Yuchen Yan
Yuchen Yan on Hugging Face Daily Papers: 19 papers, 1 in the top 3 of their day, 511 upvotes.
- IterSynth: Rethinking Deep Search Agents via Role-Decoupled Iterative Synthesis 20 upvotes, #12 of 2026-09-25
- PaperGym: Rubric-Centered Evolution for Research-Plan Generation 39 upvotes, #7 of 2026-09-01
- Ring-Zero: Scaling Zero RL to a Trillion Parameters for Emergent Reasoning 98 upvotes, #3 of 2026-07-16
- KnowU-Bench: Towards Interactive, Proactive, and Personalized Mobile Agent Evaluation 47 upvotes, #12 of 2026-04-10
- Code-A1: Adversarial Evolving of Code LLM and Test LLM via Reinforcement Learning 10 upvotes, #20 of 2026-03-17
- InftyThink+: Effective and Efficient Infinite-Horizon Reasoning via Reinforcement Learning 12 upvotes, #16 of 2026-02-09
- Every Step Evolves: Scaling Reinforcement Learning for Trillion-Scale Thinking Model 61 upvotes, #6 of 2025-10-22
- EasySteer: A Unified Framework for High-Performance and Extensible LLM Steering 26 upvotes, #16 of 2025-09-30
- GSM8K-V: Can Vision Language Models Solve Grade School Math Word Problems in Visual Contexts 27 upvotes, #13 of 2025-09-30
- Cooper: Co-Optimizing Policy and Reward Models in Reinforcement Learning for Large Language Models 16 upvotes, #9 of 2025-08-14
- Test-Time Reinforcement Learning for GUI Grounding via Region Consistency 20 upvotes, #12 of 2025-08-13
- OmniEAR: Benchmarking Agent Reasoning in Embodied Tasks 18 upvotes, #14 of 2025-08-12
- LAPO: Internalizing Reasoning Efficiency via Length-Adaptive Policy Optimization 34 upvotes, #5 of 2025-07-25
- Hierarchical Budget Policy Optimization for Adaptive Reasoning 16 upvotes, #8 of 2025-07-25
- SVGenius: Benchmarking LLMs in SVG Understanding, Editing and Generation 14 upvotes, #15 of 2025-06-05
- ViewSpatial-Bench: Evaluating Multi-perspective Spatial Localization in Vision-Language Models 11 upvotes, #33 of 2025-05-28
- Mind the Gap: Bridging Thought Leap for Improved Chain-of-Thought Tuning 23 upvotes, #12 of 2025-05-23
- Let LLMs Break Free from Overthinking via Self-Braking Tuning 23 upvotes, #12 of 2025-05-23
- VerifyBench: Benchmarking Reference-based Reward Systems for Large Language Models 17 upvotes, #13 of 2025-05-22
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.