Yuchen Yan

Yuchen Yan on Hugging Face Daily Papers: 19 papers, 1 in the top 3 of their day, 511 upvotes.

  1. IterSynth: Rethinking Deep Search Agents via Role-Decoupled Iterative Synthesis 20 upvotes, #12 of 2026-09-25
  2. PaperGym: Rubric-Centered Evolution for Research-Plan Generation 39 upvotes, #7 of 2026-09-01
  3. Ring-Zero: Scaling Zero RL to a Trillion Parameters for Emergent Reasoning 98 upvotes, #3 of 2026-07-16
  4. KnowU-Bench: Towards Interactive, Proactive, and Personalized Mobile Agent Evaluation 47 upvotes, #12 of 2026-04-10
  5. Code-A1: Adversarial Evolving of Code LLM and Test LLM via Reinforcement Learning 10 upvotes, #20 of 2026-03-17
  6. InftyThink+: Effective and Efficient Infinite-Horizon Reasoning via Reinforcement Learning 12 upvotes, #16 of 2026-02-09
  7. Every Step Evolves: Scaling Reinforcement Learning for Trillion-Scale Thinking Model 61 upvotes, #6 of 2025-10-22
  8. EasySteer: A Unified Framework for High-Performance and Extensible LLM Steering 26 upvotes, #16 of 2025-09-30
  9. GSM8K-V: Can Vision Language Models Solve Grade School Math Word Problems in Visual Contexts 27 upvotes, #13 of 2025-09-30
  10. Cooper: Co-Optimizing Policy and Reward Models in Reinforcement Learning for Large Language Models 16 upvotes, #9 of 2025-08-14
  11. Test-Time Reinforcement Learning for GUI Grounding via Region Consistency 20 upvotes, #12 of 2025-08-13
  12. OmniEAR: Benchmarking Agent Reasoning in Embodied Tasks 18 upvotes, #14 of 2025-08-12
  13. LAPO: Internalizing Reasoning Efficiency via Length-Adaptive Policy Optimization 34 upvotes, #5 of 2025-07-25
  14. Hierarchical Budget Policy Optimization for Adaptive Reasoning 16 upvotes, #8 of 2025-07-25
  15. SVGenius: Benchmarking LLMs in SVG Understanding, Editing and Generation 14 upvotes, #15 of 2025-06-05
  16. ViewSpatial-Bench: Evaluating Multi-perspective Spatial Localization in Vision-Language Models 11 upvotes, #33 of 2025-05-28
  17. Mind the Gap: Bridging Thought Leap for Improved Chain-of-Thought Tuning 23 upvotes, #12 of 2025-05-23
  18. Let LLMs Break Free from Overthinking via Self-Braking Tuning 23 upvotes, #12 of 2025-05-23
  19. VerifyBench: Benchmarking Reference-based Reward Systems for Large Language Models 17 upvotes, #13 of 2025-05-22

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.