Youbang Sun
Youbang Sun on Hugging Face Daily Papers: 13 papers, 6 in the top 3 of their day, 995 upvotes.
- Safin-1: Safety from Within through Memory-Native State Evolution 20 upvotes, #12 of 2026-09-02
- Intern-S2-Mobius: Foundation Model with Decoupled Knowledge and Reasoning 33 upvotes, #6 of 2026-08-17
- Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering 181 upvotes, #4 of 2026-07-31
- Post-Trained MoE Can Skip Half Experts via Self-Distillation 30 upvotes, #10 of 2026-05-19
- How Far Can Unsupervised RLVR Scale LLM Training? 52 upvotes, #4 of 2026-03-10
- FlowRL: Matching Reward Distributions for LLM Reasoning 100 upvotes, #2 of 2025-09-19
- SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning 73 upvotes, #3 of 2025-09-12
- A Survey of Reinforcement Learning for Large Reasoning Models 156 upvotes, #1 of 2025-09-11
- R^textbf{2AI}: Towards Resistant and Resilient AI in an Evolving World 3 upvotes, #20 of 2025-09-09
- Towards a Unified View of Large Language Model Post-Training 67 upvotes, #3 of 2025-09-05
- TTRL: Test-Time Reinforcement Learning 96 upvotes, #2 of 2025-04-23
- Technologies on Effectiveness and Efficiency: A Survey of State Spaces Models 26 upvotes, #6 of 2025-03-17
- Fourier Position Embedding: Enhancing Attention's Periodic Extension for Length Generalization 35 upvotes, #1 of 2024-12-25
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.