xuxin
xuxin on Hugging Face Daily Papers: 8 papers, 1 in the top 3 of their day, 186 upvotes.
- Learning to Foresee: Unveiling the Unlocking Efficiency of On-Policy Distillation 58 upvotes, #5 of 2026-05-18
- Progressive Residual Warmup for Language Model Pretraining 33 upvotes, #5 of 2026-03-09
- Composition-RL: Compose Your Verifiable Prompts for Reinforcement Learning of Large Language Models 92 upvotes, #2 of 2026-02-13
- EntroPIC: Towards Stable Long-Term Training of LLMs via Entropy Stabilization with Proportional-Integral Control 5 upvotes, #18 of 2025-11-21
- On Predictability of Reinforcement Learning Dynamics for Large Language Models 8 upvotes, #16 of 2025-10-02
- Thinking-Free Policy Initialization Makes Distilled Reasoning Models More Effective and Efficient Reasoners 29 upvotes, #10 of 2025-10-01
- GPAS: Accelerating Convergence of LLM Pretraining via Gradient-Preserving Activation Scaling 2 upvotes, #22 of 2025-06-30
- VerifyBench: Benchmarking Reference-based Reward Systems for Large Language Models 17 upvotes, #13 of 2025-05-22
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.