Andrew Zhao
Andrew Zhao on Hugging Face Daily Papers: 5 papers, 3 in the top 3 of their day, 557 upvotes.
- Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning 144 upvotes, #1 of 2025-06-03
- Absolute Zero: Reinforced Self-play Reasoning with Zero Data 135 upvotes, #1 of 2025-05-07
- Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model? 113 upvotes, #1 of 2025-04-21
- LLM-based Optimization of Compound AI Systems: A Survey 13 upvotes, #6 of 2024-10-23
- Model Surgery: Modulating LLM's Behavior Via Simple Parameter Editing 16 upvotes, #6 of 2024-07-15
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.