Bowen Yu
Bowen Yu on Hugging Face Daily Papers: 6 papers, 3 in the top 3 of their day, 310 upvotes.
- Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning 144 upvotes, #1 of 2025-06-03
- Self-play with Execution Feedback: Improving Instruction-following Capabilities of Large Language Models 14 upvotes, #13 of 2024-06-21
- Large Language Models are Superpositions of All Characters: Attaining Arbitrary Role-play via Self-Alignment 36 upvotes, #2 of 2024-01-24
- Qwen Technical Report 39 upvotes, #4 of 2023-09-29
- PolyLM: An Open Source Polyglot Large Language Model 27 upvotes, #2 of 2023-07-13
- Domain Incremental Lifelong Learning in an Open World 1 upvotes, #9 of 2023-05-12
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.