Le Yu
Le Yu on Hugging Face Daily Papers: 7 papers, 4 in the top 3 of their day, 1,046 upvotes.
- RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback 13 upvotes, #9 of 2025-07-23
- Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning 144 upvotes, #1 of 2025-06-03
- Qwen3 Technical Report 152 upvotes, #1 of 2025-05-19
- WorldPM: Scaling Human Preference Modeling 33 upvotes, #5 of 2025-05-16
- Qwen2.5-1M Technical Report 51 upvotes, #1 of 2025-01-28
- Qwen2.5 Technical Report 328 upvotes, #1 of 2024-12-20
- A Unified View of Delta Parameter Editing in Post-Trained Large-Scale Models 14 upvotes, #16 of 2024-10-18
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.