Yelong Shen
Yelong Shen on Hugging Face Daily Papers: 12 papers, 5 in the top 3 of their day, 511 upvotes.
- Phi-4-Mini-Reasoning: Exploring the Limits of Small Reasoning Language Models in Math 37 upvotes, #3 of 2025-05-01
- Reinforcement Learning for Reasoning in Large Language Models with One Training Example 88 upvotes, #1 of 2025-04-30
- Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs 70 upvotes, #1 of 2025-03-04
- Sigma: Differential Rescaling of Query, Key and Value for Efficient Language Models 40 upvotes, #3 of 2025-01-24
- GRIN: GRadient-INformed MoE 14 upvotes, #7 of 2024-09-19
- OmniParser for Pure Vision Based GUI Agent 16 upvotes, #6 of 2024-08-02
- Rho-1: Not All Tokens Are What You Need 72 upvotes, #1 of 2024-04-12
- Multi-LoRA Composition for Image Generation 30 upvotes, #5 of 2024-02-27
- Language Models can be Logical Solvers 19 upvotes, #7 of 2023-11-13
- An Empirical Study of Scaling Instruct-Tuned Large Multimodal Models 19 upvotes, #6 of 2023-09-19
- Efficient RLHF: Reducing the Memory Usage of PPO 16 upvotes, #5 of 2023-09-06
- AR-Diffusion: Auto-Regressive Diffusion Model for Text Generation 3 upvotes, #6 of 2023-05-17
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.