Yelong Shen

Yelong Shen on Hugging Face Daily Papers: 12 papers, 5 in the top 3 of their day, 511 upvotes.

  1. Phi-4-Mini-Reasoning: Exploring the Limits of Small Reasoning Language Models in Math 37 upvotes, #3 of 2025-05-01
  2. Reinforcement Learning for Reasoning in Large Language Models with One Training Example 88 upvotes, #1 of 2025-04-30
  3. Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs 70 upvotes, #1 of 2025-03-04
  4. Sigma: Differential Rescaling of Query, Key and Value for Efficient Language Models 40 upvotes, #3 of 2025-01-24
  5. GRIN: GRadient-INformed MoE 14 upvotes, #7 of 2024-09-19
  6. OmniParser for Pure Vision Based GUI Agent 16 upvotes, #6 of 2024-08-02
  7. Rho-1: Not All Tokens Are What You Need 72 upvotes, #1 of 2024-04-12
  8. Multi-LoRA Composition for Image Generation 30 upvotes, #5 of 2024-02-27
  9. Language Models can be Logical Solvers 19 upvotes, #7 of 2023-11-13
  10. An Empirical Study of Scaling Instruct-Tuned Large Multimodal Models 19 upvotes, #6 of 2023-09-19
  11. Efficient RLHF: Reducing the Memory Usage of PPO 16 upvotes, #5 of 2023-09-06
  12. AR-Diffusion: Auto-Regressive Diffusion Model for Text Generation 3 upvotes, #6 of 2023-05-17

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.