Weizhu Chen

Weizhu Chen on Hugging Face Daily Papers: 11 papers, 7 in the top 3 of their day, 758 upvotes.

  1. Phi-4-Mini-Reasoning: Exploring the Limits of Small Reasoning Language Models in Math 37 upvotes, #3 of 2025-05-01
  2. Reinforcement Learning for Reasoning in Large Language Models with One Training Example 88 upvotes, #1 of 2025-04-30
  3. Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs 70 upvotes, #1 of 2025-03-04
  4. LongRoPE2: Near-Lossless LLM Context Window Scaling 31 upvotes, #5 of 2025-02-28
  5. GRIN: GRadient-INformed MoE 14 upvotes, #7 of 2024-09-19
  6. Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone 222 upvotes, #1 of 2024-04-23
  7. Rho-1: Not All Tokens Are What You Need 72 upvotes, #1 of 2024-04-12
  8. Multi-LoRA Composition for Image Generation 30 upvotes, #5 of 2024-02-27
  9. Language Models can be Logical Solvers 19 upvotes, #7 of 2023-11-13
  10. Learning From Mistakes Makes LLM Better Reasoner 29 upvotes, #1 of 2023-11-01
  11. LoftQ: LoRA-Fine-Tuning-Aware Quantization for Large Language Models 30 upvotes, #2 of 2023-10-16

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.