Weizhu Chen
Weizhu Chen on Hugging Face Daily Papers: 11 papers, 7 in the top 3 of their day, 758 upvotes.
- Phi-4-Mini-Reasoning: Exploring the Limits of Small Reasoning Language Models in Math 37 upvotes, #3 of 2025-05-01
- Reinforcement Learning for Reasoning in Large Language Models with One Training Example 88 upvotes, #1 of 2025-04-30
- Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs 70 upvotes, #1 of 2025-03-04
- LongRoPE2: Near-Lossless LLM Context Window Scaling 31 upvotes, #5 of 2025-02-28
- GRIN: GRadient-INformed MoE 14 upvotes, #7 of 2024-09-19
- Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone 222 upvotes, #1 of 2024-04-23
- Rho-1: Not All Tokens Are What You Need 72 upvotes, #1 of 2024-04-12
- Multi-LoRA Composition for Image Generation 30 upvotes, #5 of 2024-02-27
- Language Models can be Logical Solvers 19 upvotes, #7 of 2023-11-13
- Learning From Mistakes Makes LLM Better Reasoner 29 upvotes, #1 of 2023-11-01
- LoftQ: LoRA-Fine-Tuning-Aware Quantization for Large Language Models 30 upvotes, #2 of 2023-10-16
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.