Fuli Luo

Fuli Luo on Hugging Face Daily Papers: 8 papers, 5 in the top 3 of their day, 790 upvotes.

  1. MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training 15 upvotes, #20 of 2026-07-01
  2. MiMo-V2-Flash Technical Report 31 upvotes, #8 of 2026-01-07
  3. Stabilizing MoE Reinforcement Learning by Aligning Training and Inference Routers 2 upvotes, #24 of 2025-10-27
  4. DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning 271 upvotes, #1 of 2025-01-23
  5. DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search 48 upvotes, #1 of 2024-08-16
  6. DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence 74 upvotes, #1 of 2024-01-26
  7. DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models 65 upvotes, #2 of 2024-01-12
  8. DeepSeek LLM: Scaling Open-Source Language Models with Longtermism 56 upvotes, #1 of 2024-01-08

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.