Wenhan Ma

Wenhan Ma on Hugging Face Daily Papers: 4 papers, 2 in the top 3 of their day, 198 upvotes.

  1. MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training 15 upvotes, #20 of 2026-07-01
  2. Stabilizing MoE Reinforcement Learning by Aligning Training and Inference Routers 2 upvotes, #24 of 2025-10-27
  3. MiMo-VL Technical Report 70 upvotes, #1 of 2025-06-05
  4. MiMo: Unlocking the Reasoning Potential of Language Model -- From Pretraining to Posttraining 74 upvotes, #2 of 2025-05-13

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.