Zhanfeng Mo

Zhanfeng Mo on Hugging Face Daily Papers: 4 papers, 1 in the top 3 of their day, 228 upvotes.

  1. First Try Matters: Revisiting the Role of Reflection in Reasoning Models 24 upvotes, #16 of 2025-10-10
  2. Multi-Agent Tool-Integrated Policy Optimization 29 upvotes, #8 of 2025-10-09
  3. MiroMind-M1: An Open-Source Advancement in Mathematical Reasoning via Context-Aware Multi-Stage Policy Optimization 116 upvotes, #2 of 2025-07-22
  4. 100 Days After DeepSeek-R1: A Survey on Replication Studies and More Directions for Reasoning Language Models 29 upvotes, #5 of 2025-05-01

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.