Fan Zhou

Fan Zhou on Hugging Face Daily Papers: 7 papers, 4 in the top 3 of their day, 275 upvotes.

  1. OctoThinker: Mid-training Incentivizes Reinforcement Learning Scaling 42 upvotes, #4 of 2025-06-26
  2. Revisiting Reinforcement Learning for LLM Reasoning from A Cross-Domain Perspective 42 upvotes, #1 of 2025-06-20
  3. MegaMath: Pushing the Limits of Open Math Corpora 29 upvotes, #2 of 2025-04-07
  4. Sailor2: Sailing in South-East Asia with Inclusive Multilingual LLMs 11 upvotes, #14 of 2025-02-18
  5. Diving into Self-Evolving Training for Multimodal Reasoning 37 upvotes, #3 of 2024-12-24
  6. Programming Every Example: Lifting Pre-training Data Quality like Experts at Scale 57 upvotes, #2 of 2024-09-26
  7. OlympicArena: Benchmarking Multi-discipline Cognitive Reasoning for Superintelligent AI 14 upvotes, #10 of 2024-06-19

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.