Fan Zhou
Fan Zhou on Hugging Face Daily Papers: 7 papers, 4 in the top 3 of their day, 275 upvotes.
- OctoThinker: Mid-training Incentivizes Reinforcement Learning Scaling 42 upvotes, #4 of 2025-06-26
- Revisiting Reinforcement Learning for LLM Reasoning from A Cross-Domain Perspective 42 upvotes, #1 of 2025-06-20
- MegaMath: Pushing the Limits of Open Math Corpora 29 upvotes, #2 of 2025-04-07
- Sailor2: Sailing in South-East Asia with Inclusive Multilingual LLMs 11 upvotes, #14 of 2025-02-18
- Diving into Self-Evolving Training for Multimodal Reasoning 37 upvotes, #3 of 2024-12-24
- Programming Every Example: Lifting Pre-training Data Quality like Experts at Scale 57 upvotes, #2 of 2024-09-26
- OlympicArena: Benchmarking Multi-discipline Cognitive Reasoning for Superintelligent AI 14 upvotes, #10 of 2024-06-19
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.