Zengzhi Wang

Zengzhi Wang on Hugging Face Daily Papers: 8 papers, 4 in the top 3 of their day, 271 upvotes.

  1. MegaScience: Pushing the Frontiers of Post-Training Datasets for Science Reasoning 48 upvotes, #3 of 2025-07-23
  2. OctoThinker: Mid-training Incentivizes Reinforcement Learning Scaling 42 upvotes, #4 of 2025-06-26
  3. MegaMath: Pushing the Limits of Open Math Corpora 29 upvotes, #2 of 2025-04-07
  4. Programming Every Example: Lifting Pre-training Data Quality like Experts at Scale 57 upvotes, #2 of 2024-09-26
  5. Data Contamination Report from the 2024 CONDA Shared Task 8 upvotes, #6 of 2024-08-01
  6. OlympicArena Medal Ranks: Who Is the Most Intelligent AI So Far? 3 upvotes, #24 of 2024-06-25
  7. OlympicArena: Benchmarking Multi-discipline Cognitive Reasoning for Superintelligent AI 14 upvotes, #10 of 2024-06-19
  8. Generative AI for Math: Part I -- MathPile: A Billion-Token-Scale Pretraining Corpus for Math 28 upvotes, #3 of 2023-12-29

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.