Zhipeng Chen

Zhipeng Chen on Hugging Face Daily Papers: 4 papers, 0 in the top 3 of their day, 104 upvotes.

  1. Pass@k Training for Adaptively Balancing Exploration and Exploitation of Large Reasoning Models 24 upvotes, #9 of 2025-08-15
  2. Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models 35 upvotes, #4 of 2025-03-28
  3. R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning 25 upvotes, #9 of 2025-03-10
  4. An Empirical Study on Eliciting and Improving R1-like Reasoning Models 8 upvotes, #18 of 2025-03-10

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.