Zhipeng Chen
Zhipeng Chen on Hugging Face Daily Papers: 4 papers, 0 in the top 3 of their day, 104 upvotes.
- Pass@k Training for Adaptively Balancing Exploration and Exploitation of Large Reasoning Models 24 upvotes, #9 of 2025-08-15
- Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models 35 upvotes, #4 of 2025-03-28
- R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning 25 upvotes, #9 of 2025-03-10
- An Empirical Study on Eliciting and Improving R1-like Reasoning Models 8 upvotes, #18 of 2025-03-10
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.