Zhibin Gou
Zhibin Gou on Hugging Face Daily Papers: 5 papers, 5 in the top 3 of their day, 704 upvotes.
- DeepSeekMath-V2: Towards Self-Verifiable Mathematical Reasoning 66 upvotes, #2 of 2025-12-01
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning 271 upvotes, #1 of 2025-01-23
- DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search 48 upvotes, #1 of 2024-08-16
- DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence 53 upvotes, #1 of 2024-06-19
- CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing 9 upvotes, #2 of 2023-05-22
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.