zhuqihao

zhuqihao on Hugging Face Daily Papers: 7 papers, 7 in the top 3 of their day, 925 upvotes.

  1. DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning 271 upvotes, #1 of 2025-01-23
  2. DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search 48 upvotes, #1 of 2024-08-16
  3. DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence 53 upvotes, #1 of 2024-06-19
  4. DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data 27 upvotes, #2 of 2024-05-24
  5. DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models 150 upvotes, #1 of 2024-02-06
  6. DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence 74 upvotes, #1 of 2024-01-26
  7. DeepSeek LLM: Scaling Open-Source Language Models with Longtermism 56 upvotes, #1 of 2024-01-08

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.