Rishabh Agarwal

Rishabh Agarwal on Hugging Face Daily Papers: 5 papers, 3 in the top 3 of their day, 114 upvotes.

  1. Not All LLM Reasoners Are Created Equal 25 upvotes, #4 of 2024-10-03
  2. Stop Regressing: Training Value Functions via Classification for Scalable Deep RL 10 upvotes, #6 of 2024-03-07
  3. Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models 29 upvotes, #2 of 2023-12-12
  4. GKD: Generalized Knowledge Distillation for Auto-regressive Sequence Models 38 upvotes, #2 of 2023-06-26
  5. Bigger, Better, Faster: Human-level Atari with human-level efficiency 5 upvotes, #2 of 2023-06-01

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.