Rishabh Agarwal
Rishabh Agarwal on Hugging Face Daily Papers: 5 papers, 3 in the top 3 of their day, 114 upvotes.
- Not All LLM Reasoners Are Created Equal 25 upvotes, #4 of 2024-10-03
- Stop Regressing: Training Value Functions via Classification for Scalable Deep RL 10 upvotes, #6 of 2024-03-07
- Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models 29 upvotes, #2 of 2023-12-12
- GKD: Generalized Knowledge Distillation for Auto-regressive Sequence Models 38 upvotes, #2 of 2023-06-26
- Bigger, Better, Faster: Human-level Atari with human-level efficiency 5 upvotes, #2 of 2023-06-01
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.