Anikait Singh
Anikait Singh on Hugging Face Daily Papers: 7 papers, 3 in the top 3 of their day, 202 upvotes.
- RLAD: Training LLMs to Discover Abstractions for Solving Reasoning Problems 7 upvotes, #28 of 2025-10-03
- Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs 31 upvotes, #4 of 2025-03-04
- FSPO: Few-Shot Preference Optimization of Synthetic Preference Data in LLMs Elicits Effective Personalization to Real Users 5 upvotes, #17 of 2025-02-27
- Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Though 81 upvotes, #2 of 2025-01-09
- D5RL: Diverse Datasets for Data-Driven Deep Reinforcement Learning 6 upvotes, #7 of 2024-08-19
- Robotic Offline RL from Internet Videos via Value-Function Pre-Training 9 upvotes, #3 of 2023-09-25
- RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control 33 upvotes, #3 of 2023-08-01
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.