Pradeep Dasigi
Pradeep Dasigi on Hugging Face Daily Papers: 6 papers, 2 in the top 3 of their day, 248 upvotes.
- DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research 53 upvotes, #4 of 2025-11-25
- Large-Scale Data Selection for Instruction Tuning 10 upvotes, #13 of 2025-03-04
- TÜLU 3: Pushing Frontiers in Open Language Model Post-Training 55 upvotes, #1 of 2024-11-25
- Hybrid Preferences: Learning to Route Instances for Human vs. AI Feedback 10 upvotes, #9 of 2024-10-28
- OLMo: Accelerating the Science of Language Models 86 upvotes, #1 of 2024-02-02
- How Far Can Camels Go? Exploring the State of Instruction Tuning on Open Resources 5 upvotes, #8 of 2023-06-09
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.