Jonas Hübotter
Jonas Hübotter on Hugging Face Daily Papers: 5 papers, 0 in the top 3 of their day, 111 upvotes.
- Tool-R0: Self-Evolving LLM Agents for Tool-Learning from Zero Data 10 upvotes, #17 of 2026-03-03
- Reinforcement Learning via Self-Distillation 36 upvotes, #5 of 2026-01-29
- Self-Distillation Enables Continual Learning 24 upvotes, #6 of 2026-01-28
- Learning on the Job: Test-Time Curricula for Targeted Reinforcement Learning 1 upvotes, #35 of 2025-10-07
- Specialization after Generalization: Towards Understanding Test-Time Training in Foundation Models 2 upvotes, #42 of 2025-10-01
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.