Hamish Ivison
Hamish Ivison on Hugging Face Daily Papers: 12 papers, 2 in the top 3 of their day, 341 upvotes.
- Learning to Solve Hard Problems in RL for LLMs by Never Giving Up 13 upvotes, #20 of 2026-09-15
- Tmax: A simple recipe for terminal agents 14 upvotes, #19 of 2026-06-23
- Meta-Reinforcement Learning with Self-Reflection for Agentic Search 8 upvotes, #21 of 2026-03-13
- Olmo 3 22 upvotes, #11 of 2025-12-17
- DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research 53 upvotes, #4 of 2025-11-25
- RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments 12 upvotes, #14 of 2025-11-11
- Large-Scale Data Selection for Instruction Tuning 10 upvotes, #13 of 2025-03-04
- TESS 2: A Large-Scale Generalist Diffusion Language Model 5 upvotes, #21 of 2025-02-20
- TÜLU 3: Pushing Frontiers in Open Language Model Post-Training 55 upvotes, #1 of 2024-11-25
- OLMo: Accelerating the Science of Language Models 86 upvotes, #1 of 2024-02-02
- How Far Can Camels Go? Exploring the State of Instruction Tuning on Open Resources 5 upvotes, #8 of 2023-06-09
- TESS: Text-to-Text Self-Conditioned Simplex Diffusion 3 upvotes, #5 of 2023-05-16
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.