Hamish Ivison

Hamish Ivison on Hugging Face Daily Papers: 12 papers, 2 in the top 3 of their day, 341 upvotes.

  1. Learning to Solve Hard Problems in RL for LLMs by Never Giving Up 13 upvotes, #20 of 2026-09-15
  2. Tmax: A simple recipe for terminal agents 14 upvotes, #19 of 2026-06-23
  3. Meta-Reinforcement Learning with Self-Reflection for Agentic Search 8 upvotes, #21 of 2026-03-13
  4. Olmo 3 22 upvotes, #11 of 2025-12-17
  5. DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research 53 upvotes, #4 of 2025-11-25
  6. RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments 12 upvotes, #14 of 2025-11-11
  7. Large-Scale Data Selection for Instruction Tuning 10 upvotes, #13 of 2025-03-04
  8. TESS 2: A Large-Scale Generalist Diffusion Language Model 5 upvotes, #21 of 2025-02-20
  9. TÜLU 3: Pushing Frontiers in Open Language Model Post-Training 55 upvotes, #1 of 2024-11-25
  10. OLMo: Accelerating the Science of Language Models 86 upvotes, #1 of 2024-02-02
  11. How Far Can Camels Go? Exploring the State of Instruction Tuning on Open Resources 5 upvotes, #8 of 2023-06-09
  12. TESS: Text-to-Text Self-Conditioned Simplex Diffusion 3 upvotes, #5 of 2023-05-16

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.