Nathan Lambert
Nathan Lambert on Hugging Face Daily Papers: 13 papers, 5 in the top 3 of their day, 590 upvotes.
- Context Language Models 40 upvotes, #29 of 2026-09-30
- Learning to Solve Hard Problems in RL for LLMs by Never Giving Up 13 upvotes, #20 of 2026-09-15
- Tmax: A simple recipe for terminal agents 14 upvotes, #19 of 2026-06-23
- Meta-Reinforcement Learning with Self-Reflection for Agentic Search 8 upvotes, #21 of 2026-03-13
- Olmo 3 22 upvotes, #11 of 2025-12-17
- TÜLU 3: Pushing Frontiers in Open Language Model Post-Training 55 upvotes, #1 of 2024-11-25
- M-RewardBench: Evaluating Reward Models in Multilingual Settings 10 upvotes, #7 of 2024-10-24
- Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Multimodal Models 89 upvotes, #1 of 2024-09-26
- OLMoE: Open Mixture-of-Experts Language Models 67 upvotes, #2 of 2024-09-04
- WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs 11 upvotes, #6 of 2024-06-27
- RewardBench: Evaluating Reward Models for Language Modeling 14 upvotes, #9 of 2024-03-21
- OLMo: Accelerating the Science of Language Models 86 upvotes, #1 of 2024-02-02
- Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research 66 upvotes, #2 of 2024-02-02
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.