dodojorid
dodojorid on Hugging Face Daily Papers: 6 papers, 3 in the top 3 of their day, 496 upvotes.
- Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening 79 upvotes, #2 of 2026-09-17
- SimpleOPD: Simple Tokenizer-Agnostic On-Policy Distillation for Long-Context Reasoning 34 upvotes, #5 of 2026-08-17
- ComBench: A Benchmark for Rigorous Proof Reasoning and Constructive Realization in Olympiad-Level Combinatorics 19 upvotes, #14 of 2026-06-11
- Achieving Gold-Medal-Level Olympiad Reasoning via Simple and Unified Scaling 154 upvotes, #1 of 2026-05-15
- P1-VL: Bridging Visual Perception and Scientific Reasoning in Physics Olympiads 57 upvotes, #6 of 2026-02-11
- P1: Mastering Physics Olympiads with Reinforcement Learning 128 upvotes, #3 of 2025-11-18
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.