dodojorid

dodojorid on Hugging Face Daily Papers: 6 papers, 3 in the top 3 of their day, 496 upvotes.

  1. Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening 79 upvotes, #2 of 2026-09-17
  2. SimpleOPD: Simple Tokenizer-Agnostic On-Policy Distillation for Long-Context Reasoning 34 upvotes, #5 of 2026-08-17
  3. ComBench: A Benchmark for Rigorous Proof Reasoning and Constructive Realization in Olympiad-Level Combinatorics 19 upvotes, #14 of 2026-06-11
  4. Achieving Gold-Medal-Level Olympiad Reasoning via Simple and Unified Scaling 154 upvotes, #1 of 2026-05-15
  5. P1-VL: Bridging Visual Perception and Scientific Reasoning in Physics Olympiads 57 upvotes, #6 of 2026-02-11
  6. P1: Mastering Physics Olympiads with Reinforcement Learning 128 upvotes, #3 of 2025-11-18

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.