Peter Belcak

Peter Belcak on Hugging Face Daily Papers: 4 papers, 3 in the top 3 of their day, 474 upvotes.

  1. GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization 191 upvotes, #1 of 2026-01-09
  2. Small Language Models are the Future of Agentic AI 3 upvotes, #36 of 2025-06-05
  3. CLIMB: CLustering-based Iterative Data Mixture Bootstrapping for Language Model Pre-training 87 upvotes, #1 of 2025-04-18
  4. Exponentially Faster Language Modelling 119 upvotes, #1 of 2023-11-21

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.