Doohyuk Jang

Doohyuk Jang on Hugging Face Daily Papers: 4 papers, 0 in the top 3 of their day, 247 upvotes.

  1. Learning Beyond What You Sample: Off-Policy-Aware Cross-Model Trajectory Exchange for RLVR 62 upvotes, #19 of 2026-09-30
  2. Meta-Awareness Enhances Reasoning Models: Self-Alignment Reinforcement Learning 54 upvotes, #7 of 2025-10-10
  3. ReviewScore: Misinformed Peer Review Detection with Large Language Models 62 upvotes, #7 of 2025-09-29
  4. Reasoning Model is Stubborn: Diagnosing Instruction Overriding in Reasoning Models 63 upvotes, #5 of 2025-05-26

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.