Kexin Huang

Kexin Huang on Hugging Face Daily Papers: 5 papers, 2 in the top 3 of their day, 267 upvotes.

  1. FIPO: Eliciting Deep Reasoning with Future-KL Influenced Policy Optimization 325 upvotes, #2 of 2026-04-01
  2. Sparse but Critical: A Token-Level Analysis of Distributional Shifts in RLVR Fine-Tuning of LLMs 7 upvotes, #18 of 2026-03-25
  3. On the Direction of RLVR Updates for LLM Reasoning: Identification and Exploitation 27 upvotes, #12 of 2026-03-24
  4. Quantile Advantage Estimation for Entropy-Safe Reasoning 113 upvotes, #3 of 2025-09-29
  5. RePO: ReLU-based Preference Optimization 1 upvotes, #46 of 2025-03-11

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.