Kexin Huang
Kexin Huang on Hugging Face Daily Papers: 5 papers, 2 in the top 3 of their day, 267 upvotes.
- FIPO: Eliciting Deep Reasoning with Future-KL Influenced Policy Optimization 325 upvotes, #2 of 2026-04-01
- Sparse but Critical: A Token-Level Analysis of Distributional Shifts in RLVR Fine-Tuning of LLMs 7 upvotes, #18 of 2026-03-25
- On the Direction of RLVR Updates for LLM Reasoning: Identification and Exploitation 27 upvotes, #12 of 2026-03-24
- Quantile Advantage Estimation for Entropy-Safe Reasoning 113 upvotes, #3 of 2025-09-29
- RePO: ReLU-based Preference Optimization 1 upvotes, #46 of 2025-03-11
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.