Wenkai Yang
Wenkai Yang on Hugging Face Daily Papers: 7 papers, 2 in the top 3 of their day, 333 upvotes.
- Rethinking Continual Experience Internalization for Self-Evolving LLM Agents 23 upvotes, #9 of 2026-06-05
- Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe 85 upvotes, #2 of 2026-04-15
- AgentProcessBench: Diagnosing Step-Level Process Quality in Tool-Using Agents 22 upvotes, #16 of 2026-03-18
- Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation 57 upvotes, #4 of 2026-02-13
- LaSeR: Reinforcement Learning with Last-Token Self-Rewarding 37 upvotes, #9 of 2025-10-17
- DeepCritic: Deliberate Critique with Large Language Models 48 upvotes, #1 of 2025-05-02
- Super(ficial)-alignment: Strong Models May Deceive Weak Models in Weak-to-Strong Generalization 4 upvotes, #24 of 2024-06-19
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.