Hongyao Tang
Hongyao Tang on Hugging Face Daily Papers: 4 papers, 2 in the top 3 of their day, 326 upvotes.
- Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement 82 upvotes, #7 of 2026-09-16
- WarpSAC: Towards the Pinnacle of Scalable Off-policy RL by Rethinking Exploration and Exploitation 141 upvotes, #5 of 2026-08-27
- The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning 161 upvotes, #1 of 2026-07-06
- Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models 143 upvotes, #1 of 2026-06-11
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.