Dingwei Chen

Dingwei Chen on Hugging Face Daily Papers: 4 papers, 1 in the top 3 of their day, 122 upvotes.

  1. A^2TGPO: Agentic Turn-Group Policy Optimization with Adaptive Turn-level Clipping 14 upvotes, #15 of 2026-05-08
  2. From Context to Skills: Can Language Models Learn from Context Skillfully? 152 upvotes, #2 of 2026-05-05
  3. AT^2PO: Agentic Turn-based Policy Optimization via Tree Search 26 upvotes, #7 of 2026-01-09
  4. ChARM: Character-based Act-adaptive Reward Modeling for Advanced Role-Playing Language Agents 7 upvotes, #26 of 2025-06-02

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.