Zhewen Tan

Zhewen Tan on Hugging Face Daily Papers: 4 papers, 0 in the top 3 of their day, 228 upvotes.

  1. S3Gym: Can LLMs Turn Self-Testing and Self-Judging into Self-Improvement? 39 upvotes, #8 of 2026-09-03
  2. ESPO: Early-Stopping Proximal Policy Optimization 19 upvotes, #17 of 2026-06-02
  3. TriPlay-RL: Tri-Role Self-Play Reinforcement Learning for LLM Safety Alignment 10 upvotes, #12 of 2026-01-28
  4. NL2Repo-Bench: Towards Long-Horizon Repository Generation Evaluation of Coding Agents 42 upvotes, #7 of 2025-12-16

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.