Zhiyuan Zeng
Zhiyuan Zeng on Hugging Face Daily Papers: 3 papers, 1 in the top 3 of their day, 122 upvotes.
- RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments 12 upvotes, #14 of 2025-11-11
- Reinforcement Learning for Reasoning in Large Language Models with One Training Example 88 upvotes, #1 of 2025-04-30
- EvalTree: Profiling Language Model Weaknesses via Hierarchical Capability Trees 5 upvotes, #22 of 2025-03-19
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.