Zhixiang Zhou
Zhixiang Zhou on Hugging Face Daily Papers: 2 papers, 0 in the top 3 of their day, 50 upvotes.
- CPGD: Toward Stable Rule-based Reinforcement Learning for Language Models 23 upvotes, #13 of 2025-05-20
- MM-PRM: Enhancing Multimodal Mathematical Reasoning with Scalable Step-Level Supervision 25 upvotes, #12 of 2025-05-20
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.