Qiaosheng ZHANG
Qiaosheng ZHANG on Hugging Face Daily Papers: 3 papers, 1 in the top 3 of their day, 111 upvotes.
- CPGD: Toward Stable Rule-based Reinforcement Learning for Language Models 23 upvotes, #13 of 2025-05-20
- MM-PRM: Enhancing Multimodal Mathematical Reasoning with Scalable Step-Level Supervision 25 upvotes, #12 of 2025-05-20
- MM-Eureka: Exploring Visual Aha Moment with Rule-based Large-scale Reinforcement Learning 53 upvotes, #3 of 2025-03-11
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.