Shenzhi Yang

Shenzhi Yang on Hugging Face Daily Papers: 2 papers, 0 in the top 3 of their day, 41 upvotes.

  1. Can LLMs Learn to Reason Robustly under Noisy Supervision? 42 upvotes, #8 of 2026-04-07
  2. TraPO: A Semi-Supervised Reinforcement Learning Framework for Boosting LLM Reasoning 3 upvotes, #29 of 2025-12-17

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.