Tiehua Mei

Tiehua Mei on Hugging Face Daily Papers: 3 papers, 1 in the top 3 of their day, 160 upvotes.

  1. ProRL: Effective Reinforcement Learning for Proactive Recommendation via Rectified Policy Gradient Estimation 87 upvotes, #2 of 2026-05-28
  2. GoLongRL: Capability-Oriented Long Context Reinforcement Learning with Multitask Alignment 56 upvotes, #6 of 2026-05-20
  3. Entropy Ratio Clipping as a Soft Global Constraint for Stable Reinforcement Learning 16 upvotes, #10 of 2025-12-08

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.