Ruotian Ma

Ruotian Ma on Hugging Face Daily Papers: 5 papers, 1 in the top 3 of their day, 116 upvotes.

  1. RLVER: Reinforcement Learning with Verifiable Emotion Rewards for Empathetic Agents 31 upvotes, #7 of 2025-07-09
  2. Two Experts Are All You Need for Steering Thinking: Reinforcing Cognitive Effort in MoE Reasoning Models Without Additional Training 9 upvotes, #23 of 2025-05-21
  3. Sentient Agent as a Judge: Evaluating Higher-Order Social Cognition in Large Language Models 25 upvotes, #3 of 2025-05-09
  4. SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning 15 upvotes, #5 of 2025-04-29
  5. S^2R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning 27 upvotes, #8 of 2025-02-21

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.