Shaohang Wei

Shaohang Wei on Hugging Face Daily Papers: 5 papers, 0 in the top 3 of their day, 40 upvotes.

  1. Verifier-Induced Support Reshaping in On-Policy Optimization 5 upvotes, #23 of 2026-08-17
  2. Sparse but Critical: A Token-Level Analysis of Distributional Shifts in RLVR Fine-Tuning of LLMs 7 upvotes, #18 of 2026-03-25
  3. Mitigating Overthinking through Reasoning Shaping 4 upvotes, #31 of 2025-10-13
  4. Well Begun is Half Done: Low-resource Preference Alignment by Weak-to-Strong Decoding 7 upvotes, #23 of 2025-06-10
  5. TIME: A Multi-level Benchmark for Temporal Reasoning of LLMs in Real-World Scenarios 2 upvotes, #41 of 2025-05-26

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.