Xiao Liang
Xiao Liang on Hugging Face Daily Papers: 8 papers, 1 in the top 3 of their day, 265 upvotes.
- Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability 9 upvotes, #37 of 2026-02-03
- Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models 46 upvotes, #4 of 2026-01-21
- Gold-Medal-Level Olympiad Geometry Solving with Efficient Heuristic Auxiliary Constructions 1 upvotes, #41 of 2025-12-03
- Beyond the Exploration-Exploitation Trade-off: A Hidden State Approach for LLM Reasoning in RLVR 47 upvotes, #5 of 2025-09-30
- Beyond Pass@1: Self-Play with Variational Problem Synthesis Sustains RLVR 116 upvotes, #2 of 2025-08-25
- Reinforcement Learning with Verifiable Rewards Implicitly Incentivizes Correct Reasoning in Base LLMs 35 upvotes, #6 of 2025-06-18
- SwS: Self-aware Weakness-driven Problem Synthesis in Reinforcement Learning for LLM Reasoning 14 upvotes, #9 of 2025-06-16
- TL;DR: Too Long, Do Re-weighting for Effcient LLM Reasoning Compression 3 upvotes, #35 of 2025-06-04
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.