TimLeung

TimLeung on Hugging Face Daily Papers: 10 papers, 3 in the top 3 of their day, 282 upvotes.

  1. Bottom-up Policy Optimization: Your Language Model Policy Secretly Contains Internal Policies 60 upvotes, #3 of 2025-12-24
  2. DeepTheorem: Advancing LLM Reasoning for Theorem Proving Through Natural Language and Reinforcement Learning 15 upvotes, #24 of 2025-05-30
  3. Two Experts Are All You Need for Steering Thinking: Reinforcing Cognitive Effort in MoE Reasoning Models Without Additional Training 9 upvotes, #23 of 2025-05-21
  4. DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning 11 upvotes, #17 of 2025-04-16
  5. Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs 51 upvotes, #2 of 2025-01-31
  6. Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs 27 upvotes, #5 of 2024-12-31
  7. Critical Tokens Matter: Token-Level Contrastive Estimation Enhence LLM's Reasoning Capability 47 upvotes, #2 of 2024-12-04
  8. Draft Model Knows When to Stop: A Self-Verification Length Policy for Speculative Decoding 6 upvotes, #17 of 2024-11-28
  9. Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training 5 upvotes, #14 of 2024-07-15
  10. Leveraging Word Guessing Games to Assess the Intelligence of Large Language Models 8 upvotes, #11 of 2023-11-01

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.