Yuzhe Gu

Yuzhe Gu on Hugging Face Daily Papers: 13 papers, 6 in the top 3 of their day, 843 upvotes.

  1. ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning 25 upvotes, #11 of 2026-06-04
  2. Intern-S1-Pro: Scientific Multimodal Foundation Model at Trillion Scale 125 upvotes, #1 of 2026-03-27
  3. OPV: Outcome-based Process Verifier for Efficient Long Chain-of-Thought Verification 32 upvotes, #4 of 2025-12-12
  4. Long-horizon Reasoning Agent for Olympiad-Level Mathematical Problem Solving 44 upvotes, #2 of 2025-12-12
  5. MM-HELIX: Boosting Multimodal Long-Chain Reflective Reasoning with Holistic Platform and Adaptive Hybrid Policy Optimization 103 upvotes, #2 of 2025-10-10
  6. Intern-S1: A Scientific Multimodal Foundation Model 242 upvotes, #1 of 2025-08-22
  7. CompassVerifier: A Unified and Robust Verifier for LLMs Evaluation and Outcome Reward 32 upvotes, #4 of 2025-08-06
  8. Semi-off-Policy Reinforcement Learning for Vision-Language Slow-thinking Reasoning 22 upvotes, #7 of 2025-07-23
  9. The Imitation Game: Turing Machine Imitator is Length Generalizable Reasoner 46 upvotes, #5 of 2025-07-18
  10. Mask-DPO: Generalizable Fine-grained Factuality Alignment of LLMs 18 upvotes, #5 of 2025-03-05
  11. Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning 55 upvotes, #3 of 2025-02-11
  12. ANAH-v2: Scaling Analytical Hallucination Annotation of Large Language Models 1 upvotes, #17 of 2024-07-09
  13. InternLM2 Technical Report 22 upvotes, #3 of 2024-03-27

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.