Yuzhe Gu
Yuzhe Gu on Hugging Face Daily Papers: 13 papers, 6 in the top 3 of their day, 843 upvotes.
- ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning 25 upvotes, #11 of 2026-06-04
- Intern-S1-Pro: Scientific Multimodal Foundation Model at Trillion Scale 125 upvotes, #1 of 2026-03-27
- OPV: Outcome-based Process Verifier for Efficient Long Chain-of-Thought Verification 32 upvotes, #4 of 2025-12-12
- Long-horizon Reasoning Agent for Olympiad-Level Mathematical Problem Solving 44 upvotes, #2 of 2025-12-12
- MM-HELIX: Boosting Multimodal Long-Chain Reflective Reasoning with Holistic Platform and Adaptive Hybrid Policy Optimization 103 upvotes, #2 of 2025-10-10
- Intern-S1: A Scientific Multimodal Foundation Model 242 upvotes, #1 of 2025-08-22
- CompassVerifier: A Unified and Robust Verifier for LLMs Evaluation and Outcome Reward 32 upvotes, #4 of 2025-08-06
- Semi-off-Policy Reinforcement Learning for Vision-Language Slow-thinking Reasoning 22 upvotes, #7 of 2025-07-23
- The Imitation Game: Turing Machine Imitator is Length Generalizable Reasoner 46 upvotes, #5 of 2025-07-18
- Mask-DPO: Generalizable Fine-grained Factuality Alignment of LLMs 18 upvotes, #5 of 2025-03-05
- Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning 55 upvotes, #3 of 2025-02-11
- ANAH-v2: Scaling Analytical Hallucination Annotation of Large Language Models 1 upvotes, #17 of 2024-07-09
- InternLM2 Technical Report 22 upvotes, #3 of 2024-03-27
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.