Wenwei Zhang
Wenwei Zhang on Hugging Face Daily Papers: 31 papers, 15 in the top 3 of their day, 1,758 upvotes.
- ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning 25 upvotes, #11 of 2026-06-04
- Intern-S1-Pro: Scientific Multimodal Foundation Model at Trillion Scale 125 upvotes, #1 of 2026-03-27
- OPV: Outcome-based Process Verifier for Efficient Long Chain-of-Thought Verification 32 upvotes, #4 of 2025-12-12
- Long-horizon Reasoning Agent for Olympiad-Level Mathematical Problem Solving 44 upvotes, #2 of 2025-12-12
- Achieving Olympia-Level Geometry Large Language Model Agent via Complexity Boosting Reinforcement Learning 31 upvotes, #5 of 2025-12-12
- MM-HELIX: Boosting Multimodal Long-Chain Reflective Reasoning with Holistic Platform and Adaptive Hybrid Policy Optimization 103 upvotes, #2 of 2025-10-10
- InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency 170 upvotes, #1 of 2025-08-26
- Intern-S1: A Scientific Multimodal Foundation Model 242 upvotes, #1 of 2025-08-22
- CompassVerifier: A Unified and Robust Verifier for LLMs Evaluation and Outcome Reward 32 upvotes, #4 of 2025-08-06
- Semi-off-Policy Reinforcement Learning for Vision-Language Slow-thinking Reasoning 22 upvotes, #7 of 2025-07-23
- The Imitation Game: Turing Machine Imitator is Length Generalizable Reasoner 46 upvotes, #5 of 2025-07-18
- Rethinking Verification for LLM Code Generation: From Generation to Testing 28 upvotes, #6 of 2025-07-10
- Pre-Trained Policy Discriminators are General Reward Models 33 upvotes, #6 of 2025-07-08
- RIG: Synergizing Reasoning and Imagination in End-to-End Generalist Policy 29 upvotes, #8 of 2025-04-01
- Mask-DPO: Generalizable Fine-grained Factuality Alignment of LLMs 18 upvotes, #5 of 2025-03-05
- Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning 55 upvotes, #3 of 2025-02-11
- InternLM-XComposer2.5-Reward: A Simple Yet Effective Multi-Modal Reward Model 39 upvotes, #6 of 2025-01-22
- Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives 23 upvotes, #4 of 2025-01-10
- Are Your LLMs Capable of Stable Reasoning? 87 upvotes, #1 of 2024-12-18
- InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions 89 upvotes, #1 of 2024-12-13
- LLaVA-3D: A Simple yet Effective Pathway to Empowering LMMs with 3D-awareness 32 upvotes, #3 of 2024-09-27
- MindSearch: Mimicking Human Minds Elicits Deep AI Searcher 37 upvotes, #6 of 2024-07-30
- ANAH-v2: Scaling Analytical Hallucination Annotation of Large Language Models 1 upvotes, #17 of 2024-07-09
- InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output 84 upvotes, #1 of 2024-07-04
- InternLM-XComposer2-4KHD: A Pioneering Large Vision-Language Model Handling Resolutions from 336 Pixels to 4K HD 24 upvotes, #3 of 2024-04-10
- InternLM2 Technical Report 22 upvotes, #3 of 2024-03-27
- Agent-FLAN: Designing Data and Methods of Effective Agent Tuning for Large Language Models 11 upvotes, #6 of 2024-03-20
- InternLM-Math: Open Math Large Language Models Toward Verifiable Reasoning 19 upvotes, #2 of 2024-02-12
- InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model 55 upvotes, #1 of 2024-01-30
- GPT4RoI: Instruction Tuning Large Language Model on Region-of-Interest 13 upvotes, #3 of 2023-07-10
- MultiModal-GPT: A Vision and Language Model for Dialogue with Humans 1 upvotes, #6 of 2023-05-09
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.