Wenwei Zhang

Wenwei Zhang on Hugging Face Daily Papers: 31 papers, 15 in the top 3 of their day, 1,758 upvotes.

  1. ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning 25 upvotes, #11 of 2026-06-04
  2. Intern-S1-Pro: Scientific Multimodal Foundation Model at Trillion Scale 125 upvotes, #1 of 2026-03-27
  3. OPV: Outcome-based Process Verifier for Efficient Long Chain-of-Thought Verification 32 upvotes, #4 of 2025-12-12
  4. Long-horizon Reasoning Agent for Olympiad-Level Mathematical Problem Solving 44 upvotes, #2 of 2025-12-12
  5. Achieving Olympia-Level Geometry Large Language Model Agent via Complexity Boosting Reinforcement Learning 31 upvotes, #5 of 2025-12-12
  6. MM-HELIX: Boosting Multimodal Long-Chain Reflective Reasoning with Holistic Platform and Adaptive Hybrid Policy Optimization 103 upvotes, #2 of 2025-10-10
  7. InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency 170 upvotes, #1 of 2025-08-26
  8. Intern-S1: A Scientific Multimodal Foundation Model 242 upvotes, #1 of 2025-08-22
  9. CompassVerifier: A Unified and Robust Verifier for LLMs Evaluation and Outcome Reward 32 upvotes, #4 of 2025-08-06
  10. Semi-off-Policy Reinforcement Learning for Vision-Language Slow-thinking Reasoning 22 upvotes, #7 of 2025-07-23
  11. The Imitation Game: Turing Machine Imitator is Length Generalizable Reasoner 46 upvotes, #5 of 2025-07-18
  12. Rethinking Verification for LLM Code Generation: From Generation to Testing 28 upvotes, #6 of 2025-07-10
  13. Pre-Trained Policy Discriminators are General Reward Models 33 upvotes, #6 of 2025-07-08
  14. RIG: Synergizing Reasoning and Imagination in End-to-End Generalist Policy 29 upvotes, #8 of 2025-04-01
  15. Mask-DPO: Generalizable Fine-grained Factuality Alignment of LLMs 18 upvotes, #5 of 2025-03-05
  16. Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning 55 upvotes, #3 of 2025-02-11
  17. InternLM-XComposer2.5-Reward: A Simple Yet Effective Multi-Modal Reward Model 39 upvotes, #6 of 2025-01-22
  18. Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives 23 upvotes, #4 of 2025-01-10
  19. Are Your LLMs Capable of Stable Reasoning? 87 upvotes, #1 of 2024-12-18
  20. InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions 89 upvotes, #1 of 2024-12-13
  21. LLaVA-3D: A Simple yet Effective Pathway to Empowering LMMs with 3D-awareness 32 upvotes, #3 of 2024-09-27
  22. MindSearch: Mimicking Human Minds Elicits Deep AI Searcher 37 upvotes, #6 of 2024-07-30
  23. ANAH-v2: Scaling Analytical Hallucination Annotation of Large Language Models 1 upvotes, #17 of 2024-07-09
  24. InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output 84 upvotes, #1 of 2024-07-04
  25. InternLM-XComposer2-4KHD: A Pioneering Large Vision-Language Model Handling Resolutions from 336 Pixels to 4K HD 24 upvotes, #3 of 2024-04-10
  26. InternLM2 Technical Report 22 upvotes, #3 of 2024-03-27
  27. Agent-FLAN: Designing Data and Methods of Effective Agent Tuning for Large Language Models 11 upvotes, #6 of 2024-03-20
  28. InternLM-Math: Open Math Large Language Models Toward Verifiable Reasoning 19 upvotes, #2 of 2024-02-12
  29. InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model 55 upvotes, #1 of 2024-01-30
  30. GPT4RoI: Instruction Tuning Large Language Model on Region-of-Interest 13 upvotes, #3 of 2023-07-10
  31. MultiModal-GPT: A Vision and Language Model for Dialogue with Humans 1 upvotes, #6 of 2023-05-09

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.