Shenzhi Wang
Shenzhi Wang on Hugging Face Daily Papers: 10 papers, 3 in the top 3 of their day, 629 upvotes.
- HopChain: Multi-Hop Data Synthesis for Generalizable Vision-Language Reasoning 106 upvotes, #1 of 2026-03-23
- Outcome Accuracy is Not Enough: Aligning the Reasoning Process of Reward Models 11 upvotes, #19 of 2026-02-09
- The Flexibility Trap: Why Arbitrary Order Limits Reasoning Potential in Diffusion Language Models 68 upvotes, #4 of 2026-01-23
- OS Agents: A Survey on MLLM-based Agents for General Computing Devices Use 9 upvotes, #9 of 2025-08-11
- Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning 144 upvotes, #1 of 2025-06-03
- Absolute Zero: Reinforced Self-play Reasoning with Zero Data 135 upvotes, #1 of 2025-05-07
- COIG-P: A High-Quality and Large-Scale Chinese Preference Dataset for Alignment with Human Values 41 upvotes, #5 of 2025-04-09
- DeeR-VLA: Dynamic Inference of Multimodal Large Language Models for Efficient Robot Execution 12 upvotes, #4 of 2024-11-06
- LLM-based Optimization of Compound AI Systems: A Survey 13 upvotes, #6 of 2024-10-23
- Model Surgery: Modulating LLM's Behavior Via Simple Parameter Editing 16 upvotes, #6 of 2024-07-15
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.