Yuzhen Huang
Yuzhen Huang on Hugging Face Daily Papers: 10 papers, 3 in the top 3 of their day, 513 upvotes.
- DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression 180 upvotes, #1 of 2026-09-18
- LOCA-bench: Benchmarking Language Agents Under Controllable and Extreme Context Growth 24 upvotes, #18 of 2026-02-10
- SWE-RM: Execution-free Feedback For Software Engineering Agents 9 upvotes, #11 of 2025-12-29
- The Tool Decathlon: Benchmarking Language Agents for Diverse, Realistic, and Long-Horizon Task Execution 44 upvotes, #6 of 2025-10-30
- Pitfalls of Rule- and Model-based Verifiers -- A Case Study on Mathematical Reasoning 6 upvotes, #32 of 2025-05-29
- Learn to Reason Efficiently with Adaptive Length-based Reward Shaping 31 upvotes, #8 of 2025-05-22
- SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild 27 upvotes, #4 of 2025-03-25
- Predictive Data Selection: The Data That Predicts Is the Data That Teaches 53 upvotes, #1 of 2025-03-03
- B-STaR: Monitoring and Balancing Exploration and Exploitation in Self-Taught Reasoners 38 upvotes, #2 of 2024-12-24
- Compression Represents Intelligence Linearly 26 upvotes, #4 of 2024-04-16
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.