Yuzhen Huang

Yuzhen Huang on Hugging Face Daily Papers: 10 papers, 3 in the top 3 of their day, 513 upvotes.

  1. DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression 180 upvotes, #1 of 2026-09-18
  2. LOCA-bench: Benchmarking Language Agents Under Controllable and Extreme Context Growth 24 upvotes, #18 of 2026-02-10
  3. SWE-RM: Execution-free Feedback For Software Engineering Agents 9 upvotes, #11 of 2025-12-29
  4. The Tool Decathlon: Benchmarking Language Agents for Diverse, Realistic, and Long-Horizon Task Execution 44 upvotes, #6 of 2025-10-30
  5. Pitfalls of Rule- and Model-based Verifiers -- A Case Study on Mathematical Reasoning 6 upvotes, #32 of 2025-05-29
  6. Learn to Reason Efficiently with Adaptive Length-based Reward Shaping 31 upvotes, #8 of 2025-05-22
  7. SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild 27 upvotes, #4 of 2025-03-25
  8. Predictive Data Selection: The Data That Predicts Is the Data That Teaches 53 upvotes, #1 of 2025-03-03
  9. B-STaR: Monitoring and Balancing Exploration and Exploitation in Self-Taught Reasoners 38 upvotes, #2 of 2024-12-24
  10. Compression Represents Intelligence Linearly 26 upvotes, #4 of 2024-04-16

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.