Kun Wan

Kun Wan on Hugging Face Daily Papers: 6 papers, 3 in the top 3 of their day, 409 upvotes.

  1. From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement 104 upvotes, #1 of 2026-08-03
  2. Vision-Zero: Scalable VLM Self-Improvement via Strategic Gamified Self-Play 123 upvotes, #3 of 2025-10-01
  3. EPO: Entropy-regularized Policy Optimization for LLM Agents Reinforcement Learning 124 upvotes, #2 of 2025-09-29
  4. Give Me FP32 or Give Me Death? Challenges and Solutions for Reproducible Reasoning 16 upvotes, #11 of 2025-06-12
  5. DUMP: Automated Distribution-Level Curriculum Learning for RL-based LLM Post-training 19 upvotes, #10 of 2025-04-15
  6. DL3DV-10K: A Large-Scale Scene Dataset for Deep Learning-based 3D Vision 18 upvotes, #7 of 2023-12-29

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.