Bo Liu

Bo Liu on Hugging Face Daily Papers: 21 papers, 11 in the top 3 of their day, 1,642 upvotes.

  1. ScholarCatalyst: A Benchmark for Retrieving Papers That Inspire New Research 11 upvotes, #55 of 2026-10-02
  2. SPADE: Self-Play in Adaptive Synthetic Executable Environments 51 upvotes, #5 of 2026-08-20
  3. From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement 104 upvotes, #1 of 2026-08-03
  4. Agents' Last Exam 345 upvotes, #1 of 2026-06-09
  5. From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space 28 upvotes, #6 of 2026-04-16
  6. Reasoning over mathematical objects: on-policy reward modeling and test time aggregation 6 upvotes, #25 of 2026-03-20
  7. Scaling Agent Learning via Experience Synthesis 72 upvotes, #3 of 2025-11-07
  8. SPICE: Self-Play In Corpus Environments Improves Reasoning 12 upvotes, #25 of 2025-10-29
  9. BigCodeArena: Unveiling More Reliable Human Preferences in Code Generation via Execution 28 upvotes, #10 of 2025-10-13
  10. Agent Learning via Early Experience 223 upvotes, #1 of 2025-10-10
  11. GEM: A Gym for Agentic LLMs 79 upvotes, #2 of 2025-10-02
  12. Vision-Zero: Scalable VLM Self-Improvement via Strategic Gamified Self-Play 123 upvotes, #3 of 2025-10-01
  13. The Era of Real-World Human Interaction: RL from User Conversations 16 upvotes, #25 of 2025-09-30
  14. LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model 76 upvotes, #4 of 2025-09-03
  15. SPIRAL: Self-Play on Zero-Sum Games Incentivizes Reasoning via Multi-Agent Multi-Turn Reinforcement Learning 43 upvotes, #2 of 2025-07-01
  16. TextArena 27 upvotes, #7 of 2025-04-16
  17. Natural Language Reinforcement Learning 25 upvotes, #5 of 2024-11-22
  18. DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search 48 upvotes, #1 of 2024-08-16
  19. DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data 27 upvotes, #2 of 2024-05-24
  20. DeepSeek-VL: Towards Real-World Vision-Language Understanding 33 upvotes, #3 of 2024-03-11
  21. DeepSeek LLM: Scaling Open-Source Language Models with Longtermism 56 upvotes, #1 of 2024-01-08

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.