Pengfei Liu

Pengfei Liu on Hugging Face Daily Papers: 25 papers, 4 in the top 3 of their day, 678 upvotes.

  1. daVinci-Env: Open SWE Environment Synthesis at Scale 29 upvotes, #7 of 2026-03-16
  2. One Sample to Rule Them All: Extreme Data Efficiency in RL Scaling 8 upvotes, #15 of 2026-01-09
  3. OctoThinker: Mid-training Incentivizes Reinforcement Learning Scaling 42 upvotes, #4 of 2025-06-26
  4. Thinking with Generated Images 13 upvotes, #26 of 2025-05-29
  5. Towards Dynamic Theory of Mind: Evaluating LLM Adaptation to Temporal Evolution of Human States 14 upvotes, #21 of 2025-05-29
  6. LIMOPro: Reasoning Refinement for Efficient and Effective Test-time Scaling 12 upvotes, #28 of 2025-05-29
  7. One RL to See Them All: Visual Triple Unified Reinforcement Learning 59 upvotes, #6 of 2025-05-26
  8. Efficient Agent Training for Computer Use 41 upvotes, #6 of 2025-05-22
  9. Generative AI Act II: Test Time Scaling Drives Cognition Engineering 16 upvotes, #8 of 2025-04-21
  10. Rethinking RL Scaling for Vision Language Models: A Transparent, From-Scratch Framework and Comprehensive Evaluation Scheme 30 upvotes, #10 of 2025-04-04
  11. LIMO: Less is More for Reasoning 47 upvotes, #3 of 2025-02-06
  12. O1 Replication Journey -- Part 3: Inference-time Scaling for Medical Reasoning 29 upvotes, #7 of 2025-01-14
  13. PC Agent: While You Sleep, AI Works -- A Cognitive Journey into Digital World 10 upvotes, #12 of 2024-12-24
  14. O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? 35 upvotes, #2 of 2024-11-26
  15. Programming Every Example: Lifting Pre-training Data Quality like Experts at Scale 57 upvotes, #2 of 2024-09-26
  16. OpenResearcher: Unleashing AI for Accelerated Scientific Research 28 upvotes, #5 of 2024-08-14
  17. Data Contamination Report from the 2024 CONDA Shared Task 8 upvotes, #6 of 2024-08-01
  18. Understanding Reference Policies in Direct Preference Optimization 13 upvotes, #6 of 2024-07-19
  19. ANOLE: An Open, Autoregressive, Native Large Multimodal Models for Interleaved Image-Text Generation 19 upvotes, #5 of 2024-07-09
  20. OlympicArena Medal Ranks: Who Is the Most Intelligent AI So Far? 3 upvotes, #24 of 2024-06-25
  21. OlympicArena: Benchmarking Multi-discipline Cognitive Reasoning for Superintelligent AI 14 upvotes, #10 of 2024-06-19
  22. Reformatted Alignment 17 upvotes, #9 of 2024-02-20
  23. Extending LLMs' Context Window with 100 Samples 16 upvotes, #6 of 2024-01-17
  24. Generative AI for Math: Part I -- MathPile: A Billion-Token-Scale Pretraining Corpus for Math 28 upvotes, #3 of 2023-12-29
  25. Alignment for Honesty 13 upvotes, #5 of 2023-12-13

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.