Pengfei Liu
Pengfei Liu on Hugging Face Daily Papers: 25 papers, 4 in the top 3 of their day, 678 upvotes.
- daVinci-Env: Open SWE Environment Synthesis at Scale 29 upvotes, #7 of 2026-03-16
- One Sample to Rule Them All: Extreme Data Efficiency in RL Scaling 8 upvotes, #15 of 2026-01-09
- OctoThinker: Mid-training Incentivizes Reinforcement Learning Scaling 42 upvotes, #4 of 2025-06-26
- Thinking with Generated Images 13 upvotes, #26 of 2025-05-29
- Towards Dynamic Theory of Mind: Evaluating LLM Adaptation to Temporal Evolution of Human States 14 upvotes, #21 of 2025-05-29
- LIMOPro: Reasoning Refinement for Efficient and Effective Test-time Scaling 12 upvotes, #28 of 2025-05-29
- One RL to See Them All: Visual Triple Unified Reinforcement Learning 59 upvotes, #6 of 2025-05-26
- Efficient Agent Training for Computer Use 41 upvotes, #6 of 2025-05-22
- Generative AI Act II: Test Time Scaling Drives Cognition Engineering 16 upvotes, #8 of 2025-04-21
- Rethinking RL Scaling for Vision Language Models: A Transparent, From-Scratch Framework and Comprehensive Evaluation Scheme 30 upvotes, #10 of 2025-04-04
- LIMO: Less is More for Reasoning 47 upvotes, #3 of 2025-02-06
- O1 Replication Journey -- Part 3: Inference-time Scaling for Medical Reasoning 29 upvotes, #7 of 2025-01-14
- PC Agent: While You Sleep, AI Works -- A Cognitive Journey into Digital World 10 upvotes, #12 of 2024-12-24
- O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? 35 upvotes, #2 of 2024-11-26
- Programming Every Example: Lifting Pre-training Data Quality like Experts at Scale 57 upvotes, #2 of 2024-09-26
- OpenResearcher: Unleashing AI for Accelerated Scientific Research 28 upvotes, #5 of 2024-08-14
- Data Contamination Report from the 2024 CONDA Shared Task 8 upvotes, #6 of 2024-08-01
- Understanding Reference Policies in Direct Preference Optimization 13 upvotes, #6 of 2024-07-19
- ANOLE: An Open, Autoregressive, Native Large Multimodal Models for Interleaved Image-Text Generation 19 upvotes, #5 of 2024-07-09
- OlympicArena Medal Ranks: Who Is the Most Intelligent AI So Far? 3 upvotes, #24 of 2024-06-25
- OlympicArena: Benchmarking Multi-discipline Cognitive Reasoning for Superintelligent AI 14 upvotes, #10 of 2024-06-19
- Reformatted Alignment 17 upvotes, #9 of 2024-02-20
- Extending LLMs' Context Window with 100 Samples 16 upvotes, #6 of 2024-01-17
- Generative AI for Math: Part I -- MathPile: A Billion-Token-Scale Pretraining Corpus for Math 28 upvotes, #3 of 2023-12-29
- Alignment for Honesty 13 upvotes, #5 of 2023-12-13
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.