Qian Liu

Qian Liu on Hugging Face Daily Papers: 26 papers, 16 in the top 3 of their day, 1,263 upvotes.

  1. Diffusion Language Models are Super Data Learners 110 upvotes, #1 of 2025-11-06
  2. SimpleTIR: End-to-End Reinforcement Learning for Multi-Turn Tool-Integrated Reasoning 80 upvotes, #3 of 2025-09-03
  3. SWE-Perf: Can Language Models Optimize Code Performance on Real-World Repositories? 37 upvotes, #2 of 2025-07-17
  4. First Return, Entropy-Eliciting Explore 23 upvotes, #8 of 2025-07-10
  5. ZeCO: Zero Communication Overhead Sequence Parallelism for Linear Attention 10 upvotes, #12 of 2025-07-04
  6. General-Reasoner: Advancing LLM Reasoning Across All Domains 20 upvotes, #10 of 2025-05-21
  7. SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild 27 upvotes, #4 of 2025-03-25
  8. SkyLadder: Better and Faster Pretraining via Context Window Scheduling 11 upvotes, #13 of 2025-03-20
  9. Predictive Data Selection: The Data That Predicts Is the Data That Teaches 53 upvotes, #1 of 2025-03-03
  10. SuperGPQA: Scaling LLM Evaluation across 285 Graduate Disciplines 92 upvotes, #3 of 2025-02-21
  11. Sailor2: Sailing in South-East Asia with Inclusive Multilingual LLMs 11 upvotes, #14 of 2025-02-18
  12. When Precision Meets Position: BFloat16 Breaks Down RoPE in Long-Context Training 13 upvotes, #5 of 2024-11-21
  13. OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models 99 upvotes, #1 of 2024-11-08
  14. Cheating Automatic LLM Benchmarks: Null Models Achieve High Win Rates 6 upvotes, #20 of 2024-10-11
  15. Programming Every Example: Lifting Pre-training Data Quality like Experts at Scale 57 upvotes, #2 of 2024-09-26
  16. Scaling Laws with Vocabulary: Larger Models Deserve Larger Vocabularies 46 upvotes, #1 of 2024-07-19
  17. Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows? 5 upvotes, #10 of 2024-07-16
  18. RegMix: Data Mixture as Regression for Language Model Pre-training 24 upvotes, #7 of 2024-07-02
  19. BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions 41 upvotes, #2 of 2024-06-24
  20. Bootstrapping Language Models with DPO Implicit Rewards 34 upvotes, #3 of 2024-06-19
  21. StarCoder 2 and The Stack v2: The Next Generation 160 upvotes, #1 of 2024-03-01
  22. Astraios: Parameter-Efficient Instruction Tuning Code Large Language Models 23 upvotes, #3 of 2024-01-02
  23. Lemur: Harmonizing Natural Language and Code for Language Agents 31 upvotes, #3 of 2023-10-13
  24. OctoPack: Instruction Tuning Code Large Language Models 33 upvotes, #2 of 2023-08-15
  25. LoraHub: Efficient Cross-Task Generalization via Dynamic LoRA Composition 34 upvotes, #1 of 2023-07-26
  26. StarCoder: may the source be with you! 34 upvotes, #1 of 2023-05-11

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.