Qian Liu
Qian Liu on Hugging Face Daily Papers: 26 papers, 16 in the top 3 of their day, 1,263 upvotes.
- Diffusion Language Models are Super Data Learners 110 upvotes, #1 of 2025-11-06
- SimpleTIR: End-to-End Reinforcement Learning for Multi-Turn Tool-Integrated Reasoning 80 upvotes, #3 of 2025-09-03
- SWE-Perf: Can Language Models Optimize Code Performance on Real-World Repositories? 37 upvotes, #2 of 2025-07-17
- First Return, Entropy-Eliciting Explore 23 upvotes, #8 of 2025-07-10
- ZeCO: Zero Communication Overhead Sequence Parallelism for Linear Attention 10 upvotes, #12 of 2025-07-04
- General-Reasoner: Advancing LLM Reasoning Across All Domains 20 upvotes, #10 of 2025-05-21
- SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild 27 upvotes, #4 of 2025-03-25
- SkyLadder: Better and Faster Pretraining via Context Window Scheduling 11 upvotes, #13 of 2025-03-20
- Predictive Data Selection: The Data That Predicts Is the Data That Teaches 53 upvotes, #1 of 2025-03-03
- SuperGPQA: Scaling LLM Evaluation across 285 Graduate Disciplines 92 upvotes, #3 of 2025-02-21
- Sailor2: Sailing in South-East Asia with Inclusive Multilingual LLMs 11 upvotes, #14 of 2025-02-18
- When Precision Meets Position: BFloat16 Breaks Down RoPE in Long-Context Training 13 upvotes, #5 of 2024-11-21
- OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models 99 upvotes, #1 of 2024-11-08
- Cheating Automatic LLM Benchmarks: Null Models Achieve High Win Rates 6 upvotes, #20 of 2024-10-11
- Programming Every Example: Lifting Pre-training Data Quality like Experts at Scale 57 upvotes, #2 of 2024-09-26
- Scaling Laws with Vocabulary: Larger Models Deserve Larger Vocabularies 46 upvotes, #1 of 2024-07-19
- Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows? 5 upvotes, #10 of 2024-07-16
- RegMix: Data Mixture as Regression for Language Model Pre-training 24 upvotes, #7 of 2024-07-02
- BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions 41 upvotes, #2 of 2024-06-24
- Bootstrapping Language Models with DPO Implicit Rewards 34 upvotes, #3 of 2024-06-19
- StarCoder 2 and The Stack v2: The Next Generation 160 upvotes, #1 of 2024-03-01
- Astraios: Parameter-Efficient Instruction Tuning Code Large Language Models 23 upvotes, #3 of 2024-01-02
- Lemur: Harmonizing Natural Language and Code for Language Agents 31 upvotes, #3 of 2023-10-13
- OctoPack: Instruction Tuning Code Large Language Models 33 upvotes, #2 of 2023-08-15
- LoraHub: Efficient Cross-Task Generalization via Dynamic LoRA Composition 34 upvotes, #1 of 2023-07-26
- StarCoder: may the source be with you! 34 upvotes, #1 of 2023-05-11
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.