Yuling
Yuling on Hugging Face Daily Papers: 28 papers, 7 in the top 3 of their day, 999 upvotes.
- Schrödinger's Code Repository: Have LLMs Learned SWE-bench or Memorized It? 31 upvotes, #7 of 2026-09-24
- EarlyEval: Cheaper Agent Evaluation via Early Outcome Prediction 118 upvotes, #5 of 2026-09-03
- ParaTempo: Efficient Parallel Reasoning via Temporal Confidence 38 upvotes, #3 of 2026-08-24
- Repo0: Design-Driven Zero-to-All Code Generation 20 upvotes, #8 of 2026-08-21
- SkillForge: Self-Distilling Agents for Project-Specific Issue Resolution 11 upvotes, #17 of 2026-08-19
- SWE-Bench ProMax: Benchmarking Agents on Large-Scale Multilingual Code Refactoring 129 upvotes, #4 of 2026-08-11
- SWE-Pruner Pro: The Coder LLM Already Knows What to Prune 77 upvotes, #5 of 2026-07-21
- Know Before Fix: QA-Driven Repository Knowledge Acquisition for Software Issue Resolution 22 upvotes, #6 of 2026-07-15
- Are Performance-Optimization Benchmarks Reliably Measuring Coding Agents? 8 upvotes, #22 of 2026-07-02
- Dockerless: Environment-Free Program Verifier for Coding Agents 108 upvotes, #1 of 2026-07-01
- FastContext: Training Efficient Repository Explorer for Coding Agents 91 upvotes, #5 of 2026-06-16
- SWE-Explore: Benchmarking How Coding Agents Explore Repositories 112 upvotes, #2 of 2026-06-09
- Rethinking the Value of Agent-Generated Tests for LLM-Based Software Engineering Agents 4 upvotes, #34 of 2026-02-12
- DLLM-Searcher: Adapting Diffusion Large Language Model for Search Agents 30 upvotes, #9 of 2026-02-11
- CodeOCR: On the Effectiveness of Vision Language Models in Code Understanding 91 upvotes, #1 of 2026-02-04
- SWE-Pruner: Self-Adaptive Context Pruning for Coding Agents 87 upvotes, #2 of 2026-01-26
- GlimpRouter: Efficient Collaborative Inference by Glimpsing One Token of Thoughts 27 upvotes, #7 of 2026-01-13
- Fed-SE: Federated Self-Evolution for Privacy-Constrained Multi-Environment LLM Agents 3 upvotes, #19 of 2025-12-12
- GraphTracer: Graph-Guided Failure Tracing in LLM Agents for Robust Multi-Turn Deep Search 2 upvotes, #31 of 2025-10-16
- HyperAgent: Leveraging Hypergraphs for Topology Optimization in Multi-Agent Communication 4 upvotes, #26 of 2025-10-16
- LongCodeZip: Compress Long Context for Code Language Models 102 upvotes, #1 of 2025-10-03
- Attention as a Compass: Efficient Exploration for Process-Supervised RL in Reasoning Models 12 upvotes, #24 of 2025-10-01
- SWE-QA: Can Language Models Answer Repository-level Code Questions? 34 upvotes, #5 of 2025-09-24
- Pruning the Unsurprising: Efficient Code Reasoning via First-Token Surprisal 18 upvotes, #5 of 2025-08-11
- EVOC2RUST: A Skeleton-guided Framework for Project-Level C-to-Rust Translation 6 upvotes, #23 of 2025-08-07
- SWE-Debate: Competitive Multi-Agent Debate for Software Issue Resolution 10 upvotes, #7 of 2025-08-04
- SWE-Exp: Experience-Driven Software Issue Resolution 12 upvotes, #6 of 2025-08-04
- From Code to Correctness: Closing the Last Mile of Code Generation with Hierarchical Debugging 30 upvotes, #3 of 2024-10-03
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.