Yuling

Yuling on Hugging Face Daily Papers: 28 papers, 7 in the top 3 of their day, 999 upvotes.

  1. Schrödinger's Code Repository: Have LLMs Learned SWE-bench or Memorized It? 31 upvotes, #7 of 2026-09-24
  2. EarlyEval: Cheaper Agent Evaluation via Early Outcome Prediction 118 upvotes, #5 of 2026-09-03
  3. ParaTempo: Efficient Parallel Reasoning via Temporal Confidence 38 upvotes, #3 of 2026-08-24
  4. Repo0: Design-Driven Zero-to-All Code Generation 20 upvotes, #8 of 2026-08-21
  5. SkillForge: Self-Distilling Agents for Project-Specific Issue Resolution 11 upvotes, #17 of 2026-08-19
  6. SWE-Bench ProMax: Benchmarking Agents on Large-Scale Multilingual Code Refactoring 129 upvotes, #4 of 2026-08-11
  7. SWE-Pruner Pro: The Coder LLM Already Knows What to Prune 77 upvotes, #5 of 2026-07-21
  8. Know Before Fix: QA-Driven Repository Knowledge Acquisition for Software Issue Resolution 22 upvotes, #6 of 2026-07-15
  9. Are Performance-Optimization Benchmarks Reliably Measuring Coding Agents? 8 upvotes, #22 of 2026-07-02
  10. Dockerless: Environment-Free Program Verifier for Coding Agents 108 upvotes, #1 of 2026-07-01
  11. FastContext: Training Efficient Repository Explorer for Coding Agents 91 upvotes, #5 of 2026-06-16
  12. SWE-Explore: Benchmarking How Coding Agents Explore Repositories 112 upvotes, #2 of 2026-06-09
  13. Rethinking the Value of Agent-Generated Tests for LLM-Based Software Engineering Agents 4 upvotes, #34 of 2026-02-12
  14. DLLM-Searcher: Adapting Diffusion Large Language Model for Search Agents 30 upvotes, #9 of 2026-02-11
  15. CodeOCR: On the Effectiveness of Vision Language Models in Code Understanding 91 upvotes, #1 of 2026-02-04
  16. SWE-Pruner: Self-Adaptive Context Pruning for Coding Agents 87 upvotes, #2 of 2026-01-26
  17. GlimpRouter: Efficient Collaborative Inference by Glimpsing One Token of Thoughts 27 upvotes, #7 of 2026-01-13
  18. Fed-SE: Federated Self-Evolution for Privacy-Constrained Multi-Environment LLM Agents 3 upvotes, #19 of 2025-12-12
  19. GraphTracer: Graph-Guided Failure Tracing in LLM Agents for Robust Multi-Turn Deep Search 2 upvotes, #31 of 2025-10-16
  20. HyperAgent: Leveraging Hypergraphs for Topology Optimization in Multi-Agent Communication 4 upvotes, #26 of 2025-10-16
  21. LongCodeZip: Compress Long Context for Code Language Models 102 upvotes, #1 of 2025-10-03
  22. Attention as a Compass: Efficient Exploration for Process-Supervised RL in Reasoning Models 12 upvotes, #24 of 2025-10-01
  23. SWE-QA: Can Language Models Answer Repository-level Code Questions? 34 upvotes, #5 of 2025-09-24
  24. Pruning the Unsurprising: Efficient Code Reasoning via First-Token Surprisal 18 upvotes, #5 of 2025-08-11
  25. EVOC2RUST: A Skeleton-guided Framework for Project-Level C-to-Rust Translation 6 upvotes, #23 of 2025-08-07
  26. SWE-Debate: Competitive Multi-Agent Debate for Software Issue Resolution 10 upvotes, #7 of 2025-08-04
  27. SWE-Exp: Experience-Driven Software Issue Resolution 12 upvotes, #6 of 2025-08-04
  28. From Code to Correctness: Closing the Last Mile of Code Generation with Hierarchical Debugging 30 upvotes, #3 of 2024-10-03

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.