Daily Papers of 2026-08-14
- Alaya-EVOKE: From Linear-Scaling Supervision to Endless World 132 upvotes, #1 of 2026-08-14
- DarwinX: Evolving Agent Harnesses Through Natural Selection 110 upvotes, #2 of 2026-08-14
- LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers 108 upvotes, #3 of 2026-08-14
- DreamX-Phi 1.0: Action-Conditioned Video World Model for Robotic Manipulation 98 upvotes, #4 of 2026-08-14
- OmniScientist: An Omni-Modal Omni-Discipline AI Scientist 87 upvotes, #5 of 2026-08-14
- Intern-S2-Preview: Scientific Agentic Foundation Model 68 upvotes, #6 of 2026-08-14
- AutoDesign: Meta-Harness Optimization for Long-Horizon Agentic Design 53 upvotes, #7 of 2026-08-14
- How Can Rhetoric Reward-Hack AI Reviewers? Dissecting Rhetorical Sensitivity in AI-Based Peer Review 48 upvotes, #8 of 2026-08-14
- PlayWorld: Benchmarking World Models with Agent Players over Long-Horizon Objectives 45 upvotes, #9 of 2026-08-14
- Spatial Memory Agent: Experience-Grounded Procedure Memory for Spatial Intelligence 43 upvotes, #10 of 2026-08-14
- Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus 30 upvotes, #11 of 2026-08-14
- LiveAnimate: Stable Long-Form Streaming Human Animation in Real-Time 23 upvotes, #12 of 2026-08-14
- Full-bandwidth transformer 22 upvotes, #13 of 2026-08-14
- UniSwap: Streaming Audio-Visual Identity Swapping for Talking Videos 22 upvotes, #13 of 2026-08-14
- Context-Matched Distillation: Teacher Causality for Autoregressive Video Distillation 18 upvotes, #15 of 2026-08-14
- H2R-Bench: Benchmarking Human-to-Robot Manipulation Video Generation in World Models 17 upvotes, #16 of 2026-08-14
- An AI4AI Framework for Visual Token Pruning 15 upvotes, #17 of 2026-08-14
- Thought-Level Beam Search for Reasoning 15 upvotes, #17 of 2026-08-14
- SKILLER: Language-Level Reinforcement Learning for Reusable Skill Extraction in Small Language Models 15 upvotes, #17 of 2026-08-14
- Knowing When to Quit: Diagnosing and Training LLMs to Abort Futile Reasoning 14 upvotes, #20 of 2026-08-14
- Maglev: Sliding Recurrent Memory 14 upvotes, #20 of 2026-08-14
- LycheeMemory V2: Efficient Long-Term Memory for LLM Agents via Semantic Segment-Level Consolidation 14 upvotes, #20 of 2026-08-14
- Are You Sure You're Sure? On the Impact of Instruction Tuning on Confidence and Lexical Diversity 12 upvotes, #23 of 2026-08-14
- Hybrid-Policy Self-Editing for Composable Unstructured Knowledge Editing 11 upvotes, #24 of 2026-08-14
- From Inaudible Inputs to Model Failures: Low-Frequency Safety Risks in LALMs 10 upvotes, #25 of 2026-08-14
- Specification-first convergence with an AI coding agent: a case study of dismantling a core architectural invariant across 189 files in a 717k-line codebase with no test oracle and no human code review 10 upvotes, #25 of 2026-08-14
- PixSDS: Why Latent SDS Makes Noisy Pixels 10 upvotes, #25 of 2026-08-14
- RibAssist 3D: Biplanar Rib-Fracture Detection, Addressing, and Selective 3D Localization from CT-Derived Projections 9 upvotes, #28 of 2026-08-14
- Mitigating Gender Bias in English to Romanian Machine Translation 9 upvotes, #28 of 2026-08-14
- TailBooster: A Dual-Layer Generative Framework for Extreme Value Augmentation with Operational Validity Enforcement 8 upvotes, #30 of 2026-08-14
- CW-BASS v2: Saturation-Aware Pseudo-Label Selection for Semi-Supervised Segmentation under Foundation-Model Teachers 8 upvotes, #30 of 2026-08-14
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.