Daily Papers of 2026-06-12
- MiniMax Sparse Attention 140 upvotes, #1 of 2026-06-12
- EvoArena: Tracking Memory Evolution for Robust LLM Agents in Dynamic Environments 137 upvotes, #2 of 2026-06-12
- WeaveBench: A Long-Horizon, Real-World Benchmark for Computer-Use Agents with Hybrid Interfaces 101 upvotes, #3 of 2026-06-12
- SpatialClaw: Rethinking Action Interface for Agentic Spatial Reasoning 101 upvotes, #3 of 2026-06-12
- MaxProof: Scaling Mathematical Proof with Generative-Verifier RL and Population-Level Test-Time Scaling 89 upvotes, #5 of 2026-06-12
- InterleaveThinker: Reinforcing Agentic Interleaved Generation 79 upvotes, #6 of 2026-06-12
- Robust-U1: Can MLLMs Self-Recover Corrupted Visual Content for Robust Understanding? 78 upvotes, #7 of 2026-06-12
- FORT-Searcher: Synthesizing Shortcut-Resistant Search Tasks for Training Deep Search Agents 74 upvotes, #8 of 2026-06-12
- LabVLA: Grounding Vision-Language-Action Models in Scientific Laboratories 53 upvotes, #9 of 2026-06-12
- VIA-SD: Verification via Intra-Model Routing for Speculative Decoding 35 upvotes, #10 of 2026-06-12
- From 2D Grids to 1D Tokens: Reforming Shared Representations for Multimodal Image Fusion 32 upvotes, #11 of 2026-06-12
- HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers 28 upvotes, #12 of 2026-06-12
- EurekAgent: Agent Environment Engineering is All You Need For Autonomous Scientific Discovery 27 upvotes, #13 of 2026-06-12
- N-GRPO: Embedding-Level Neighbor Mixing for Enhanced Policy Optimization 24 upvotes, #14 of 2026-06-12
- Demystifying Hidden-State Recurrence: Switchable Latent Reasoning with On-Policy Reinforcement Learning 21 upvotes, #15 of 2026-06-12
- VideoMDM: Towards 3D Human Motion Generation From 2D Supervision 20 upvotes, #16 of 2026-06-12
- Where, What, Why, and Importance: Structured Defect Grounding for Text-to-Image Feedback 15 upvotes, #17 of 2026-06-12
- MoVerse: Real-Time Video World Modeling with Panoramic Gaussian Scaffold 13 upvotes, #18 of 2026-06-12
- HarnessBridge: Learnable Bidirectional Controller for LLM Agent Harness 12 upvotes, #19 of 2026-06-12
- High-Fidelity Two-Step Image Generation via Teacher-Aligned End-to-End Distillation 11 upvotes, #20 of 2026-06-12
- TreeSeeker: Tree-Structured Trial, Error, and Return in Deep Search 10 upvotes, #21 of 2026-06-12
- Risk Under Pressure: Compute-Aware Evaluation of Adversarial Robustness in Language Models 9 upvotes, #22 of 2026-06-12
- Visual Para-Thinker++: A Single-Policy Multi-Agent Framework for Visual Reasoning 7 upvotes, #23 of 2026-06-12
- SG-OPD: Sign-Gated On-Policy Distillation via Sign-Consistency Gating and Phased Teacher Sampling 6 upvotes, #24 of 2026-06-12
- Rethinking Psychometric Evaluation of LLMs: When and Why Self-Reports Predict Behavior 6 upvotes, #24 of 2026-06-12
- Evoflux: Inference-Time Evolution of Executable Tool Workflows for Compact Agents 5 upvotes, #26 of 2026-06-12
- MuJoCo-Drones-Gym: A GPU-Accelerated Multi-Drone Simulator for Control and Reinforcement Learning 4 upvotes, #27 of 2026-06-12
- MaskAlign: Token-Subset Representation Alignment for Efficient Diffusion Training 4 upvotes, #27 of 2026-06-12
- Flash-GMM: A Memory-Efficient Kernel for Scalable Soft Clustering 4 upvotes, #27 of 2026-06-12
- EvoBrowseComp: Benchmarking Search Agents on Evolving Knowledge 4 upvotes, #27 of 2026-06-12
- Getting Better at Working With You: Compiling User Corrections into Runtime Enforcement for Coding Agents 4 upvotes, #27 of 2026-06-12
- Surflo: Consistent 3D Surface Flow Model with Global State 4 upvotes, #27 of 2026-06-12
- See What I See, Know What I Think: Dense Latent Communication Across Heterogeneous Agents 3 upvotes, #33 of 2026-06-12
- WEAVER, Better, Faster, Longer: An Effective World Model for Robotic Manipulation 3 upvotes, #33 of 2026-06-12
- The Cold-Start Safety Gap in LLM Agents 2 upvotes, #35 of 2026-06-12
- Revisiting Articulated Parts Perception in Robot Manipulation 2 upvotes, #35 of 2026-06-12
- WebChallenger: A Reliable and Efficient Generalist Web Agent 2 upvotes, #35 of 2026-06-12
- IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder 2 upvotes, #35 of 2026-06-12
- ToolSense: A Diagnostic Framework for Auditing Parametric Tool Knowledge in LLMs 2 upvotes, #35 of 2026-06-12
- ArogyaSutra: A Multi-Agent Framework for Multimodal Medical Reasoning in Indic Languages 2 upvotes, #35 of 2026-06-12
- On the Limits of LLM Adaptability: Impact of Model-Internalized Priors on Annotation Task Performance 1 upvotes, #41 of 2026-06-12
- Leveraging Morphology for Historical Script Metrological Analysis 1 upvotes, #41 of 2026-06-12
- PianoKontext: Expressive Performance Rendering from Deadpan Context 1 upvotes, #41 of 2026-06-12
- A Stationary (and Therefore Compatible) Representation is All You Need 1 upvotes, #41 of 2026-06-12
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.