Daily Papers of 2026-07-30
- TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM 139 upvotes, #1 of 2026-07-30
- DistillAlign: Coordinating Mode Covering and Mode Seeking in Autoregressive Video Distillation 92 upvotes, #2 of 2026-07-30
- CoRT: Counterfactual Replay for Token-Level Rubric-Guided Policy Optimization 83 upvotes, #3 of 2026-07-30
- HumanCLAW: Can Vision-Language Models Act Through a Body? 76 upvotes, #4 of 2026-07-30
- DecoEvo: Score-Decoupled Co-Evolution of Solver and Rubric-Generator Skills in Text Space 65 upvotes, #5 of 2026-07-30
- CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition 49 upvotes, #6 of 2026-07-30
- CAST: Game Solvers as Turn-Level Teachers for LLM Agents 41 upvotes, #7 of 2026-07-30
- SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution 28 upvotes, #8 of 2026-07-30
- MindForge: Teaching Small Language Models Whole-Life-Cycle Software Engineering via Source-Free Program Synthesis 27 upvotes, #9 of 2026-07-30
- SpecFirst: Behavioral Specification Elicitation as a First-Class Step in Agent-Based Program Synthesis from Scratch 20 upvotes, #10 of 2026-07-30
- StatePlay: State-Aware Game World Models for Mechanics-Consistent Generation 18 upvotes, #11 of 2026-07-30
- Can AI agents conduct open-ended AI research? Early evidence from two case studies 18 upvotes, #11 of 2026-07-30
- OVEarth-Bench: Evaluating Category Breadth and Query Diversity for Open-Vocabulary Earth Observation 14 upvotes, #13 of 2026-07-30
- Memory for Large Language Models 13 upvotes, #14 of 2026-07-30
- OmegaUse-OfficeVal: Benchmarking LLM Agents on Long-Horizon Office-Suite Tasks with Economic Grounding 13 upvotes, #14 of 2026-07-30
- Voice Memory for Agentic Speech Recognition 11 upvotes, #16 of 2026-07-30
- Explicit Layer Modeling for Video Object Insertion and Layer Decomposition 8 upvotes, #17 of 2026-07-30
- GPT-Red: Automated Red Teaming via Self-Play at Scale 7 upvotes, #18 of 2026-07-30
- CADENCE: Closing the Reasoning Gap via Coverage-Adaptive On-Policy Distillation 6 upvotes, #19 of 2026-07-30
- πR^2: Reactive Real-time Flow Policies 5 upvotes, #20 of 2026-07-30
- Grading the Narrators: An Isnad-Rijal Framework for Claim-Level Provenance in Multi-Agent Knowledge Systems 4 upvotes, #21 of 2026-07-30
- StealthBench: Measuring Operational Stealth in Autonomous Offensive-Security Agents 4 upvotes, #21 of 2026-07-30
- SecRespond: Benchmarking AI Agents for Real-World Post-Compromise Incident Response 4 upvotes, #21 of 2026-07-30
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.