Daily Papers of 2026-07-02
- PerceptionRubrics: Calibrating Multimodal Evaluation to Human Perception 41 upvotes, #1 of 2026-07-02
- TurboServe: Serving Streaming Video Generation Efficiently and Economically 34 upvotes, #2 of 2026-07-02
- ELDR: Expert-Locality-Aware Decode Routing for PD-Disaggregated MoE Serving 31 upvotes, #3 of 2026-07-02
- MemSyco-Bench: Benchmarking Sycophancy in Agent Memory 29 upvotes, #4 of 2026-07-02
- Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity 28 upvotes, #5 of 2026-07-02
- Multimodal Continuous Reasoning via Asymmetric Mutual Variational Learning 27 upvotes, #6 of 2026-07-02
- AutoTrainess: Teaching Language Models to Improve Language Models Autonomously 22 upvotes, #7 of 2026-07-02
- ASPIRE: Agentic /Skills Discovery for Robotics 22 upvotes, #7 of 2026-07-02
- Domain Arithmetic: One-Shot VLA Adaptation under Environmental Shifts 22 upvotes, #7 of 2026-07-02
- ABot-M0.5: Unified Mobility-and-Manipulation World Action Model 19 upvotes, #10 of 2026-07-02
- CausalMix: Data Mixture as Causal Inference for Language Model Training 19 upvotes, #10 of 2026-07-02
- Perceive-to-Reason: Decoupling Perception and Reasoning for Fine-Grained Visual Reasoning 18 upvotes, #12 of 2026-07-02
- Valdi: Value Diffusion World Models 15 upvotes, #13 of 2026-07-02
- When LLMs Read Tables Carelessly: Measuring and Reducing Data Referencing Errors 14 upvotes, #14 of 2026-07-02
- BioInsight: Multi-Agent Orchestration for Interactive Biomedical Knowledge Discovery 13 upvotes, #15 of 2026-07-02
- PixelEyes: Decoupling Perception and Reasoning for Pinpoint Visual Evidence Seeking 12 upvotes, #16 of 2026-07-02
- Cross-Domain Generalization Failure in Lightweight Intrusion Detection Models for IIoT Networks 11 upvotes, #17 of 2026-07-02
- Graph-Native Reinforcement Learning Enables Traceable Scientific Hypothesis Generation through Conceptual Recombination 11 upvotes, #17 of 2026-07-02
- The State-Prediction Separation Hypothesis 11 upvotes, #17 of 2026-07-02
- Personalization as Inverse Planning: Learning Latent Design Intents for Agentic Slide Generation via Structural Denoising 10 upvotes, #20 of 2026-07-02
- NoPA: Non-Parametric Online 3D Scene Graph Generation 9 upvotes, #21 of 2026-07-02
- AI translation of literary texts is "fine", but readers still prefer human translations 8 upvotes, #22 of 2026-07-02
- Building to the Test: Coding Agents Deliver What You Check, Not What You Requested 8 upvotes, #22 of 2026-07-02
- When More Sampling Hurts: The Modal Ceiling and Correlation Ceiling of Test-Time Scaling 8 upvotes, #22 of 2026-07-02
- AtomiMed: Hierarchical Atomic Fact-Checking for Universal Clinical-Aware Medical Report Evaluation 8 upvotes, #22 of 2026-07-02
- GRPO, Dr. GRPO, and DAPO Are Three Operations on One Number: The Group-Standard-Deviation Identity 8 upvotes, #22 of 2026-07-02
- Are Performance-Optimization Benchmarks Reliably Measuring Coding Agents? 8 upvotes, #22 of 2026-07-02
- Rank-Aware Hyperbolic Alignment for Vision-Language Dataset Distillation 7 upvotes, #28 of 2026-07-02
- SciIR: A Large-scale Training Dataset and Benchmark for Scientific Image Reasoning Generation 7 upvotes, #28 of 2026-07-02
- Seeing Is Not Sharing: Some Vision-Language Models Overestimate Common Ground in Asymmetric Dialogue 7 upvotes, #28 of 2026-07-02
- Autonomous Scientific Discovery via Iterative Meta-Reflection 7 upvotes, #28 of 2026-07-02
- CogSENet: Blind Image Deblurring with Blur-Conditioned Semantic Routing and Explicit Frequency Fusion 6 upvotes, #32 of 2026-07-02
- HealthAgentBench: A Unified Benchmark Suite of Realistic Agentic Healthcare Environments for Challenging Frontier AI Agents 6 upvotes, #32 of 2026-07-02
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.