Daily Papers of 2026-05-22
- DelTA: Discriminative Token Credit Assignment for Reinforcement Learning from Verifiable Rewards 204 upvotes, #1 of 2026-05-22
- TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generation 174 upvotes, #2 of 2026-05-22
- Perception or Prejudice: Can MLLMs Go Beyond First Impressions of Personality? 169 upvotes, #3 of 2026-05-22
- π-Bench: Evaluating Proactive Personal Assistant Agents in Long-Horizon Workflows 102 upvotes, #4 of 2026-05-22
- Full Attention Strikes Back: Transferring Full Attention into Sparse within Hundred Training Steps 93 upvotes, #5 of 2026-05-22
- ACC: Compiling Agent Trajectories for Long-Context Training 59 upvotes, #6 of 2026-05-22
- PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects 51 upvotes, #7 of 2026-05-22
- LatentOmni: Rethinking Omni-Modal Understanding via Unified Audio-Visual Latent Reasoning 46 upvotes, #8 of 2026-05-22
- Forecasting Scientific Progress with Artificial Intelligence 42 upvotes, #9 of 2026-05-22
- WorldKV: Efficient World Memory with World Retrieval and Compression 41 upvotes, #10 of 2026-05-22
- SEGA: Spectral-Energy Guided Attention for Resolution Extrapolation in Diffusion Transformers 40 upvotes, #11 of 2026-05-22
- Spreadsheet-RL: Advancing Large Language Model Agents on Realistic Spreadsheet Tasks via Reinforcement Learning 35 upvotes, #12 of 2026-05-22
- Gated DeltaNet-2: Decoupling Erase and Write in Linear Attention 30 upvotes, #13 of 2026-05-22
- FlowLong: Inference-time Long Video Generation via Manifold-constrained Tweedie Matching 29 upvotes, #14 of 2026-05-22
- SpaceDG: Benchmarking Spatial Intelligence under Visual Degradation 28 upvotes, #15 of 2026-05-22
- Sensor2Sensor: Cross-Embodiment Sensor Conversion for Autonomous Driving 27 upvotes, #16 of 2026-05-22
- Unsupervised Process Reward Models 25 upvotes, #17 of 2026-05-22
- Q-ARVD: Quantizing Autoregressive Video Diffusion Models 21 upvotes, #18 of 2026-05-22
- Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles 20 upvotes, #19 of 2026-05-22
- AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment 19 upvotes, #20 of 2026-05-22
- Training Large Language Models to Predict Clinical Events 17 upvotes, #21 of 2026-05-22
- Forecasting Downstream Performance of LLMs With Proxy Metrics 14 upvotes, #22 of 2026-05-22
- GenEvolve: Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation 13 upvotes, #23 of 2026-05-22
- KVServe: Service-Aware KV Cache Compression for Communication-Efficient Disaggregated LLM Serving 12 upvotes, #24 of 2026-05-22
- ClinSeekAgent: Automating Multimodal Evidence Seeking for Agentic Clinical Reasoning 12 upvotes, #24 of 2026-05-22
- Bernini: Latent Semantic Planning for Video Diffusion 12 upvotes, #24 of 2026-05-22
- Efficient Agentic Reasoning Through Self-Regulated Simulative Planning 11 upvotes, #27 of 2026-05-22
- Swift Sampling: Selecting Temporal Surprises via Taylor Series 11 upvotes, #27 of 2026-05-22
- One Sentence, One Drama: Personalized Short-Form Drama Generation via Multi-Agent Systems 10 upvotes, #29 of 2026-05-22
- LoREnc: Low-Rank Encryption for Securing Foundation Models and LoRA Adapters 9 upvotes, #30 of 2026-05-22
- TerminalWorld: Benchmarking Agents on Real-World Terminal Tasks 9 upvotes, #30 of 2026-05-22
- Rule2DRC: Benchmarking LLM Agents for DRC Script Synthesis with Execution-Guided Test Generation 8 upvotes, #32 of 2026-05-22
- SceneAligner: 3D-Grounded Floorplan Localization in the Wild 8 upvotes, #32 of 2026-05-22
- "I didn't Make the Micro Decisions": Measuring, Inducing, and Exposing Goal-Level AI Contributions in Collaboration 6 upvotes, #34 of 2026-05-22
- Segment Anything with Motion, Geometry, and Semantic Adaptation for Complex Nonlinear Visual Object Tracking 6 upvotes, #34 of 2026-05-22
- Diversed Model Discovery via Structured Table Discovery 6 upvotes, #34 of 2026-05-22
- OmniPro: A Comprehensive Benchmark for Omni-Proactive Streaming Video Understanding 5 upvotes, #37 of 2026-05-22
- DecQ: Detail-Condensing Queries for Enhanced Reconstruction and Generation in Representation Autoencoders 5 upvotes, #37 of 2026-05-22
- SAM 3D Animal: Promptable Animal 3D Reconstruction from Images in the Wild 4 upvotes, #39 of 2026-05-22
- Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search 4 upvotes, #39 of 2026-05-22
- Platonic Representations in the Human Brain: Unsupervised Recovery of Universal Geometry 4 upvotes, #39 of 2026-05-22
- Same Architecture, Different Capacity: Optimizer-Induced Spectral Scaling Laws 4 upvotes, #39 of 2026-05-22
- From Reasoning Chains to Verifiable Subproblems: Curriculum Reinforcement Learning Enables Credit Assignment for LLM Reasoning 4 upvotes, #39 of 2026-05-22
- More Context, Larger Models, or Moral Knowledge? A Systematic Study of Schwartz Value Detection in Political Texts 4 upvotes, #39 of 2026-05-22
- AnyMo: Geometry-Aware Setup-Agnostic Modeling of Human Motion in the Wild 4 upvotes, #39 of 2026-05-22
- Disentangling Sampling from Training Budget in Class-Imbalanced CT Body Composition Segmentation 3 upvotes, #46 of 2026-05-22
- FashionLens: Toward Versatile Fashion Image Retrieval via Task-Adaptive Learning 2 upvotes, #47 of 2026-05-22
- Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators 2 upvotes, #47 of 2026-05-22
- Minimalist Visual Inertial Odometry 1 upvotes, #49 of 2026-05-22
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.