Daily Papers of 2026-04-28
- From Skills to Talent: Organising Heterogeneous Agents as a Real-World Company 99 upvotes, #1 of 2026-04-28
- World-R1: Reinforcing 3D Constraints for Text-to-Video Generation 34 upvotes, #2 of 2026-04-28
- ReVSI: Rebuilding Visual Spatial Intelligence Evaluation for Accurate Assessment of VLM 3D Reasoning 31 upvotes, #3 of 2026-04-28
- Vision-Language-Action Safety: Threats, Challenges, Evaluations, and Mechanisms 26 upvotes, #4 of 2026-04-28
- Why Fine-Tuning Encourages Hallucinations and How to Fix It 25 upvotes, #5 of 2026-04-28
- ClawMark: A Living-World Benchmark for Multi-Turn, Multi-Day, Multimodal Coworker Agents 20 upvotes, #6 of 2026-04-28
- For-Value: Efficient Forward-Only Data Valuation for finetuning LLMs and VLMs 19 upvotes, #7 of 2026-04-28
- Sapiens2 16 upvotes, #8 of 2026-04-28
- Rewarding the Scientific Process: Process-Level Reward Modeling for Agentic Data Analysis 14 upvotes, #9 of 2026-04-28
- Towards Understanding the Robustness of Sparse Autoencoders 11 upvotes, #10 of 2026-04-28
- SketchVLM: Vision language models can annotate images to explain thoughts and guide users 11 upvotes, #10 of 2026-04-28
- UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models 9 upvotes, #12 of 2026-04-28
- Taming Actor-Observer Asymmetry in Agents via Dialectical Alignment 9 upvotes, #12 of 2026-04-28
- How Much Is One Recurrence Worth? Iso-Depth Scaling Laws for Looped Language Models 9 upvotes, #12 of 2026-04-28
- Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation 8 upvotes, #15 of 2026-04-28
- Efficient Agent Evaluation via Diversity-Guided User Simulation 7 upvotes, #16 of 2026-04-28
- TexOCR: Advancing Document OCR Models for Compilable Page-to-LaTeX Reconstruction 6 upvotes, #17 of 2026-04-28
- Learning to Identify Out-of-Distribution Objects for 3D LiDAR Anomaly Segmentation 6 upvotes, #17 of 2026-04-28
- PageGuide: Browser extension to assist users in navigating a webpage and locating information 6 upvotes, #17 of 2026-04-28
- ATTN-FIQA: Interpretable Attention-based Face Image Quality Assessment with Vision Transformers 4 upvotes, #20 of 2026-04-28
- RaV-IDP: A Reconstruction-as-Validation Framework for Faithful Intelligent Document Processing 4 upvotes, #20 of 2026-04-28
- Stabilizing Efficient Reasoning with Step-Level Advantage Selection 4 upvotes, #20 of 2026-04-28
- IndustryAssetEQA: A Neurosymbolic Operational Intelligence System for Embodied Question Answering in Industrial Asset Maintenance 3 upvotes, #23 of 2026-04-28
- OmniShotCut: Holistic Relational Shot Boundary Detection with Shot-Query Transformer 3 upvotes, #23 of 2026-04-28
- Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining 2 upvotes, #25 of 2026-04-28
- EX-FIQA: Leveraging Intermediate Early eXit Representations from Vision Transformers for Face Image Quality Assessment 2 upvotes, #25 of 2026-04-28
- Discovering Agentic Safety Specifications from 1-Bit Danger Signals 2 upvotes, #25 of 2026-04-28
- Improving Robustness of Tabular Retrieval via Representational Stability 2 upvotes, #25 of 2026-04-28
- Improving Vision-language Models with Perception-centric Process Reward Models 2 upvotes, #25 of 2026-04-28
- ProEval: Proactive Failure Discovery and Efficient Performance Estimation for Generative AI Evaluation 1 upvotes, #30 of 2026-04-28
- Personality Shapes Gender Bias in Persona-Conditioned LLM Narratives Across English and Hindi: An Empirical Investigation 1 upvotes, #30 of 2026-04-28
- Credal Concept Bottleneck Models for Epistemic-Aleatoric Uncertainty Decomposition 1 upvotes, #30 of 2026-04-28
- Quantum Kernel Advantage over Classical Collapse in Medical Foundation Model Embeddings 1 upvotes, #30 of 2026-04-28
- Stochastic KV Routing: Enabling Adaptive Depth-Wise Cache Sharing 8 upvotes, #34 of 2026-04-28
- Zero-to-CAD: Agentic Synthesis of Interpretable CAD Programs at Million-Scale Without Real Data 9 upvotes, #34 of 2026-04-28
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.