Daily Papers of 2026-04-03
- DataFlex: A Unified Framework for Data-Centric Dynamic Training of Large Language Models 345 upvotes, #1 of 2026-04-03
- The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook 136 upvotes, #2 of 2026-04-03
- Generative World Renderer 101 upvotes, #3 of 2026-04-03
- SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization 92 upvotes, #4 of 2026-04-03
- CORAL: Towards Autonomous Multi-Agent Evolution for Open-Ended Discovery 52 upvotes, #5 of 2026-04-03
- VOID: Video Object and Interaction Deletion 51 upvotes, #6 of 2026-04-03
- Steerable Visual Representations 51 upvotes, #6 of 2026-04-03
- EgoSim: Egocentric World Simulator for Embodied Interaction Generation 36 upvotes, #8 of 2026-04-03
- Therefore I am. I Think 30 upvotes, #9 of 2026-04-03
- LatentUM: Unleashing the Potential of Interleaved Cross-Modal Reasoning via a Latent-Space Unified Model 30 upvotes, #9 of 2026-04-03
- Omni-SimpleMem: Autoresearch-Guided Discovery of Lifelong Multimodal Agent Memory 29 upvotes, #11 of 2026-04-03
- NearID: Identity Representation Learning via Near-identity Distractors 29 upvotes, #11 of 2026-04-03
- UniDriveVLA: Unifying Understanding, Perception, and Action Planning for Autonomous Driving 25 upvotes, #13 of 2026-04-03
- ASI-Evolve: AI Accelerates AI 24 upvotes, #14 of 2026-04-03
- Gated Condition Injection without Multimodal Attention: Towards Controllable Linear-Attention Transformers 18 upvotes, #15 of 2026-04-03
- Investigating Autonomous Agent Contributions in the Wild: Activity Patterns and Code Change over Time 17 upvotes, #16 of 2026-04-03
- GPA: Learning GUI Process Automation from Demonstrations 16 upvotes, #17 of 2026-04-03
- Video Models Reason Early: Exploiting Plan Commitment for Maze Solving 14 upvotes, #18 of 2026-04-03
- AIBench: Evaluating Visual-Logical Consistency in Academic Illustration Generation 13 upvotes, #19 of 2026-04-03
- Omni123: Exploring 3D Native Foundation Models with Limited 3D Data by Unifying Text to 2D and 3D Generation 13 upvotes, #19 of 2026-04-03
- VideoZeroBench: Probing the Limits of Video MLLMs with Spatio-Temporal Evidence Verification 12 upvotes, #21 of 2026-04-03
- Tex3D: Objects as Attack Surfaces via Adversarial 3D Textures for Vision-Language-Action Models 12 upvotes, #21 of 2026-04-03
- Woosh: A Sound Effects Foundation Model 12 upvotes, #21 of 2026-04-03
- Apriel-Reasoner: RL Post-Training for General-Purpose and Efficient Reasoning 12 upvotes, #21 of 2026-04-03
- AutoMIA: Improved Baselines for Membership Inference Attack via Agentic Self-Exploration 11 upvotes, #25 of 2026-04-03
- T5Gemma-TTS Technical Report 11 upvotes, #25 of 2026-04-03
- MDPBench: A Benchmark for Multilingual Document Parsing in Real-World Scenarios 10 upvotes, #27 of 2026-04-03
- DynaVid: Learning to Generate Highly Dynamic Videos using Synthetic Motion Data 10 upvotes, #27 of 2026-04-03
- Forecasting Supply Chain Disruptions with Foresight Learning 9 upvotes, #29 of 2026-04-03
- Efficient and Principled Scientific Discovery through Bayesian Optimization: A Tutorial 9 upvotes, #29 of 2026-04-03
- Ask or Assume? Uncertainty-Aware Clarification-Seeking in Coding Agents 8 upvotes, #31 of 2026-04-03
- LinguDistill: Recovering Linguistic Ability in Vision- Language Models via Selective Cross-Modal Distillation 8 upvotes, #31 of 2026-04-03
- Working Notes on Late Interaction Dynamics: Analyzing Targeted Behaviors of Late Interaction Models 7 upvotes, #33 of 2026-04-03
- Signals: Trajectory Sampling and Triage for Agentic Interactions 7 upvotes, #33 of 2026-04-03
- UniRecGen: Unifying Multi-View 3D Reconstruction and Generation 7 upvotes, #33 of 2026-04-03
- Memory-Augmented Vision-Language Agents for Persistent and Semantically Consistent Object Captioning 6 upvotes, #36 of 2026-04-03
- LOME: Learning Human-Object Manipulation with Action-Conditioned Egocentric World Model 6 upvotes, #36 of 2026-04-03
- Executing as You Generate: Hiding Execution Latency in LLM Code Generation 6 upvotes, #36 of 2026-04-03
- FlowSlider: Training-Free Continuous Image Editing via Fidelity-Steering Decomposition 6 upvotes, #36 of 2026-04-03
- ActionParty: Multi-Subject Action Binding in Generative Video Games 6 upvotes, #36 of 2026-04-03
- MultiGen: Level-Design for Editable Multiplayer Worlds in Diffusion Game Engines 5 upvotes, #41 of 2026-04-03
- An Empirical Recipe for Universal Phone Recognition 5 upvotes, #41 of 2026-04-03
- Brainstacks: Cross-Domain Cognitive Capabilities via Frozen MoE-LoRA Stacks for Continual LLM Learning 5 upvotes, #41 of 2026-04-03
- Friends and Grandmothers in Silico: Localizing Entity Cells in Language Models 5 upvotes, #41 of 2026-04-03
- Automatic Image-Level Morphological Trait Annotation for Organismal Images 5 upvotes, #41 of 2026-04-03
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.