Daily Papers of 2026-04-16
- Seedance 2.0: Advancing Video Generation for World Complexity 151 upvotes, #1 of 2026-04-16
- RationalRewards: Reasoning Rewards Scale Visual Generation Both Training and Test Time 100 upvotes, #2 of 2026-04-16
- OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models 62 upvotes, #3 of 2026-04-16
- SpatialEvo: Self-Evolving Spatial Intelligence via Deterministic Geometric Environments 62 upvotes, #3 of 2026-04-16
- Memory Transfer Learning: How Memories are Transferred Across Domains in Coding Agents 29 upvotes, #5 of 2026-04-16
- From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space 28 upvotes, #6 of 2026-04-16
- Sema Code: Decoupling AI Coding Agents into Programmable, Embeddable Infrastructure 24 upvotes, #7 of 2026-04-16
- Exploration and Exploitation Errors Are Measurable for Language Model Agents 24 upvotes, #7 of 2026-04-16
- Target Policy Optimization 22 upvotes, #9 of 2026-04-16
- SemaClaw: A Step Towards General-Purpose Personal AI Agents through Harness Engineering 20 upvotes, #10 of 2026-04-16
- Geometric Context Transformer for Streaming 3D Reconstruction 17 upvotes, #11 of 2026-04-16
- Free Geometry: Refining 3D Reconstruction from Longer Versions of Itself 15 upvotes, #12 of 2026-04-16
- LangFlow: Continuous Diffusion Rivals Discrete in Language Modeling 14 upvotes, #13 of 2026-04-16
- Anthropogenic Regional Adaptation in Multimodal Vision-Language Model 13 upvotes, #14 of 2026-04-16
- TIP: Token Importance in On-Policy Distillation 13 upvotes, #14 of 2026-04-16
- TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration 13 upvotes, #14 of 2026-04-16
- ROSE: An Intent-Centered Evaluation Metric for NL2SQL 11 upvotes, #17 of 2026-04-16
- Self-Distillation Zero: Self-Revision Turns Binary Rewards into Dense Supervision 10 upvotes, #18 of 2026-04-16
- UI-Zoomer: Uncertainty-Driven Adaptive Zoom-In for GUI Grounding 10 upvotes, #18 of 2026-04-16
- SkVM: Compiling Skills for Efficient Execution Everywhere 9 upvotes, #20 of 2026-04-16
- ReconPhys: Reconstruct Appearance and Physical Attributes from Single Video 9 upvotes, #20 of 2026-04-16
- Narrative-Driven Paper-to-Slide Generation via ArcDeck 7 upvotes, #22 of 2026-04-16
- HDR Video Generation via Latent Alignment with Logarithmic Encoding 6 upvotes, #23 of 2026-04-16
- MERRIN: A Benchmark for Multimodal Evidence Retrieval and Reasoning in Noisy Web Environments 6 upvotes, #23 of 2026-04-16
- UI-Copilot: Advancing Long-Horizon GUI Automation via Tool-Integrated Policy Optimization 6 upvotes, #23 of 2026-04-16
- Self-Sovereign Agent 5 upvotes, #26 of 2026-04-16
- What do Language Models Learn and When? The Implicit Curriculum Hypothesis 4 upvotes, #27 of 2026-04-16
- Mobile GUI Agents under Real-world Threats: Are We There Yet? 3 upvotes, #28 of 2026-04-16
- Do AI Coding Agents Log Like Humans? An Empirical Study 3 upvotes, #28 of 2026-04-16
- ROSE: Retrieval-Oriented Segmentation Enhancement 3 upvotes, #28 of 2026-04-16
- InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysis 2 upvotes, #31 of 2026-04-16
- A Temporally Augmented Graph Attention Network for Affordance Classification 1 upvotes, #32 of 2026-04-16
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.