Daily Papers of 2026-07-10
- Vidu S1: A Real-Time Interactive Video Generation Model 138 upvotes, #1 of 2026-07-10
- Video-Oasis: Rethinking Evaluation of Video Understanding 64 upvotes, #2 of 2026-07-10
- Why Can't I Open My Drawer? Mitigating Object-Driven Shortcuts in Zero-Shot Compositional Action Recognition 54 upvotes, #3 of 2026-07-10
- Ideas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea Generation 39 upvotes, #4 of 2026-07-10
- LongE2V: Long-Horizon Event-based Video Reconstruction, Prediction, and Frame Interpolation with Video Diffusion Models 35 upvotes, #5 of 2026-07-10
- UniClawBench: A Universal Benchmark for Proactive Agents on Real-World Tasks 34 upvotes, #6 of 2026-07-10
- OpenCoF: Learning to Reason Through Video Generation 28 upvotes, #7 of 2026-07-10
- Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE 23 upvotes, #8 of 2026-07-10
- Enhancing In-context Panoramic Generation via Geometric-aware Pretraining 22 upvotes, #9 of 2026-07-10
- DrugGen 2: A disease-aware language model for enhancing drug discovery 20 upvotes, #10 of 2026-07-10
- CineMobile: On-Device Image-to-Video Diffusion for Cinematic Camera Motion Generation 18 upvotes, #11 of 2026-07-10
- Linear Attention Architectures: Mechanisms, Trade-offs, and Cross-Layer Routing 14 upvotes, #12 of 2026-07-10
- Remember When It Matters: Proactive Memory Agent for Long-Horizon Agents 14 upvotes, #12 of 2026-07-10
- ARDY: Autoregressive Diffusion with Hybrid Representation for Interactive Human Motion Generation 10 upvotes, #14 of 2026-07-10
- Flash-BoN: Instant Drafts for Inference-Time Scaling in Diffusion Models 9 upvotes, #15 of 2026-07-10
- UP: Unbounded Positive Asymmetric Optimization for Breaking the Exploration-Stability Dilemma 9 upvotes, #15 of 2026-07-10
- A Sparse and Truncated State Vector Simulator for Peaked Circuits 9 upvotes, #15 of 2026-07-10
- SAM-MT: Real-Time Interactive Multi-Target Video Segmentation 9 upvotes, #15 of 2026-07-10
- PhyMRI-SR: Toward Physics-Aware MRI Image Super-Resolution 5 upvotes, #19 of 2026-07-10
- Can Dialects Be Steered Like Languages? Sparse Neurons and Distributed Directions in Arabic LLMs 4 upvotes, #20 of 2026-07-10
- CausalDS: Benchmarking Causal Reasoning in Data-Science Agents 4 upvotes, #20 of 2026-07-10
- A Quantized Native Runtime for On-Device Semantic Audio Generation 4 upvotes, #20 of 2026-07-10
- PAST-TIDE: Prototype-Anchored Statement Tuning with Topic-Invariant Normalization for Stance Detection 3 upvotes, #23 of 2026-07-10
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.