Daily Papers of 2025-11-20
- Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation 209 upvotes, #1 of 2025-11-20
- Reasoning via Video: The First Evaluation of Video Models' Reasoning Abilities through Maze-Solving Tasks 72 upvotes, #2 of 2025-11-20
- What Does It Take to Be a Good AI Research Agent? Studying the Role of Ideation Diversity 54 upvotes, #3 of 2025-11-20
- VisPlay: Self-Evolving Vision-Language Models from Images 41 upvotes, #4 of 2025-11-20
- Instruction-Guided Lesion Segmentation for Chest X-rays with Automatically Generated Large-Scale Dataset 25 upvotes, #5 of 2025-11-20
- ARC-Chapter: Structuring Hour-Long Videos into Navigable Chapters and Hierarchical Summaries 16 upvotes, #6 of 2025-11-20
- MHR: Momentum Human Rig 13 upvotes, #7 of 2025-11-20
- Mixture of States: Routing Token-Level Dynamics for Multimodal Generation 6 upvotes, #8 of 2025-11-20
- FreeAskWorld: An Interactive and Closed-Loop Simulator for Human-Centric Embodied AI 6 upvotes, #8 of 2025-11-20
- RoMa v2: Harder Better Faster Denser Feature Matching 6 upvotes, #8 of 2025-11-20
- Aligning Generative Music AI with Human Preferences: Methods and Challenges 2 upvotes, #11 of 2025-11-20
- Medal S: Spatio-Textual Prompt Model for Medical Segmentation 1 upvotes, #12 of 2025-11-20
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.