Daily Papers of 2025-11-19
- VIDEOP2R: Video Understanding from Perception to Reasoning 107 upvotes, #1 of 2025-11-19
- Think-at-Hard: Selective Latent Iterations to Improve Reasoning Language Models 102 upvotes, #2 of 2025-11-19
- AraLingBench A Human-Annotated Benchmark for Evaluating Arabic Linguistic Capabilities of Large Language Models 69 upvotes, #3 of 2025-11-19
- A Style is Worth One Code: Unlocking Code-to-Style Image Generation with Discrete Style Space 58 upvotes, #4 of 2025-11-19
- Large Language Models Meet Extreme Multi-label Classification: Scaling and Multi-modal Framework 37 upvotes, #5 of 2025-11-19
- Can World Simulators Reason? Gen-ViRe: A Generative Visual Reasoning Benchmark 34 upvotes, #6 of 2025-11-19
- REVISOR: Beyond Textual Reflection, Towards Multimodal Introspective Reasoning in Long-Form Video Understanding 24 upvotes, #7 of 2025-11-19
- MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs 24 upvotes, #7 of 2025-11-19
- Orion: A Unified Visual Agent for Multimodal Perception, Advanced Visual Reasoning and Execution 19 upvotes, #9 of 2025-11-19
- OmniZip: Audio-Guided Dynamic Token Compression for Fast Omnimodal Large Language Models 16 upvotes, #10 of 2025-11-19
- Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning 15 upvotes, #11 of 2025-11-19
- ATLAS: A High-Difficulty, Multidisciplinary Benchmark for Frontier Scientific Reasoning 14 upvotes, #12 of 2025-11-19
- Φeat: Physically-Grounded Feature Representation 10 upvotes, #13 of 2025-11-19
- Mitigating Label Length Bias in Large Language Models 6 upvotes, #14 of 2025-11-19
- Proactive Hearing Assistants that Isolate Egocentric Conversations 5 upvotes, #15 of 2025-11-19
- Agent READMEs: An Empirical Study of Context Files for Agentic Coding 5 upvotes, #15 of 2025-11-19
- Error-Driven Scene Editing for 3D Grounding in Large Language Models 4 upvotes, #17 of 2025-11-19
- LLM-Powered Fully Automated Chaos Engineering: Towards Enabling Anyone to Build Resilient Software Systems at Low Cost 3 upvotes, #18 of 2025-11-19
- A Brain Wave Encodes a Thousand Tokens: Modeling Inter-Cortical Neural Interactions for Effective EEG-based Emotion Recognition 3 upvotes, #18 of 2025-11-19
- TopoPerception: A Shortcut-Free Evaluation of Global Visual Perception in Large Vision-Language Models 1 upvotes, #20 of 2025-11-19
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.