Daily Papers of 2025-10-09
- Cache-to-Cache: Direct Semantic Communication Between Large Language Models 89 upvotes, #1 of 2025-10-09
- Ming-UniVision: Joint Image Understanding and Generation with a Unified Continuous Tokenizer 69 upvotes, #2 of 2025-10-09
- Lumina-DiMOO: An Omni Diffusion Large Language Model for Multi-Modal Generation and Understanding 49 upvotes, #3 of 2025-10-09
- RLinf-VLA: A Unified and Efficient Framework for VLA+RL Training 35 upvotes, #4 of 2025-10-09
- MATRIX: Mask Track Alignment for Interaction-aware Video Generation 35 upvotes, #4 of 2025-10-09
- SHANKS: Simultaneous Hearing and Thinking for Spoken Language Models 34 upvotes, #6 of 2025-10-09
- Vibe Checker: Aligning Code Evaluation with Human Preference 30 upvotes, #7 of 2025-10-09
- Multi-Agent Tool-Integrated Policy Optimization 29 upvotes, #8 of 2025-10-09
- The Markovian Thinker 27 upvotes, #9 of 2025-10-09
- Artificial Hippocampus Networks for Efficient Long-Context Modeling 26 upvotes, #10 of 2025-10-09
- Pushing on Multilingual Reasoning Models with Language-Mixed Chain-of-Thought 24 upvotes, #11 of 2025-10-09
- Why Low-Precision Transformer Training Fails: An Analysis on Flash Attention 22 upvotes, #12 of 2025-10-09
- The African Languages Lab: A Collaborative Approach to Advancing Low-Resource African NLP 22 upvotes, #12 of 2025-10-09
- OBS-Diff: Accurate Pruning For Diffusion Models in One-Shot 21 upvotes, #14 of 2025-10-09
- Revisiting Long-context Modeling from Context Denoising Perspective 20 upvotes, #15 of 2025-10-09
- CALM Before the STORM: Unlocking Native Reasoning for Optimization Modeling 19 upvotes, #16 of 2025-10-09
- Native Hybrid Attention for Efficient Sequence Modeling 16 upvotes, #17 of 2025-10-09
- When Benchmarks Age: Temporal Misalignment through Large Language Model Factuality Evaluation 14 upvotes, #18 of 2025-10-09
- StaMo: Unsupervised Learning of Generalizable Robot Motion from Compact State Representation 12 upvotes, #19 of 2025-10-09
- Patch-as-Decodable-Token: Towards Unified Multi-Modal Vision Tasks in MLLMs 11 upvotes, #20 of 2025-10-09
- TTRV: Test-Time Reinforcement Learning for Vision Language Models 11 upvotes, #20 of 2025-10-09
- Are We Using the Right Benchmark: An Evaluation Framework for Visual Token Compression Methods 11 upvotes, #20 of 2025-10-09
- Reinforcement Mid-Training 8 upvotes, #23 of 2025-10-09
- NorMuon: Making Muon more efficient and scalable 6 upvotes, #24 of 2025-10-09
- Revisiting the Uniform Information Density Hypothesis in LLM Reasoning Traces 6 upvotes, #24 of 2025-10-09
- G^2RPO: Granular GRPO for Precise Reward in Flow Models 5 upvotes, #26 of 2025-10-09
- MLE-Smith: Scaling MLE Tasks with Automated Multi-Agent Pipeline 5 upvotes, #26 of 2025-10-09
- WristWorld: Generating Wrist-Views via 4D World Models for Robotic Manipulation 5 upvotes, #26 of 2025-10-09
- AlphaApollo: Orchestrating Foundation Models and Professional Tools into a Self-Evolving System for Deep Agentic Reasoning 4 upvotes, #29 of 2025-10-09
- Bridging Text and Video Generation: A Survey 3 upvotes, #30 of 2025-10-09
- A Single Character can Make or Break Your LLM Evals 3 upvotes, #30 of 2025-10-09
- Code Agent can be an End-to-end System Hacker: Benchmarking Real-world Threats of Computer-use Agent 3 upvotes, #30 of 2025-10-09
- Heptapod: Language Modeling on Visual Signals 3 upvotes, #30 of 2025-10-09
- Online Generic Event Boundary Detection 3 upvotes, #30 of 2025-10-09
- Beyond Monolingual Assumptions: A Survey of Code-Switched NLP in the Era of Large Language Models 3 upvotes, #30 of 2025-10-09
- U-Bench: A Comprehensive Understanding of U-Net through 100-Variant Benchmarking 3 upvotes, #30 of 2025-10-09
- DeepTravel: An End-to-End Agentic Reinforcement Learning Framework for Autonomous Travel Planning Agents 2 upvotes, #37 of 2025-10-09
- FinLFQA: Evaluating Attributed Text Generation of LLMs in Financial Long-Form Question Answering 2 upvotes, #37 of 2025-10-09
- M3Retrieve: Benchmarking Multimodal Retrieval for Medicine 2 upvotes, #37 of 2025-10-09
- TRAVL: A Recipe for Making Video-Language Models Better Judges of Physics Implausibility 2 upvotes, #37 of 2025-10-09
- D^3QE: Learning Discrete Distribution Discrepancy-aware Quantization Error for Autoregressive-Generated Image Detection 1 upvotes, #41 of 2025-10-09
- PuzzlePlex: Benchmarking Foundation Models on Reasoning and Planning with Puzzles 1 upvotes, #41 of 2025-10-09
- Glocal Information Bottleneck for Time Series Imputation 1 upvotes, #43 of 2025-10-09
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.