Daily Papers of 2026-02-13
- The Devil Behind Moltbook: Anthropic Safety is Always Vanishing in Self-Evolving AI Societies 188 upvotes, #1 of 2026-02-13
- Composition-RL: Compose Your Verifiable Prompts for Reinforcement Learning of Large Language Models 92 upvotes, #2 of 2026-02-13
- DeepGen 1.0: A Lightweight Unified Multimodal Model for Advancing Image Generation and Editing 78 upvotes, #3 of 2026-02-13
- Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation 57 upvotes, #4 of 2026-02-13
- GigaBrain-0.5M*: a VLA That Learns From World Model-Based Reinforcement Learning 55 upvotes, #5 of 2026-02-13
- MOSS-Audio-Tokenizer: Scaling Audio Tokenizers for Future Audio Foundation Models 49 upvotes, #6 of 2026-02-13
- NarraScore: Bridging Visual Narrative and Musical Dynamics via Hierarchical Affective Control 43 upvotes, #7 of 2026-02-13
- LawThinker: A Deep Research Legal Agent in Dynamic Environments 34 upvotes, #8 of 2026-02-13
- Thinking with Drafting: Optical Decompression via Logical Reconstruction 32 upvotes, #9 of 2026-02-13
- Stroke of Surprise: Progressive Semantic Illusions in Vector Sketching 31 upvotes, #10 of 2026-02-13
- Think Longer to Explore Deeper: Learn to Explore In-Context via Length-Incentivized Reinforcement Learning 30 upvotes, #11 of 2026-02-13
- RISE: Self-Improving Robot Policy with Compositional World Model 28 upvotes, #12 of 2026-02-13
- χ_{0}: Resource-Aware Robust Manipulation via Taming Distributional Inconsistencies 25 upvotes, #13 of 2026-02-13
- EgoHumanoid: Unlocking In-the-Wild Loco-Manipulation with Robot-Free Egocentric Demonstration 20 upvotes, #14 of 2026-02-13
- dVoting: Fast Voting for dLLMs 20 upvotes, #14 of 2026-02-13
- Sparse Video Generation Propels Real-World Beyond-the-View Vision-Language Navigation 18 upvotes, #16 of 2026-02-13
- Voxtral Realtime 15 upvotes, #17 of 2026-02-13
- DeepSight: An All-in-One LM Safety Toolkit 13 upvotes, #18 of 2026-02-13
- Unveiling Implicit Advantage Symmetry: Why GRPO Struggles with Exploration and Difficulty Adaptation 12 upvotes, #19 of 2026-02-13
- Adapting Vision-Language Models for E-commerce Understanding at Scale 12 upvotes, #19 of 2026-02-13
- Gaia2: Benchmarking LLM Agents on Dynamic and Asynchronous Environments 12 upvotes, #19 of 2026-02-13
- PISCO: Precise Video Instance Insertion with Sparse Control 11 upvotes, #22 of 2026-02-13
- T3D: Few-Step Diffusion Language Models via Trajectory Self-Distillation with Direct Discriminative Optimization 8 upvotes, #23 of 2026-02-13
- MemFly: On-the-Fly Memory Optimization via Information Bottleneck 7 upvotes, #24 of 2026-02-13
- ThinkRouter: Efficient Reasoning via Routing Thinking between Latent and Discrete Spaces 7 upvotes, #24 of 2026-02-13
- Single-minus gluon tree amplitudes are nonzero 7 upvotes, #24 of 2026-02-13
- Dreaming in Code for Curriculum Learning in Open-Ended Worlds 6 upvotes, #27 of 2026-02-13
- MiniCPM-SALA: Hybridizing Sparse and Linear Attention for Efficient Long-Context Modeling 6 upvotes, #27 of 2026-02-13
- MolmoSpaces: A Large-Scale Open Ecosystem for Robot Navigation and Manipulation 5 upvotes, #29 of 2026-02-13
- MetaphorStar: Image Metaphor Understanding and Reasoning with End-to-End Visual Reinforcement Learning 4 upvotes, #30 of 2026-02-13
- Multimodal Fact-Level Attribution for Verifiable Reasoning 4 upvotes, #30 of 2026-02-13
- Pretraining A Large Language Model using Distributed GPUs: A Memory-Efficient Decentralized Paradigm 4 upvotes, #30 of 2026-02-13
- P-GenRM: Personalized Generative Reward Model with Test-time User-based Scaling 4 upvotes, #30 of 2026-02-13
- Sci-CoE: Co-evolving Scientific Reasoning LLMs via Geometric Consensus with Sparse Supervision 4 upvotes, #30 of 2026-02-13
- Budget-Constrained Agentic Large Language Models: Intention-Based Planning for Costly Tool Use 3 upvotes, #35 of 2026-02-13
- ExStrucTiny: A Benchmark for Schema-Variable Structured Information Extraction from Document Images 3 upvotes, #35 of 2026-02-13
- Stemphonic: All-at-once Flexible Multi-stem Music Generation 2 upvotes, #37 of 2026-02-13
- Neural Additive Experts: Context-Gated Experts for Controllable Model Additivity 2 upvotes, #37 of 2026-02-13
- ABot-N0: Technical Report on the VLA Foundation Model for Versatile Embodied Navigation 2 upvotes, #37 of 2026-02-13
- ScalSelect: Scalable Training-Free Multimodal Data Selection for Efficient Visual Instruction Tuning 2 upvotes, #37 of 2026-02-13
- Detecting RLVR Training Data via Structural Convergence of Reasoning 2 upvotes, #37 of 2026-02-13
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.