Daily Papers of 2026-02-16
- Less is Enough: Synthesizing Diverse Data in Feature Space of LLMs 225 upvotes, #1 of 2026-02-16
- SQuTR: A Robustness Benchmark for Spoken Query to Text Retrieval under Acoustic Noise 142 upvotes, #2 of 2026-02-16
- MedXIAOHE: A Comprehensive Recipe for Building Medical MLLMs 59 upvotes, #3 of 2026-02-16
- Zooming without Zooming: Region-to-Image Distillation for Fine-Grained Multimodal Perception 58 upvotes, #4 of 2026-02-16
- OneVision-Encoder: Codec-Aligned Sparsity as a Foundational Principle for Multimodal Intelligence 46 upvotes, #5 of 2026-02-16
- CoPE-VideoLM: Codec Primitives For Efficient Video Language Models 29 upvotes, #6 of 2026-02-16
- SemanticMoments: Training-Free Motion Similarity via Third Moment Features 21 upvotes, #7 of 2026-02-16
- GeoAgent: Learning to Geolocate Everywhere with Reinforced Geographic Characteristics 20 upvotes, #8 of 2026-02-16
- What does RL improve for Visual Reasoning? A Frankenstein-Style Analysis 14 upvotes, #9 of 2026-02-16
- Intelligent AI Delegation 13 upvotes, #10 of 2026-02-16
- ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning 11 upvotes, #11 of 2026-02-16
- RLinf-Co: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models 11 upvotes, #11 of 2026-02-16
- BPDQ: Bit-Plane Decomposition Quantization on a Variable Grid for Large Language Models 9 upvotes, #13 of 2026-02-16
- Towards Universal Video MLLMs with Attribute-Structured and Quality-Verified Instructions 8 upvotes, #14 of 2026-02-16
- Xiaomi-Robotics-0: An Open-Sourced Vision-Language-Action Model with Real-Time Execution 6 upvotes, #15 of 2026-02-16
- DICE: Diffusion Large Language Models Excel at Generating CUDA Kernels 5 upvotes, #16 of 2026-02-16
- SciAgentGym: Benchmarking Multi-Step Scientific Tool-use in LLM Agents 5 upvotes, #16 of 2026-02-16
- Principled Synthetic Data Enables the First Scaling Laws for LLMs in Recommendation 4 upvotes, #18 of 2026-02-16
- Code2Worlds: Empowering Coding LLMs for 4D World Generation 4 upvotes, #18 of 2026-02-16
- FLAC: Maximum Entropy RL via Kinetic Energy Regularized Bridge Matching 4 upvotes, #18 of 2026-02-16
- Best of Both Worlds: Multimodal Reasoning and Generation via Unified Discrete Flow Matching 3 upvotes, #21 of 2026-02-16
- On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs 3 upvotes, #21 of 2026-02-16
- Light4D: Training-Free Extreme Viewpoint 4D Video Relighting 2 upvotes, #23 of 2026-02-16
- TADA! Tuning Audio Diffusion Models through Activation Steering 2 upvotes, #23 of 2026-02-16
- Favia: Forensic Agent for Vulnerability-fix Identification and Analysis 2 upvotes, #23 of 2026-02-16
- Self-EvolveRec: Self-Evolving Recommender Systems with LLM-based Directional Feedback 2 upvotes, #23 of 2026-02-16
- Quantized Evolution Strategies: High-precision Fine-tuning of Quantized LLMs at Low-precision Cost 1 upvotes, #27 of 2026-02-16
- GeneralVLA: Generalizable Vision-Language-Action Models with Knowledge-Guided Trajectory Planning 1 upvotes, #27 of 2026-02-16
- Steer2Edit: From Activation Steering to Component-Level Editing 1 upvotes, #27 of 2026-02-16
- scPilot: Large Language Model Reasoning Toward Automated Single-Cell Analysis and Discovery 1 upvotes, #27 of 2026-02-16
- Learning Image-based Tree Crown Segmentation from Enhanced Lidar-based Pseudo-labels 1 upvotes, #27 of 2026-02-16
- OpenLID-v3: Improving the Precision of Closely Related Language Identification -- An Experience Report 0 upvotes, #32 of 2026-02-16
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.