Daily Papers of 2026-03-04
- Utonia: Toward One Encoder for All Point Clouds 161 upvotes, #1 of 2026-03-04
- Beyond Language Modeling: An Exploration of Multimodal Pretraining 85 upvotes, #2 of 2026-03-04
- UniG2U-Bench: Do Unified Models Advance Multimodal Understanding? 81 upvotes, #3 of 2026-03-04
- BeyondSWE: Can Current Code Agent Survive Beyond Single-Repo Bug Fixing? 53 upvotes, #4 of 2026-03-04
- Qwen3-Coder-Next Technical Report 44 upvotes, #5 of 2026-03-04
- Beyond Length Scaling: Synergizing Breadth and Depth for Generative Reward Models 33 upvotes, #6 of 2026-03-04
- Kling-MotionControl Technical Report 26 upvotes, #7 of 2026-03-04
- Kiwi-Edit: Versatile Video Editing via Instruction and Reference Guidance 22 upvotes, #8 of 2026-03-04
- How Controllable Are Large Language Models? A Unified Evaluation across Behavioral Granularities 22 upvotes, #8 of 2026-03-04
- PRISM: Pushing the Frontier of Deep Think via Process Reward Model-Guided Inference 18 upvotes, #10 of 2026-03-04
- Next Embedding Prediction Makes World Models Stronger 17 upvotes, #11 of 2026-03-04
- Code2Math: Can Your Code Agent Effectively Evolve Math Problems Through Exploration? 17 upvotes, #11 of 2026-03-04
- Spilled Energy in Large Language Models 11 upvotes, #13 of 2026-03-04
- Humans and LLMs Diverge on Probabilistic Inferences 11 upvotes, #13 of 2026-03-04
- Surgical Post-Training: Cutting Errors, Keeping Knowledge 11 upvotes, #13 of 2026-03-04
- Track4World: Feedforward World-centric Dense 3D Tracking of All Pixels 11 upvotes, #13 of 2026-03-04
- Learning When to Act or Refuse: Guarding Agentic Reasoning Models for Safe Multi-Step Tool Use 11 upvotes, #13 of 2026-03-04
- InfoPO: Information-Driven Policy Optimization for User-Centric Agents 9 upvotes, #18 of 2026-03-04
- BBQ-to-Image: Numeric Bounding Box and Qolor Control in Large-Scale Text-to-Image Models 8 upvotes, #19 of 2026-03-04
- NOVA: Sparse Control, Dense Synthesis for Pair-Free Video Editing 7 upvotes, #20 of 2026-03-04
- Towards Simulating Social Media Users with LLMs: Evaluating the Operational Validity of Conditioned Comment Prediction 6 upvotes, #21 of 2026-03-04
- SciDER: Scientific Data-centric End-to-end Researcher 6 upvotes, #21 of 2026-03-04
- Chain of World: World Model Thinking in Latent Motion 6 upvotes, #21 of 2026-03-04
- CFG-Ctrl: Control-Based Classifier-Free Diffusion Guidance 6 upvotes, #21 of 2026-03-04
- DREAM: Where Visual Understanding Meets Text-to-Image Generation 4 upvotes, #25 of 2026-03-04
- QEDBENCH: Quantifying the Alignment Gap in Automated Evaluation of University-Level Mathematical Proofs 3 upvotes, #26 of 2026-03-04
- Transformers converge to invariant algorithmic cores 3 upvotes, #26 of 2026-03-04
- ParEVO: Synthesizing Code for Irregular Data: High-Performance Parallelism through Agentic Evolution 3 upvotes, #26 of 2026-03-04
- AgentConductor: Topology Evolution for Multi-Agent Competition-Level Code Generation 2 upvotes, #29 of 2026-03-04
- LFPO: Likelihood-Free Policy Optimization for Masked Diffusion Models 2 upvotes, #29 of 2026-03-04
- DynaMoE: Dynamic Token-Level Expert Activation with Layer-Wise Adaptive Capacity for Mixture-of-Experts Neural Networks 2 upvotes, #29 of 2026-03-04
- APRES: An Agentic Paper Revision and Evaluation System 2 upvotes, #29 of 2026-03-04
- Easy to Learn, Yet Hard to Forget: Towards Robust Unlearning Under Bias 1 upvotes, #33 of 2026-03-04
- SGDC: Structurally-Guided Dynamic Convolution for Medical Image Segmentation 1 upvotes, #33 of 2026-03-04
- GroupGPT: A Token-efficient and Privacy-preserving Agentic Framework for Multi-User Chat Assistant 1 upvotes, #33 of 2026-03-04
- Words & Weights: Streamlining Multi-Turn Interactions via Co-Adaptation 1 upvotes, #33 of 2026-03-04
- Token Reduction via Local and Global Contexts Optimization for Efficient Video Large Language Models 1 upvotes, #33 of 2026-03-04
- Whisper-RIR-Mega: A Paired Clean-Reverberant Speech Benchmark for ASR Robustness to Room Acoustics 1 upvotes, #33 of 2026-03-04
- Fast Matrix Multiplication in Small Formats: Discovering New Schemes with an Open-Source Flip Graph Framework 1 upvotes, #33 of 2026-03-04
- HateMirage: An Explainable Multi-Dimensional Dataset for Decoding Faux Hate and Subtle Online Abuse 1 upvotes, #33 of 2026-03-04
- Conditioned Activation Transport for T2I Safety Steering 1 upvotes, #33 of 2026-03-04
- Multi-Domain Riemannian Graph Gluing for Building Graph Foundation Models 0 upvotes, #42 of 2026-03-04
- Transform-Invariant Generative Ray Path Sampling for Efficient Radio Propagation Modeling 0 upvotes, #42 of 2026-03-04
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.