Daily Papers of 2025-08-13
- WebWatcher: Breaking New Frontier of Vision-Language Deep Research Agent 114 upvotes, #1 of 2025-08-13
- Matrix-3D: Omnidirectional Explorable 3D World Generation 67 upvotes, #2 of 2025-08-13
- Beyond Ten Turns: Unlocking Long-Horizon Agentic Search with Large-Scale Asynchronous RL 45 upvotes, #3 of 2025-08-13
- Complex Logical Instruction Generation 38 upvotes, #4 of 2025-08-13
- CharacterShot: Controllable and Consistent 4D Character Animation 37 upvotes, #5 of 2025-08-13
- Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models 34 upvotes, #6 of 2025-08-13
- VertexRegen: Mesh Generation with Continuous Level of Detail 33 upvotes, #7 of 2025-08-13
- HierSearch: A Hierarchical Enterprise Deep Search Framework Integrating Local and Web Searches 28 upvotes, #8 of 2025-08-13
- OpenCUA: Open Foundations for Computer-Use Agents 25 upvotes, #9 of 2025-08-13
- StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation 24 upvotes, #10 of 2025-08-13
- Train Long, Think Short: Curriculum Learning for Efficient Reasoning 22 upvotes, #11 of 2025-08-13
- Test-Time Reinforcement Learning for GUI Grounding via Region Consistency 20 upvotes, #12 of 2025-08-13
- UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation 16 upvotes, #13 of 2025-08-13
- Aryabhata: An exam-focused language model for JEE Math 16 upvotes, #13 of 2025-08-13
- Feedback-Driven Tool-Use Improvements in Large Language Models via Automated Build Environments 16 upvotes, #13 of 2025-08-13
- Cut2Next: Generating Next Shot via In-Context Tuning 12 upvotes, #16 of 2025-08-13
- Democratizing Diplomacy: A Harness for Evaluating Any Large Language Model on Full-Press Diplomacy 10 upvotes, #17 of 2025-08-13
- Towards Affordance-Aware Robotic Dexterous Grasping with Human-like Priors 10 upvotes, #17 of 2025-08-13
- ASTRA: Autonomous Spatial-Temporal Red-teaming for AI Software Assistants 9 upvotes, #19 of 2025-08-13
- Adversarial Video Promotion Against Text-to-Video Retrieval 9 upvotes, #19 of 2025-08-13
- AimBot: A Simple Auxiliary Visual Cue to Enhance Spatial Awareness of Visuomotor Policies 9 upvotes, #19 of 2025-08-13
- DeCRED: Decoder-Centric Regularization for Encoder-Decoder Based Speech Recognition 8 upvotes, #22 of 2025-08-13
- AutoCodeBench: Large Language Models are Automatic Code Benchmark Generators 7 upvotes, #23 of 2025-08-13
- Technical Report: Full-Stack Fine-Tuning for the Q Programming Language 5 upvotes, #24 of 2025-08-13
- GeRe: Towards Efficient Anti-Forgetting in Continual Learning of LLM via General Samples Replay 4 upvotes, #25 of 2025-08-13
- Optimization-Free Style Transfer for 3D Gaussian Splats 4 upvotes, #25 of 2025-08-13
- TopXGen: Topic-Diverse Parallel Data Generation for Low-Resource Machine Translation 3 upvotes, #27 of 2025-08-13
- RedDino: A foundation model for red blood cell analysis 2 upvotes, #28 of 2025-08-13
- Bridging Theory and Practice in Quantum Game Theory: Optimized Implementation of the Battle of the Sexes with Error Mitigation on NISQ Hardware 2 upvotes, #28 of 2025-08-13
- NVSpeech: An Integrated and Scalable Pipeline for Human-Like Speech Modeling with Paralinguistic Vocalizations 1 upvotes, #30 of 2025-08-13
- WGAST: Weakly-Supervised Generative Network for Daily 10 m Land Surface Temperature Estimation via Spatio-Temporal Fusion 1 upvotes, #30 of 2025-08-13
- Putnam-AXIOM: A Functional and Static Benchmark 1 upvotes, #30 of 2025-08-13
- BiasGym: Fantastic Biases and How to Find (and Remove) Them 1 upvotes, #30 of 2025-08-13
- Improving Masked Style Transfer using Blended Partial Convolution 0 upvotes, #34 of 2025-08-13
- Text-conditioned State Space Model For Domain-generalized Change Detection Visual Question Answering 0 upvotes, #34 of 2025-08-13
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.