Daily Papers of 2025-04-01
- MoCha: Towards Movie-Grade Talking Character Synthesis 103 upvotes, #1 of 2025-04-01
- TextCrafter: Accurately Rendering Multiple Texts in Complex Visual Scenes 87 upvotes, #2 of 2025-04-01
- Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model 59 upvotes, #3 of 2025-04-01
- What, How, Where, and How Well? A Survey on Test-Time Scaling in Large Language Models 49 upvotes, #4 of 2025-04-01
- Efficient Inference for Large Reasoning Models: A Survey 45 upvotes, #5 of 2025-04-01
- Unicorn: Text-Only Data Synthesis for Vision Language Model Training 37 upvotes, #6 of 2025-04-01
- TokenHSI: Unified Synthesis of Physical Human-Scene Interactions through Task Tokenization 33 upvotes, #7 of 2025-04-01
- RIG: Synergizing Reasoning and Imagination in End-to-End Generalist Policy 29 upvotes, #8 of 2025-04-01
- SketchVideo: Sketch-based Video Generation and Editing 21 upvotes, #9 of 2025-04-01
- Effectively Controlling Reasoning Models through Thinking Intervention 18 upvotes, #10 of 2025-04-01
- Expanding RL with Verifiable Rewards Across Diverse Domains 17 upvotes, #11 of 2025-04-01
- Query and Conquer: Execution-Guided SQL Generation 17 upvotes, #11 of 2025-04-01
- Progressive Rendering Distillation: Adapting Stable Diffusion for Instant Text-to-Mesh Generation without 3D Data 15 upvotes, #13 of 2025-04-01
- ActionStudio: A Lightweight Framework for Data and Training of Large Action Models 12 upvotes, #14 of 2025-04-01
- TeleAntiFraud-28k: A Audio-Text Slow-Thinking Dataset for Telecom Fraud Detection 11 upvotes, #15 of 2025-04-01
- Classical Planning with LLM-Generated Heuristics: Challenging the State of the Art with Python Code 10 upvotes, #16 of 2025-04-01
- AvatarArtist: Open-Domain 4D Avatarization 8 upvotes, #17 of 2025-04-01
- Easi3R: Estimating Disentangled Motion from DUSt3R Without Training 7 upvotes, #18 of 2025-04-01
- UPME: An Unsupervised Peer Review Framework for Multimodal Large Language Model Evaluation 6 upvotes, #19 of 2025-04-01
- MeshCraft: Exploring Efficient and Controllable Mesh Generation with Flow-based DiTs 6 upvotes, #19 of 2025-04-01
- DSO: Aligning 3D Generators with Simulation Feedback for Physical Soundness 5 upvotes, #21 of 2025-04-01
- Decoupling Angles and Strength in Low-rank Adaptation 4 upvotes, #22 of 2025-04-01
- PAVE: Patching and Adapting Video Large Language Models 4 upvotes, #22 of 2025-04-01
- Bridging Evolutionary Multiobjective Optimization and GPU Acceleration via Tensorization 4 upvotes, #22 of 2025-04-01
- KOFFVQA: An Objectively Evaluated Free-form VQA Benchmark for Large Vision-Language Models in the Korean Language 4 upvotes, #22 of 2025-04-01
- Entropy-Based Adaptive Weighting for Self-Training 4 upvotes, #22 of 2025-04-01
- Understanding Co-speech Gestures in-the-wild 1 upvotes, #27 of 2025-04-01
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.