Daily Papers of 2025-04-30
- Reinforcement Learning for Reasoning in Large Language Models with One Training Example 88 upvotes, #1 of 2025-04-30
- The Leaderboard Illusion 66 upvotes, #2 of 2025-04-30
- UniversalRAG: Retrieval-Augmented Generation over Multiple Corpora with Diverse Modalities and Granularities 60 upvotes, #3 of 2025-04-30
- ReasonIR: Training Retrievers for Reasoning Tasks 50 upvotes, #4 of 2025-04-30
- Toward Evaluative Thinking: Meta Policy Optimization with Evolving Reward Models 34 upvotes, #5 of 2025-04-30
- TesserAct: Learning 4D Embodied World Models 19 upvotes, #6 of 2025-04-30
- In-Context Edit: Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer 16 upvotes, #7 of 2025-04-30
- Certified Mitigation of Worst-Case LLM Copyright Infringement 12 upvotes, #8 of 2025-04-30
- X-Fusion: Introducing New Modality to Frozen Large Language Models 11 upvotes, #9 of 2025-04-30
- YoChameleon: Personalized Vision and Language Generation 11 upvotes, #9 of 2025-04-30
- RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning 8 upvotes, #11 of 2025-04-30
- ISDrama: Immersive Spatial Drama Generation through Multimodal Prompting 8 upvotes, #11 of 2025-04-30
- Learning Explainable Dense Reward Shapes via Bayesian Optimization 5 upvotes, #13 of 2025-04-30
- TreeHop: Generate and Filter Next Query Embeddings Efficiently for Multi-hop Question Answering 5 upvotes, #13 of 2025-04-30
- Disentangle Identity, Cooperate Emotion: Correlation-Aware Emotional Talking Portrait Generation 4 upvotes, #15 of 2025-04-30
- LawFlow : Collecting and Simulating Lawyers' Thought Processes 4 upvotes, #15 of 2025-04-30
- Chain-of-Defensive-Thought: Structured Reasoning Elicits Robustness in Large Language Models against Reference Corruption 3 upvotes, #17 of 2025-04-30
- CaRL: Learning Scalable Planning Policies with Simple Rewards 2 upvotes, #18 of 2025-04-30
- A Review of 3D Object Detection with Vision-Language Models 2 upvotes, #18 of 2025-04-30
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.