KABI
KABI on Hugging Face Daily Papers: 34 papers, 16 in the top 3 of their day, 2,326 upvotes.
- Atria Dawn: The Dawn of Agentic Superintelligence 421 upvotes, #2 of 2026-09-15
- Toward Generalist Autonomous Research via Hypothesis-Tree Refinement 112 upvotes, #2 of 2026-06-11
- Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence 80 upvotes, #3 of 2026-04-21
- OmniGAIA: Towards Native Omni-Modal AI Agents 51 upvotes, #4 of 2026-02-27
- ET-Agent: Incentivizing Effective Tool-Integrated Reasoning Agent via Behavior Calibration 15 upvotes, #15 of 2026-01-13
- EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis 35 upvotes, #7 of 2026-01-12
- SmartSearch: Process Reward-Guided Query Refinement for Search Agents 8 upvotes, #17 of 2026-01-12
- V-Thinker: Interactive Thinking with Images 93 upvotes, #2 of 2025-11-07
- ToolScope: An Agentic Framework for Vision-Guided and Long-Horizon Tool Use 22 upvotes, #12 of 2025-11-04
- DeepAgent: A General Reasoning Agent with Scalable Toolsets 92 upvotes, #1 of 2025-10-27
- Towards Mixed-Modal Retrieval for Universal Retrieval-Augmented Generation 32 upvotes, #7 of 2025-10-21
- Agentic Entropy-Balanced Policy Optimization 95 upvotes, #2 of 2025-10-17
- Toward Effective Tool-Integrated Reasoning via Self-Evolved Preference Learning 13 upvotes, #30 of 2025-09-30
- We-Math 2.0: A Versatile MathBook System for Incentivizing Visual Mathematical Reasoning 141 upvotes, #1 of 2025-08-15
- Klear-Reasoner: Advancing Reasoning Capability via Gradient-Preserving Clipping Policy Optimization 37 upvotes, #7 of 2025-08-12
- Agentic Reinforced Policy Optimization 127 upvotes, #1 of 2025-07-29
- Decoupled Planning and Execution: A Hierarchical Reasoning Framework for Deep Search 22 upvotes, #9 of 2025-07-04
- Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning 53 upvotes, #3 of 2025-05-23
- WebThinker: Empowering Large Reasoning Models with Deep Research Capability 41 upvotes, #2 of 2025-05-01
- Search-o1: Agentic Search-Enhanced Large Reasoning Models 75 upvotes, #4 of 2025-01-09
- Progressive Multimodal Reasoning via Active Retrieval 67 upvotes, #2 of 2024-12-20
- Multi-Dimensional Insights: Benchmarking Real-World Personalization in Large Multimodal Models 41 upvotes, #2 of 2024-12-18
- Smaller Language Models Are Better Instruction Evolvers 24 upvotes, #6 of 2024-12-17
- CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation 52 upvotes, #1 of 2024-10-31
- Toward General Instruction-Following Alignment for Retrieval-Augmented Generation 43 upvotes, #4 of 2024-10-15
- MSI-Agent: Incorporating Multi-Scale Insight into Embodied Agents for Superior Planning and Decision-Making 7 upvotes, #10 of 2024-09-30
- How Do Your Code LLMs Perform? Empowering Code Instruction Tuning with High-Quality Data 29 upvotes, #1 of 2024-09-09
- Qwen2 Technical Report 142 upvotes, #1 of 2024-07-16
- DotaMath: Decomposition of Thought with Code Assistance and Self-correction for Mathematical Reasoning 13 upvotes, #8 of 2024-07-08
- We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning? 69 upvotes, #1 of 2024-07-02
- Understand What LLM Needs: Dual Preference Alignment for Retrieval-Augmented Generation 5 upvotes, #15 of 2024-06-28
- Self-play with Execution Feedback: Improving Instruction-following Capabilities of Large Language Models 14 upvotes, #13 of 2024-06-21
- CS-Bench: A Comprehensive Benchmark for Large Language Models towards Computer Science Mastery 14 upvotes, #12 of 2024-06-14
- Scaling Relationship on Learning Mathematical Reasoning with Large Language Models 23 upvotes, #4 of 2023-08-04
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.