KABI

KABI on Hugging Face Daily Papers: 34 papers, 16 in the top 3 of their day, 2,326 upvotes.

  1. Atria Dawn: The Dawn of Agentic Superintelligence 421 upvotes, #2 of 2026-09-15
  2. Toward Generalist Autonomous Research via Hypothesis-Tree Refinement 112 upvotes, #2 of 2026-06-11
  3. Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence 80 upvotes, #3 of 2026-04-21
  4. OmniGAIA: Towards Native Omni-Modal AI Agents 51 upvotes, #4 of 2026-02-27
  5. ET-Agent: Incentivizing Effective Tool-Integrated Reasoning Agent via Behavior Calibration 15 upvotes, #15 of 2026-01-13
  6. EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis 35 upvotes, #7 of 2026-01-12
  7. SmartSearch: Process Reward-Guided Query Refinement for Search Agents 8 upvotes, #17 of 2026-01-12
  8. V-Thinker: Interactive Thinking with Images 93 upvotes, #2 of 2025-11-07
  9. ToolScope: An Agentic Framework for Vision-Guided and Long-Horizon Tool Use 22 upvotes, #12 of 2025-11-04
  10. DeepAgent: A General Reasoning Agent with Scalable Toolsets 92 upvotes, #1 of 2025-10-27
  11. Towards Mixed-Modal Retrieval for Universal Retrieval-Augmented Generation 32 upvotes, #7 of 2025-10-21
  12. Agentic Entropy-Balanced Policy Optimization 95 upvotes, #2 of 2025-10-17
  13. Toward Effective Tool-Integrated Reasoning via Self-Evolved Preference Learning 13 upvotes, #30 of 2025-09-30
  14. We-Math 2.0: A Versatile MathBook System for Incentivizing Visual Mathematical Reasoning 141 upvotes, #1 of 2025-08-15
  15. Klear-Reasoner: Advancing Reasoning Capability via Gradient-Preserving Clipping Policy Optimization 37 upvotes, #7 of 2025-08-12
  16. Agentic Reinforced Policy Optimization 127 upvotes, #1 of 2025-07-29
  17. Decoupled Planning and Execution: A Hierarchical Reasoning Framework for Deep Search 22 upvotes, #9 of 2025-07-04
  18. Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning 53 upvotes, #3 of 2025-05-23
  19. WebThinker: Empowering Large Reasoning Models with Deep Research Capability 41 upvotes, #2 of 2025-05-01
  20. Search-o1: Agentic Search-Enhanced Large Reasoning Models 75 upvotes, #4 of 2025-01-09
  21. Progressive Multimodal Reasoning via Active Retrieval 67 upvotes, #2 of 2024-12-20
  22. Multi-Dimensional Insights: Benchmarking Real-World Personalization in Large Multimodal Models 41 upvotes, #2 of 2024-12-18
  23. Smaller Language Models Are Better Instruction Evolvers 24 upvotes, #6 of 2024-12-17
  24. CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation 52 upvotes, #1 of 2024-10-31
  25. Toward General Instruction-Following Alignment for Retrieval-Augmented Generation 43 upvotes, #4 of 2024-10-15
  26. MSI-Agent: Incorporating Multi-Scale Insight into Embodied Agents for Superior Planning and Decision-Making 7 upvotes, #10 of 2024-09-30
  27. How Do Your Code LLMs Perform? Empowering Code Instruction Tuning with High-Quality Data 29 upvotes, #1 of 2024-09-09
  28. Qwen2 Technical Report 142 upvotes, #1 of 2024-07-16
  29. DotaMath: Decomposition of Thought with Code Assistance and Self-correction for Mathematical Reasoning 13 upvotes, #8 of 2024-07-08
  30. We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning? 69 upvotes, #1 of 2024-07-02
  31. Understand What LLM Needs: Dual Preference Alignment for Retrieval-Augmented Generation 5 upvotes, #15 of 2024-06-28
  32. Self-play with Execution Feedback: Improving Instruction-following Capabilities of Large Language Models 14 upvotes, #13 of 2024-06-21
  33. CS-Bench: A Comprehensive Benchmark for Large Language Models towards Computer Science Mastery 14 upvotes, #12 of 2024-06-14
  34. Scaling Relationship on Learning Mathematical Reasoning with Large Language Models 23 upvotes, #4 of 2023-08-04

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.