Jinyang Wu

Jinyang Wu on Hugging Face Daily Papers: 21 papers, 7 in the top 3 of their day, 950 upvotes.

  1. WideSWE: Can Coding Agents Coordinate Changes Across Repositories? 17 upvotes, #44 of 2026-09-29
  2. VideoGAIA: A Benchmark for General AI Assistants on Agentic Video Understanding 10 upvotes, #23 of 2026-08-18
  3. AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning 94 upvotes, #1 of 2026-08-07
  4. PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement Learning 38 upvotes, #7 of 2026-08-05
  5. From Proprietary to Open-Source: Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic Search 82 upvotes, #4 of 2026-07-28
  6. SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning 99 upvotes, #3 of 2026-07-17
  7. TACO: Tool-Augmented Credit Optimization for Agentic Tool Use 21 upvotes, #14 of 2026-06-30
  8. OPID: On-Policy Skill Distillation for Agentic Reinforcement Learning 54 upvotes, #3 of 2026-06-26
  9. Orchestra-o1: Omnimodal Agent Orchestration 45 upvotes, #7 of 2026-06-15
  10. Late-Layer Fusion is Enough: Dual-Path Vision Token Routing for Multimodal Large Language Models under Visual Saturation 3 upvotes, #36 of 2026-06-10
  11. Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles 20 upvotes, #19 of 2026-05-22
  12. Self-Distilled Agentic Reinforcement Learning 107 upvotes, #2 of 2026-05-15
  13. SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization 92 upvotes, #4 of 2026-04-03
  14. Exploring Knowledge Purification in Multi-Teacher Knowledge Distillation for LLMs 2 upvotes, #35 of 2026-02-09
  15. OdysseyArena: Benchmarking Large Language Models For Long-Horizon, Active and Inductive Interactions 57 upvotes, #4 of 2026-02-09
  16. Spark: Strategic Policy-Aware Exploration via Dynamic Branching for Long-Horizon Agentic Learning 22 upvotes, #6 of 2026-01-29
  17. Atlas: Orchestrating Heterogeneous Models and Tools for Multi-Domain Complex Reasoning 40 upvotes, #3 of 2026-01-08
  18. From Imitation to Discrimination: Toward A Generalized Curriculum Advantage Mechanism Enhancing Cross-Domain Reasoning Tasks 27 upvotes, #3 of 2025-12-08
  19. Thought-Augmented Policy Optimization: Bridging External Guidance and Internal Capabilities 14 upvotes, #19 of 2025-05-26
  20. Boosting Multimodal Reasoning with MCTS-Automated Structured Thinking 19 upvotes, #5 of 2025-02-06
  21. Beyond Examples: High-level Automated Reasoning Paradigm in In-Context Learning via MCTS 30 upvotes, #2 of 2024-12-02

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.