TongyiLab

TongyiLab on Hugging Face Daily Papers: 31 papers, 7 in the top 3 of their day, 2 paper of the day.

  1. ReWorld: An Interactive World Model with Long-Horizon Memory 24 upvotes, #11 of 2026-08-25
  2. Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents 302 upvotes, #1 of 2026-07-31
  3. Focusing on What Matters: Saliency-Harnessing Accurate Routing for Diffusion MoE 6 upvotes, #34 of 2026-06-30
  4. Connect the Dots: Training LLMs for Long-Lifecycle Agents with Cross-Domain Generalization Via Reinforcement Learning 8 upvotes, #27 of 2026-06-23
  5. EvoTrainer: Co-Evolving LLM Policies and Training Harnesses for Autonomous Agentic Reinforcement Learning 11 upvotes, #22 of 2026-06-11
  6. ESPO: Early-Stopping Proximal Policy Optimization 19 upvotes, #17 of 2026-06-02
  7. See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding 33 upvotes, #7 of 2026-05-25
  8. MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation 14 upvotes, #15 of 2026-05-20
  9. ToolCUA: Towards Optimal GUI-Tool Path Orchestration for Computer Use Agents 27 upvotes, #14 of 2026-05-13
  10. TCOD: Exploring Temporal Curriculum in On-Policy Distillation for Multi-turn Autonomous Agents 7 upvotes, #11 of 2026-04-29
  11. DreamVideo-Omni: Omni-Motion Controlled Multi-Subject Video Customization with Latent Identity Reinforcement Learning 31 upvotes, #6 of 2026-03-13
  12. LaSER: Internalizing Explicit Reasoning into Latent Space for Dense Retrieval 5 upvotes, #27 of 2026-03-03
  13. Mobile-Agent-v3.5: Multi-platform Fundamental GUI Agents 46 upvotes, #2 of 2026-02-20
  14. DenseGRPO: From Sparse to Dense Reward for Flow Matching Model Alignment 15 upvotes, #14 of 2026-02-02
  15. ExpSeek: Self-Triggered Experience Seeking for Web Agents 16 upvotes, #12 of 2026-01-15
  16. Nested Browser-Use Learning for Agentic Information Seeking 17 upvotes, #13 of 2025-12-30
  17. Knot Forcing: Taming Autoregressive Video Diffusion Models for Real-time Infinite Interactive Portrait Animation 4 upvotes, #27 of 2025-12-30
  18. MAI-UI Technical Report: Real-World Centric Foundation GUI Agents 26 upvotes, #3 of 2025-12-29
  19. MobileWorld: Benchmarking Autonomous Mobile Agents in Agent-User Interactive, and MCP-Augmented Environments 11 upvotes, #15 of 2025-12-23
  20. QwenLong-L1.5: Post-Training Recipe for Long-Context Reasoning and Memory Management 98 upvotes, #2 of 2025-12-16
  21. EcomBench: Towards Holistic Evaluation of Foundation Agents in E-commerce 2 upvotes, #20 of 2025-12-10
  22. Wan-Move: Motion-controllable Video Generation via Latent Trajectory Guidance 123 upvotes, #1 of 2025-12-10
  23. AgentFrontier: Expanding the Capability Frontier of LLM Agents with ZPD-Guided Data Synthesis 22 upvotes, #9 of 2025-10-29
  24. Tongyi DeepResearch Technical Report 89 upvotes, #2 of 2025-10-29
  25. AgentFold: Long-Horizon Web Agents with Proactive Context Management 65 upvotes, #3 of 2025-10-29
  26. WebLeaper: Empowering Efficiency and Efficacy in WebAgent via Enabling Info-Rich Seeking 20 upvotes, #14 of 2025-10-29
  27. ParallelMuse: Agentic Parallel Thinking for Deep Information Seeking 20 upvotes, #14 of 2025-10-29
  28. Repurposing Synthetic Data for Fine-grained Search Agent Supervision 22 upvotes, #9 of 2025-10-29
  29. OSWorld-MCP: Benchmarking MCP Tool Invocation In Computer-Use Agents 22 upvotes, #9 of 2025-10-29
  30. UI-Ins: Enhancing GUI Grounding with Multi-Perspective Instruction-as-Reasoning 22 upvotes, #9 of 2025-10-27
  31. LOVE-R1: Advancing Long Video Understanding with an Adaptive Zoom-in Mechanism via Multi-Step Reasoning 5 upvotes, #54 of 2025-09-30

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.