TongyiLab
TongyiLab on Hugging Face Daily Papers: 31 papers, 7 in the top 3 of their day, 2 paper of the day.
- ReWorld: An Interactive World Model with Long-Horizon Memory 24 upvotes, #11 of 2026-08-25
- Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents 302 upvotes, #1 of 2026-07-31
- Focusing on What Matters: Saliency-Harnessing Accurate Routing for Diffusion MoE 6 upvotes, #34 of 2026-06-30
- Connect the Dots: Training LLMs for Long-Lifecycle Agents with Cross-Domain Generalization Via Reinforcement Learning 8 upvotes, #27 of 2026-06-23
- EvoTrainer: Co-Evolving LLM Policies and Training Harnesses for Autonomous Agentic Reinforcement Learning 11 upvotes, #22 of 2026-06-11
- ESPO: Early-Stopping Proximal Policy Optimization 19 upvotes, #17 of 2026-06-02
- See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding 33 upvotes, #7 of 2026-05-25
- MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation 14 upvotes, #15 of 2026-05-20
- ToolCUA: Towards Optimal GUI-Tool Path Orchestration for Computer Use Agents 27 upvotes, #14 of 2026-05-13
- TCOD: Exploring Temporal Curriculum in On-Policy Distillation for Multi-turn Autonomous Agents 7 upvotes, #11 of 2026-04-29
- DreamVideo-Omni: Omni-Motion Controlled Multi-Subject Video Customization with Latent Identity Reinforcement Learning 31 upvotes, #6 of 2026-03-13
- LaSER: Internalizing Explicit Reasoning into Latent Space for Dense Retrieval 5 upvotes, #27 of 2026-03-03
- Mobile-Agent-v3.5: Multi-platform Fundamental GUI Agents 46 upvotes, #2 of 2026-02-20
- DenseGRPO: From Sparse to Dense Reward for Flow Matching Model Alignment 15 upvotes, #14 of 2026-02-02
- ExpSeek: Self-Triggered Experience Seeking for Web Agents 16 upvotes, #12 of 2026-01-15
- Nested Browser-Use Learning for Agentic Information Seeking 17 upvotes, #13 of 2025-12-30
- Knot Forcing: Taming Autoregressive Video Diffusion Models for Real-time Infinite Interactive Portrait Animation 4 upvotes, #27 of 2025-12-30
- MAI-UI Technical Report: Real-World Centric Foundation GUI Agents 26 upvotes, #3 of 2025-12-29
- MobileWorld: Benchmarking Autonomous Mobile Agents in Agent-User Interactive, and MCP-Augmented Environments 11 upvotes, #15 of 2025-12-23
- QwenLong-L1.5: Post-Training Recipe for Long-Context Reasoning and Memory Management 98 upvotes, #2 of 2025-12-16
- EcomBench: Towards Holistic Evaluation of Foundation Agents in E-commerce 2 upvotes, #20 of 2025-12-10
- Wan-Move: Motion-controllable Video Generation via Latent Trajectory Guidance 123 upvotes, #1 of 2025-12-10
- AgentFrontier: Expanding the Capability Frontier of LLM Agents with ZPD-Guided Data Synthesis 22 upvotes, #9 of 2025-10-29
- Tongyi DeepResearch Technical Report 89 upvotes, #2 of 2025-10-29
- AgentFold: Long-Horizon Web Agents with Proactive Context Management 65 upvotes, #3 of 2025-10-29
- WebLeaper: Empowering Efficiency and Efficacy in WebAgent via Enabling Info-Rich Seeking 20 upvotes, #14 of 2025-10-29
- ParallelMuse: Agentic Parallel Thinking for Deep Information Seeking 20 upvotes, #14 of 2025-10-29
- Repurposing Synthetic Data for Fine-grained Search Agent Supervision 22 upvotes, #9 of 2025-10-29
- OSWorld-MCP: Benchmarking MCP Tool Invocation In Computer-Use Agents 22 upvotes, #9 of 2025-10-29
- UI-Ins: Enhancing GUI Grounding with Multi-Perspective Instruction-as-Reasoning 22 upvotes, #9 of 2025-10-27
- LOVE-R1: Advancing Long Video Understanding with an Adaptive Zoom-in Mechanism via Multi-Step Reasoning 5 upvotes, #54 of 2025-09-30
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.