Qwen
Qwen on Hugging Face Daily Papers: 55 papers, 15 in the top 3 of their day, 8 paper of the day.
- QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon Agents 43 upvotes, #15 of 2026-09-29
- Verifiable Hidden Dynamics Play: Generating Agentic RL Environments from Solved Mechanisms 22 upvotes, #10 of 2026-09-24
- OmniVChat: Synthesizing, Benchmarking, and Training for Native Audio-Visual Dialogue 149 upvotes, #2 of 2026-09-21
- RecreationWorld: Scalable and Verifiable Environments for Hybrid Computer-Use Agents 78 upvotes, #5 of 2026-09-21
- Terminal-Universe: Turning Agent Trajectories into Scalable Terminal Environments 291 upvotes, #2 of 2026-09-04
- Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving 378 upvotes, #2 of 2026-09-02
- E-Commerce Bench: Evaluating LLM Agents on Long-Horizon Autonomous Business Operation 17 upvotes, #14 of 2026-09-02
- On the Design of Qwen3.8-Next Architecture: Evaluation, Efficiency, and Training Stability 54 upvotes, #5 of 2026-09-01
- Are the Financial Reasoning from LLMs Credible? A Real World Test over Long-Horizon Statements 14 upvotes, #20 of 2026-08-05
- Qwen-Music Technical Report 27 upvotes, #11 of 2026-07-20
- Qwen-Image-2.0-RL Technical Report 48 upvotes, #3 of 2026-06-29
- Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models 28 upvotes, #5 of 2026-06-29
- Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System 26 upvotes, #6 of 2026-06-29
- The Verification Horizon: No Silver Bullet for Coding Agent Rewards 47 upvotes, #5 of 2026-06-26
- Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation 49 upvotes, #4 of 2026-06-26
- Qwen-AgentWorld: Language World Models for General Agents 144 upvotes, #1 of 2026-06-24
- Deeper is Not Always Better: Mitigating the Alignment Tax via Confident Layer Decoding 24 upvotes, #11 of 2026-06-23
- Native Active Perception as Reasoning for Omni-Modal Understanding 17 upvotes, #8 of 2026-06-18
- Unified Multimodal Autoregressive Modeling with Shared Context-Visual Tokenizer is Key to Unification 14 upvotes, #16 of 2026-06-17
- Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation 29 upvotes, #8 of 2026-06-16
- Breaking Entropy Bounds: Accelerating RL Training via MTP with Rejection Sampling 21 upvotes, #13 of 2026-06-11
- Verifiable Environments Are LEGO Bricks: Recursive Composition for Reasoning Generalization 7 upvotes, #24 of 2026-06-11
- Qwen-Image-Flash: Beyond Objective Design 35 upvotes, #5 of 2026-06-04
- Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments 138 upvotes, #2 of 2026-05-29
- CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents 31 upvotes, #11 of 2026-05-26
- Qwen-Image-VAE-2.0 Technical Report 58 upvotes, #6 of 2026-05-14
- Qwen-Image-2.0 Technical Report 106 upvotes, #1 of 2026-05-12
- OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models 62 upvotes, #3 of 2026-04-16
- FIPO: Eliciting Deep Reasoning with Future-KL Influenced Policy Optimization 325 upvotes, #2 of 2026-04-01
- RealChart2Code: Advancing Chart-to-Code Generation with Real Data and Multi-Task Evaluation 28 upvotes, #7 of 2026-03-30
- Sparse but Critical: A Token-Level Analysis of Distributional Shifts in RLVR Fine-Tuning of LLMs 7 upvotes, #18 of 2026-03-25
- On the Direction of RLVR Updates for LLM Reasoning: Identification and Exploitation 27 upvotes, #12 of 2026-03-24
- HopChain: Multi-Hop Data Synthesis for Generalizable Vision-Language Reasoning 106 upvotes, #1 of 2026-03-23
- CodePercept: Code-Grounded Visual STEM Perception for MLLMs 13 upvotes, #11 of 2026-03-12
- From Narrow to Panoramic Vision: Attention-Guided Cold-Start Reshapes Multimodal Reasoning 10 upvotes, #20 of 2026-03-10
- Qwen3-Coder-Next Technical Report 44 upvotes, #5 of 2026-03-04
- WebWorld: A Large-Scale World Model for Web Agent Training 7 upvotes, #16 of 2026-02-17
- OPUS: Towards Efficient and Principled Data Selection in Large Language Model Pre-training in Every Iteration 315 upvotes, #1 of 2026-02-11
- Outcome Accuracy is Not Enough: Aligning the Reasoning Process of Reward Models 11 upvotes, #19 of 2026-02-09
- Canzona: A Unified, Asynchronous, and Load-Balanced Framework for Distributed Matrix-based Optimizers 18 upvotes, #11 of 2026-02-09
- SWE-Universe: Scale Real-World Verifiable Environments to Millions 59 upvotes, #7 of 2026-02-03
- Qwen3-ASR Technical Report 33 upvotes, #8 of 2026-01-30
- DeepPlanning: Benchmarking Long-Horizon Agentic Planning with Verifiable Constraints 24 upvotes, #8 of 2026-01-27
- Qwen3-TTS Technical Report 54 upvotes, #5 of 2026-01-23
- MegaFlow: Large-Scale Distributed Orchestration System for the Agentic Era 19 upvotes, #11 of 2026-01-13
- Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking 45 upvotes, #5 of 2026-01-12
- Qwen3-VL Technical Report 120 upvotes, #1 of 2025-12-04
- Stabilizing Reinforcement Learning with LLMs: Formulation and Practices 83 upvotes, #4 of 2025-12-02
- Soft Adaptive Policy Optimization 33 upvotes, #6 of 2025-11-26
- Revisiting Multimodal Positional Encoding in Vision-Language Models 19 upvotes, #9 of 2025-11-03
- Qwen3Guard Technical Report 12 upvotes, #20 of 2025-10-17
- Scaling Generalist Data-Analytic Agents 16 upvotes, #25 of 2025-09-30
- Qwen3-Omni Technical Report 121 upvotes, #1 of 2025-09-23
- Group Sequence Policy Optimization 257 upvotes, #1 of 2025-07-25
- Qwen3 Technical Report 152 upvotes, #1 of 2025-05-19
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.