Qwen

Qwen on Hugging Face Daily Papers: 55 papers, 15 in the top 3 of their day, 8 paper of the day.

  1. QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon Agents 43 upvotes, #15 of 2026-09-29
  2. Verifiable Hidden Dynamics Play: Generating Agentic RL Environments from Solved Mechanisms 22 upvotes, #10 of 2026-09-24
  3. OmniVChat: Synthesizing, Benchmarking, and Training for Native Audio-Visual Dialogue 149 upvotes, #2 of 2026-09-21
  4. RecreationWorld: Scalable and Verifiable Environments for Hybrid Computer-Use Agents 78 upvotes, #5 of 2026-09-21
  5. Terminal-Universe: Turning Agent Trajectories into Scalable Terminal Environments 291 upvotes, #2 of 2026-09-04
  6. Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving 378 upvotes, #2 of 2026-09-02
  7. E-Commerce Bench: Evaluating LLM Agents on Long-Horizon Autonomous Business Operation 17 upvotes, #14 of 2026-09-02
  8. On the Design of Qwen3.8-Next Architecture: Evaluation, Efficiency, and Training Stability 54 upvotes, #5 of 2026-09-01
  9. Are the Financial Reasoning from LLMs Credible? A Real World Test over Long-Horizon Statements 14 upvotes, #20 of 2026-08-05
  10. Qwen-Music Technical Report 27 upvotes, #11 of 2026-07-20
  11. Qwen-Image-2.0-RL Technical Report 48 upvotes, #3 of 2026-06-29
  12. Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models 28 upvotes, #5 of 2026-06-29
  13. Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System 26 upvotes, #6 of 2026-06-29
  14. The Verification Horizon: No Silver Bullet for Coding Agent Rewards 47 upvotes, #5 of 2026-06-26
  15. Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation 49 upvotes, #4 of 2026-06-26
  16. Qwen-AgentWorld: Language World Models for General Agents 144 upvotes, #1 of 2026-06-24
  17. Deeper is Not Always Better: Mitigating the Alignment Tax via Confident Layer Decoding 24 upvotes, #11 of 2026-06-23
  18. Native Active Perception as Reasoning for Omni-Modal Understanding 17 upvotes, #8 of 2026-06-18
  19. Unified Multimodal Autoregressive Modeling with Shared Context-Visual Tokenizer is Key to Unification 14 upvotes, #16 of 2026-06-17
  20. Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation 29 upvotes, #8 of 2026-06-16
  21. Breaking Entropy Bounds: Accelerating RL Training via MTP with Rejection Sampling 21 upvotes, #13 of 2026-06-11
  22. Verifiable Environments Are LEGO Bricks: Recursive Composition for Reasoning Generalization 7 upvotes, #24 of 2026-06-11
  23. Qwen-Image-Flash: Beyond Objective Design 35 upvotes, #5 of 2026-06-04
  24. Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments 138 upvotes, #2 of 2026-05-29
  25. CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents 31 upvotes, #11 of 2026-05-26
  26. Qwen-Image-VAE-2.0 Technical Report 58 upvotes, #6 of 2026-05-14
  27. Qwen-Image-2.0 Technical Report 106 upvotes, #1 of 2026-05-12
  28. OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models 62 upvotes, #3 of 2026-04-16
  29. FIPO: Eliciting Deep Reasoning with Future-KL Influenced Policy Optimization 325 upvotes, #2 of 2026-04-01
  30. RealChart2Code: Advancing Chart-to-Code Generation with Real Data and Multi-Task Evaluation 28 upvotes, #7 of 2026-03-30
  31. Sparse but Critical: A Token-Level Analysis of Distributional Shifts in RLVR Fine-Tuning of LLMs 7 upvotes, #18 of 2026-03-25
  32. On the Direction of RLVR Updates for LLM Reasoning: Identification and Exploitation 27 upvotes, #12 of 2026-03-24
  33. HopChain: Multi-Hop Data Synthesis for Generalizable Vision-Language Reasoning 106 upvotes, #1 of 2026-03-23
  34. CodePercept: Code-Grounded Visual STEM Perception for MLLMs 13 upvotes, #11 of 2026-03-12
  35. From Narrow to Panoramic Vision: Attention-Guided Cold-Start Reshapes Multimodal Reasoning 10 upvotes, #20 of 2026-03-10
  36. Qwen3-Coder-Next Technical Report 44 upvotes, #5 of 2026-03-04
  37. WebWorld: A Large-Scale World Model for Web Agent Training 7 upvotes, #16 of 2026-02-17
  38. OPUS: Towards Efficient and Principled Data Selection in Large Language Model Pre-training in Every Iteration 315 upvotes, #1 of 2026-02-11
  39. Outcome Accuracy is Not Enough: Aligning the Reasoning Process of Reward Models 11 upvotes, #19 of 2026-02-09
  40. Canzona: A Unified, Asynchronous, and Load-Balanced Framework for Distributed Matrix-based Optimizers 18 upvotes, #11 of 2026-02-09
  41. SWE-Universe: Scale Real-World Verifiable Environments to Millions 59 upvotes, #7 of 2026-02-03
  42. Qwen3-ASR Technical Report 33 upvotes, #8 of 2026-01-30
  43. DeepPlanning: Benchmarking Long-Horizon Agentic Planning with Verifiable Constraints 24 upvotes, #8 of 2026-01-27
  44. Qwen3-TTS Technical Report 54 upvotes, #5 of 2026-01-23
  45. MegaFlow: Large-Scale Distributed Orchestration System for the Agentic Era 19 upvotes, #11 of 2026-01-13
  46. Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking 45 upvotes, #5 of 2026-01-12
  47. Qwen3-VL Technical Report 120 upvotes, #1 of 2025-12-04
  48. Stabilizing Reinforcement Learning with LLMs: Formulation and Practices 83 upvotes, #4 of 2025-12-02
  49. Soft Adaptive Policy Optimization 33 upvotes, #6 of 2025-11-26
  50. Revisiting Multimodal Positional Encoding in Vision-Language Models 19 upvotes, #9 of 2025-11-03
  51. Qwen3Guard Technical Report 12 upvotes, #20 of 2025-10-17
  52. Scaling Generalist Data-Analytic Agents 16 upvotes, #25 of 2025-09-30
  53. Qwen3-Omni Technical Report 121 upvotes, #1 of 2025-09-23
  54. Group Sequence Policy Optimization 257 upvotes, #1 of 2025-07-25
  55. Qwen3 Technical Report 152 upvotes, #1 of 2025-05-19

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.