Zhixuan Liang

Zhixuan Liang on Hugging Face Daily Papers: 15 papers, 1 in the top 3 of their day, 394 upvotes.

  1. Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System 26 upvotes, #6 of 2026-06-29
  2. Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models 28 upvotes, #5 of 2026-06-29
  3. Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation 29 upvotes, #8 of 2026-06-16
  4. AHA-WAM:Asynchronous Horizon-Adaptive World-Action Modeling with Observation-Guided Context Routing 14 upvotes, #18 of 2026-06-09
  5. Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments 138 upvotes, #2 of 2026-05-29
  6. HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System 20 upvotes, #7 of 2026-04-17
  7. From Passive Observer to Active Critic: Reinforcement Learning Elicits Process Reasoning for Robotic Manipulation 6 upvotes, #29 of 2026-03-18
  8. UltraDexGrasp: Learning Universal Dexterous Grasping for Bimanual Robots with Synthetic Data 7 upvotes, #14 of 2026-03-06
  9. BiManiBench: A Hierarchical Benchmark for Evaluating Bimanual Coordination of Multimodal Large Language Models 3 upvotes, #19 of 2026-02-19
  10. Expertise need not monopolize: Action-Specialized Mixture of Experts for Vision-Language-Action Learning 8 upvotes, #27 of 2025-10-17
  11. VER: Vision Expert Transformer for Robot Learning via Foundation Distillation and Dynamic Routing 5 upvotes, #37 of 2025-10-14
  12. Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies 28 upvotes, #5 of 2025-08-28
  13. HyCodePolicy: Hybrid Language Controllers for Multimodal Monitoring and Decision in Embodied Agents 5 upvotes, #13 of 2025-08-06
  14. RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation 16 upvotes, #7 of 2025-06-26
  15. Plot2Code: A Comprehensive Benchmark for Evaluating Multi-modal Large Language Models in Code Generation from Scientific Plots 15 upvotes, #5 of 2024-05-14

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.