Tencent Hunyuan

Tencent Hunyuan on Hugging Face Daily Papers: 84 papers, 14 in the top 3 of their day, 4 paper of the day.

  1. Scaling Trajectories for Complex Tasks through Recursive Self-Rewrite 10 upvotes, #2 of 2026-10-05
  2. AutoGUIWorld: Image Generators as Visual World Models for GUI Agent 55 upvotes, #21 of 2026-10-02
  3. Does Native 3D Texture Generation Necessarily Require 3D Assets for Training? 8 upvotes, #67 of 2026-10-02
  4. How Far Are We from Removing the Visual Encoder? Scaling Laws for Encoder-Free Multimodal Pretraining 65 upvotes, #11 of 2026-09-29
  5. ExplorationBench: Measuring AI Systems' Exploration in Verifiable Alien Worlds 28 upvotes, #8 of 2026-09-25
  6. Hunyuan-A13B Technical Report 24 upvotes, #9 of 2026-09-24
  7. SAS: Simple Attention Sparsification via End-to-End Optimization of Context Ranking 65 upvotes, #7 of 2026-09-14
  8. T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks 59 upvotes, #5 of 2026-09-10
  9. AuK Technical Report: An Open-Source Foundational Model for Speech Generation and Editing 218 upvotes, #2 of 2026-09-09
  10. Omni Interaction Agent Technical Report 134 upvotes, #3 of 2026-09-09
  11. FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience 98 upvotes, #2 of 2026-09-08
  12. Environment Evolution for Terminal Agents 22 upvotes, #21 of 2026-09-04
  13. DramaChain Bench: An End-to-End Benchmark for Short-Drama Generation 6 upvotes, #26 of 2026-09-02
  14. Dynamic Important Example Mining for Reinforcement Finetuning 4 upvotes, #34 of 2026-09-01
  15. CAFE: Self-Improving Search Agents Need Co-Evolving Feedback 6 upvotes, #17 of 2026-08-26
  16. GameXpert-Bench: How Far Are Coding Agents from Expert Game Development? 17 upvotes, #12 of 2026-08-25
  17. WithEveryone: Unified Planning and Identity Grounding for Group Image Generation 42 upvotes, #4 of 2026-08-21
  18. UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations 47 upvotes, #6 of 2026-08-18
  19. WorldClaw: Agentic 3D Open-World Generation at Scale 68 upvotes, #4 of 2026-08-07
  20. Recursive Synthesis for Long-Horizon Terminal Tasks 239 upvotes, #1 of 2026-08-06
  21. Hunyuan3D-Buffalo 1.0: A Unified Multimodal Model for Scalable 3D Generation, Understanding, and Editing 88 upvotes, #3 of 2026-08-05
  22. Scaling Native Multimodal Pre-Training From Scratch 28 upvotes, #4 of 2026-07-27
  23. Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning 35 upvotes, #8 of 2026-07-22
  24. MeanFlowNFT: Bringing Forward-Process RL to Average-Velocity Generators 17 upvotes, #16 of 2026-07-17
  25. Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable 216 upvotes, #1 of 2026-07-16
  26. Long-Horizon-Terminal-Bench: Testing the Limits of Agents on Long-Horizon Terminal Tasks with Dense Reward-Based Grading 74 upvotes, #2 of 2026-07-13
  27. Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling 76 upvotes, #3 of 2026-07-08
  28. TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training 18 upvotes, #13 of 2026-07-08
  29. Optimizing Visual Generative Models via Distribution-wise Rewards 16 upvotes, #13 of 2026-07-03
  30. When Search Agents Should Ask: DiscoBench for Clarification-Aware Deep Search 15 upvotes, #15 of 2026-07-03
  31. GEAR: Guided End-to-End AutoRegression for Image Synthesis 34 upvotes, #9 of 2026-07-01
  32. PolyFlow: Continuous Topology Embedding Flow Matching for Artist-style Mesh Generation 12 upvotes, #23 of 2026-07-01
  33. GUICrafter: Weakly-Supervised GUI Agent Leveraging Massive Unannotated Screenshots 16 upvotes, #17 of 2026-06-30
  34. ViQ: Text-Aligned Visual Quantized Representations at Any Resolution 38 upvotes, #6 of 2026-06-26
  35. VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct 6 upvotes, #17 of 2026-06-24
  36. FastMix: Fast Data Mixture Optimization via Gradient Descent 3 upvotes, #38 of 2026-06-23
  37. Training Open Models for Agentic Phone Use 16 upvotes, #17 of 2026-06-23
  38. STARE: Surprisal-Guided Token-Level Advantage Reweighting for Policy Entropy Stability 12 upvotes, #13 of 2026-06-18
  39. Hy-Embodied-0.5-VLA: From Vision-Language-Action Models to a Real-World Robot Learning Stack 15 upvotes, #17 of 2026-06-15
  40. Stream3D-VLM: Online 3D Spatial Understanding with Incremental Geometry Priors 4 upvotes, #27 of 2026-06-08
  41. GenClaw: Code-Driven Agentic Image Generation 38 upvotes, #9 of 2026-05-29
  42. GEM: Generative Supervision Helps Embodied Intelligence 41 upvotes, #8 of 2026-05-28
  43. Reinforcing Few-step Generators via Reward-Tilted Distribution Matching 5 upvotes, #40 of 2026-05-26
  44. PlanningBench: Generating Scalable and Verifiable Planning Data for Evaluating and Training Large Language Models 29 upvotes, #11 of 2026-05-21
  45. Unlocking Dense Metric Depth Estimation in VLMs 12 upvotes, #18 of 2026-05-18
  46. Learning to Foresee: Unveiling the Unlocking Efficiency of On-Policy Distillation 58 upvotes, #5 of 2026-05-18
  47. Implicit Preference Alignment for Human Image Animation 1 upvotes, #56 of 2026-05-13
  48. FaithfulFaces: Pose-Faithful Facial Identity Preservation for Text-to-Video Generation 1 upvotes, #56 of 2026-05-13
  49. Debiased Model-based Representations for Sample-efficient Continuous Control 9 upvotes, #33 of 2026-05-13
  50. Reinforcing Multimodal Reasoning Against Visual Degradation 7 upvotes, #31 of 2026-05-12
  51. DeltaRubric: Generative Multimodal Reward Modeling via Joint Planning and Verification 6 upvotes, #34 of 2026-05-12
  52. Listwise Policy Optimization: Group-based RLVR as Target-Projection on the LLM Response Simplex 65 upvotes, #4 of 2026-05-11
  53. OpenSearch-VL: An Open Recipe for Frontier Multimodal Search Agents 96 upvotes, #4 of 2026-05-07
  54. Toward Scalable Terminal Task Synthesis via Skill Graphs 11 upvotes, #10 of 2026-04-29
  55. Hierarchical SVG Tokenization: Learning Compact Visual Programs for Scalable Vector Graphics Modeling 18 upvotes, #12 of 2026-04-15
  56. HY-Embodied-0.5: Embodied Foundation Models for Real-World Agents 182 upvotes, #4 of 2026-04-10
  57. Unify-Agent: A Unified Multimodal Agent for World-Grounded Image Synthesis 46 upvotes, #10 of 2026-04-01
  58. Manifold-Aware Exploration for Reinforcement Learning in Video Generation 33 upvotes, #10 of 2026-03-24
  59. Spatial-TTT: Streaming Visual-based Spatial Intelligence with Test-Time Training 90 upvotes, #1 of 2026-03-13
  60. UniCom: Unified Multimodal Modeling via Compressed Continuous Semantic Representations 4 upvotes, #22 of 2026-03-12
  61. HiAR: Efficient Autoregressive Long Video Generation via Hierarchical Denoising 31 upvotes, #8 of 2026-03-10
  62. HY-WU (Part I): An Extensible Functional Neural Memory Framework and An Instantiation in Text-Guided Image Editing 3 upvotes, #30 of 2026-03-10
  63. Beyond Length Scaling: Synergizing Breadth and Depth for Generative Reward Models 33 upvotes, #6 of 2026-03-04
  64. ArtLLM: Generating Articulated Assets via 3D LLM 3 upvotes, #32 of 2026-03-03
  65. RubricBench: Aligning Model-Generated Rubrics with Human Standards 55 upvotes, #4 of 2026-03-03
  66. Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation 57 upvotes, #4 of 2026-02-13
  67. Composition-RL: Compose Your Verifiable Prompts for Reinforcement Learning of Large Language Models 92 upvotes, #2 of 2026-02-13
  68. WorldCompass: Reinforcement Learning for Long-Horizon World Models 20 upvotes, #21 of 2026-02-10
  69. ProAct: Agentic Lookahead in Interactive Environments 25 upvotes, #11 of 2026-02-06
  70. HY3D-Bench: Generation of 3D Assets 22 upvotes, #16 of 2026-02-05
  71. Search-R2: Enhancing Search-Integrated Reasoning via Actor-Refiner Collaboration 7 upvotes, #29 of 2026-02-04
  72. iFSQ: Improving FSQ for Image Generation with 1 Line of Code 31 upvotes, #7 of 2026-01-27
  73. The Script is All You Need: An Agentic Framework for Long-Horizon Dialogue-to-Cinematic Video Generation 55 upvotes, #3 of 2026-01-27
  74. TAG-MoE: Task-Aware Gating for Unified Generative Mixture-of-Experts 13 upvotes, #22 of 2026-01-16
  75. Diversity or Precision? A Deep Dive into Next Token Prediction 7 upvotes, #11 of 2026-01-05
  76. Video Generation Models Are Good Latent Reward Models 44 upvotes, #1 of 2025-11-28
  77. Harmony: Harmonizing Audio and Video Generation through Cross-Task Synergy 21 upvotes, #4 of 2025-11-27
  78. HunyuanOCR Technical Report 19 upvotes, #12 of 2025-11-26
  79. GeoVista: Web-Augmented Agentic Visual Reasoning for Geolocalization 89 upvotes, #2 of 2025-11-24
  80. NaTex: Seamless Texture Generation as Latent Color Diffusion 15 upvotes, #12 of 2025-11-21
  81. Part-X-MLLM: Part-aware 3D Multimodal Large Language Model 70 upvotes, #5 of 2025-11-18
  82. LaSeR: Reinforcement Learning with Last-Token Self-Rewarding 37 upvotes, #9 of 2025-10-17
  83. DA^2: Depth Anything in Any Direction 20 upvotes, #14 of 2025-10-01
  84. P3-SAM: Native 3D Part Segmentation 16 upvotes, #6 of 2025-09-11

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.