Tencent Hunyuan
Tencent Hunyuan on Hugging Face Daily Papers: 84 papers, 14 in the top 3 of their day, 4 paper of the day.
- Scaling Trajectories for Complex Tasks through Recursive Self-Rewrite 10 upvotes, #2 of 2026-10-05
- AutoGUIWorld: Image Generators as Visual World Models for GUI Agent 55 upvotes, #21 of 2026-10-02
- Does Native 3D Texture Generation Necessarily Require 3D Assets for Training? 8 upvotes, #67 of 2026-10-02
- How Far Are We from Removing the Visual Encoder? Scaling Laws for Encoder-Free Multimodal Pretraining 65 upvotes, #11 of 2026-09-29
- ExplorationBench: Measuring AI Systems' Exploration in Verifiable Alien Worlds 28 upvotes, #8 of 2026-09-25
- Hunyuan-A13B Technical Report 24 upvotes, #9 of 2026-09-24
- SAS: Simple Attention Sparsification via End-to-End Optimization of Context Ranking 65 upvotes, #7 of 2026-09-14
- T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks 59 upvotes, #5 of 2026-09-10
- AuK Technical Report: An Open-Source Foundational Model for Speech Generation and Editing 218 upvotes, #2 of 2026-09-09
- Omni Interaction Agent Technical Report 134 upvotes, #3 of 2026-09-09
- FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience 98 upvotes, #2 of 2026-09-08
- Environment Evolution for Terminal Agents 22 upvotes, #21 of 2026-09-04
- DramaChain Bench: An End-to-End Benchmark for Short-Drama Generation 6 upvotes, #26 of 2026-09-02
- Dynamic Important Example Mining for Reinforcement Finetuning 4 upvotes, #34 of 2026-09-01
- CAFE: Self-Improving Search Agents Need Co-Evolving Feedback 6 upvotes, #17 of 2026-08-26
- GameXpert-Bench: How Far Are Coding Agents from Expert Game Development? 17 upvotes, #12 of 2026-08-25
- WithEveryone: Unified Planning and Identity Grounding for Group Image Generation 42 upvotes, #4 of 2026-08-21
- UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations 47 upvotes, #6 of 2026-08-18
- WorldClaw: Agentic 3D Open-World Generation at Scale 68 upvotes, #4 of 2026-08-07
- Recursive Synthesis for Long-Horizon Terminal Tasks 239 upvotes, #1 of 2026-08-06
- Hunyuan3D-Buffalo 1.0: A Unified Multimodal Model for Scalable 3D Generation, Understanding, and Editing 88 upvotes, #3 of 2026-08-05
- Scaling Native Multimodal Pre-Training From Scratch 28 upvotes, #4 of 2026-07-27
- Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning 35 upvotes, #8 of 2026-07-22
- MeanFlowNFT: Bringing Forward-Process RL to Average-Velocity Generators 17 upvotes, #16 of 2026-07-17
- Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable 216 upvotes, #1 of 2026-07-16
- Long-Horizon-Terminal-Bench: Testing the Limits of Agents on Long-Horizon Terminal Tasks with Dense Reward-Based Grading 74 upvotes, #2 of 2026-07-13
- Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling 76 upvotes, #3 of 2026-07-08
- TurnOPD: Making On-Policy Distillation Turn-Aware for Efficient Long-Horizon Agent Training 18 upvotes, #13 of 2026-07-08
- Optimizing Visual Generative Models via Distribution-wise Rewards 16 upvotes, #13 of 2026-07-03
- When Search Agents Should Ask: DiscoBench for Clarification-Aware Deep Search 15 upvotes, #15 of 2026-07-03
- GEAR: Guided End-to-End AutoRegression for Image Synthesis 34 upvotes, #9 of 2026-07-01
- PolyFlow: Continuous Topology Embedding Flow Matching for Artist-style Mesh Generation 12 upvotes, #23 of 2026-07-01
- GUICrafter: Weakly-Supervised GUI Agent Leveraging Massive Unannotated Screenshots 16 upvotes, #17 of 2026-06-30
- ViQ: Text-Aligned Visual Quantized Representations at Any Resolution 38 upvotes, #6 of 2026-06-26
- VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct 6 upvotes, #17 of 2026-06-24
- FastMix: Fast Data Mixture Optimization via Gradient Descent 3 upvotes, #38 of 2026-06-23
- Training Open Models for Agentic Phone Use 16 upvotes, #17 of 2026-06-23
- STARE: Surprisal-Guided Token-Level Advantage Reweighting for Policy Entropy Stability 12 upvotes, #13 of 2026-06-18
- Hy-Embodied-0.5-VLA: From Vision-Language-Action Models to a Real-World Robot Learning Stack 15 upvotes, #17 of 2026-06-15
- Stream3D-VLM: Online 3D Spatial Understanding with Incremental Geometry Priors 4 upvotes, #27 of 2026-06-08
- GenClaw: Code-Driven Agentic Image Generation 38 upvotes, #9 of 2026-05-29
- GEM: Generative Supervision Helps Embodied Intelligence 41 upvotes, #8 of 2026-05-28
- Reinforcing Few-step Generators via Reward-Tilted Distribution Matching 5 upvotes, #40 of 2026-05-26
- PlanningBench: Generating Scalable and Verifiable Planning Data for Evaluating and Training Large Language Models 29 upvotes, #11 of 2026-05-21
- Unlocking Dense Metric Depth Estimation in VLMs 12 upvotes, #18 of 2026-05-18
- Learning to Foresee: Unveiling the Unlocking Efficiency of On-Policy Distillation 58 upvotes, #5 of 2026-05-18
- Implicit Preference Alignment for Human Image Animation 1 upvotes, #56 of 2026-05-13
- FaithfulFaces: Pose-Faithful Facial Identity Preservation for Text-to-Video Generation 1 upvotes, #56 of 2026-05-13
- Debiased Model-based Representations for Sample-efficient Continuous Control 9 upvotes, #33 of 2026-05-13
- Reinforcing Multimodal Reasoning Against Visual Degradation 7 upvotes, #31 of 2026-05-12
- DeltaRubric: Generative Multimodal Reward Modeling via Joint Planning and Verification 6 upvotes, #34 of 2026-05-12
- Listwise Policy Optimization: Group-based RLVR as Target-Projection on the LLM Response Simplex 65 upvotes, #4 of 2026-05-11
- OpenSearch-VL: An Open Recipe for Frontier Multimodal Search Agents 96 upvotes, #4 of 2026-05-07
- Toward Scalable Terminal Task Synthesis via Skill Graphs 11 upvotes, #10 of 2026-04-29
- Hierarchical SVG Tokenization: Learning Compact Visual Programs for Scalable Vector Graphics Modeling 18 upvotes, #12 of 2026-04-15
- HY-Embodied-0.5: Embodied Foundation Models for Real-World Agents 182 upvotes, #4 of 2026-04-10
- Unify-Agent: A Unified Multimodal Agent for World-Grounded Image Synthesis 46 upvotes, #10 of 2026-04-01
- Manifold-Aware Exploration for Reinforcement Learning in Video Generation 33 upvotes, #10 of 2026-03-24
- Spatial-TTT: Streaming Visual-based Spatial Intelligence with Test-Time Training 90 upvotes, #1 of 2026-03-13
- UniCom: Unified Multimodal Modeling via Compressed Continuous Semantic Representations 4 upvotes, #22 of 2026-03-12
- HiAR: Efficient Autoregressive Long Video Generation via Hierarchical Denoising 31 upvotes, #8 of 2026-03-10
- HY-WU (Part I): An Extensible Functional Neural Memory Framework and An Instantiation in Text-Guided Image Editing 3 upvotes, #30 of 2026-03-10
- Beyond Length Scaling: Synergizing Breadth and Depth for Generative Reward Models 33 upvotes, #6 of 2026-03-04
- ArtLLM: Generating Articulated Assets via 3D LLM 3 upvotes, #32 of 2026-03-03
- RubricBench: Aligning Model-Generated Rubrics with Human Standards 55 upvotes, #4 of 2026-03-03
- Learning beyond Teacher: Generalized On-Policy Distillation with Reward Extrapolation 57 upvotes, #4 of 2026-02-13
- Composition-RL: Compose Your Verifiable Prompts for Reinforcement Learning of Large Language Models 92 upvotes, #2 of 2026-02-13
- WorldCompass: Reinforcement Learning for Long-Horizon World Models 20 upvotes, #21 of 2026-02-10
- ProAct: Agentic Lookahead in Interactive Environments 25 upvotes, #11 of 2026-02-06
- HY3D-Bench: Generation of 3D Assets 22 upvotes, #16 of 2026-02-05
- Search-R2: Enhancing Search-Integrated Reasoning via Actor-Refiner Collaboration 7 upvotes, #29 of 2026-02-04
- iFSQ: Improving FSQ for Image Generation with 1 Line of Code 31 upvotes, #7 of 2026-01-27
- The Script is All You Need: An Agentic Framework for Long-Horizon Dialogue-to-Cinematic Video Generation 55 upvotes, #3 of 2026-01-27
- TAG-MoE: Task-Aware Gating for Unified Generative Mixture-of-Experts 13 upvotes, #22 of 2026-01-16
- Diversity or Precision? A Deep Dive into Next Token Prediction 7 upvotes, #11 of 2026-01-05
- Video Generation Models Are Good Latent Reward Models 44 upvotes, #1 of 2025-11-28
- Harmony: Harmonizing Audio and Video Generation through Cross-Task Synergy 21 upvotes, #4 of 2025-11-27
- HunyuanOCR Technical Report 19 upvotes, #12 of 2025-11-26
- GeoVista: Web-Augmented Agentic Visual Reasoning for Geolocalization 89 upvotes, #2 of 2025-11-24
- NaTex: Seamless Texture Generation as Latent Color Diffusion 15 upvotes, #12 of 2025-11-21
- Part-X-MLLM: Part-aware 3D Multimodal Large Language Model 70 upvotes, #5 of 2025-11-18
- LaSeR: Reinforcement Learning with Last-Token Self-Rewarding 37 upvotes, #9 of 2025-10-17
- DA^2: Depth Anything in Any Direction 20 upvotes, #14 of 2025-10-01
- P3-SAM: Native 3D Part Segmentation 16 upvotes, #6 of 2025-09-11
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.