Xiangyu Yue
Xiangyu Yue on Hugging Face Daily Papers: 13 papers, 4 in the top 3 of their day, 535 upvotes.
- NaTex: Seamless Texture Generation as Latent Color Diffusion 15 upvotes, #12 of 2025-11-21
- SciReasoner: Laying the Scientific Reasoning Ground Across Disciplines 93 upvotes, #3 of 2025-09-26
- ScaleCUA: Scaling Open-Source Computer Use Agents with Cross-Platform Data 101 upvotes, #1 of 2025-09-19
- Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations 25 upvotes, #9 of 2025-06-24
- MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs 81 upvotes, #3 of 2025-05-28
- Video-R1: Reinforcing Video Reasoning in MLLMs 74 upvotes, #1 of 2025-03-28
- Unleashing Vecset Diffusion Model for Fast Shape Generation 40 upvotes, #7 of 2025-03-21
- Unposed Sparse Views Room Layout Reconstruction in the Age of Pretrain Model 2 upvotes, #26 of 2025-03-04
- Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation 11 upvotes, #13 of 2025-02-25
- Chimera: Improving Generalist Model with Domain-Specific Experts 9 upvotes, #19 of 2024-12-11
- AV-Odyssey Bench: Can Your Multimodal LLMs Really Understand Audio-Visual Information? 21 upvotes, #6 of 2024-12-04
- Remember, Retrieve and Generate: Understanding Infinite Visual Concepts as Your Personalized Assistant 8 upvotes, #20 of 2024-10-18
- Scaling Up Your Kernels: Large Kernel Design in ConvNets towards Universal Representations 8 upvotes, #16 of 2024-10-11
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.