Xiaohui Shen

Xiaohui Shen on Hugging Face Daily Papers: 7 papers, 3 in the top 3 of their day, 255 upvotes.

  1. Vidi: Large Multimodal Models for Video Understanding and Editing 15 upvotes, #14 of 2025-04-23
  2. Democratizing Text-to-Image Masked Generative Models with Compact Text-Aware One-Dimensional Tokens 16 upvotes, #9 of 2025-01-15
  3. 1.58-bit FLUX 67 upvotes, #2 of 2024-12-30
  4. Randomized Autoregressive Visual Generation 17 upvotes, #5 of 2024-11-04
  5. Alleviating Distortion in Image Generation via Multi-Resolution Diffusion Models 27 upvotes, #6 of 2024-06-14
  6. An Image is Worth 32 Tokens for Reconstruction and Generation 49 upvotes, #1 of 2024-06-12
  7. COCONut: Modernizing COCO Segmentation 25 upvotes, #2 of 2024-04-15

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.