Yizhuo Li
Yizhuo Li on Hugging Face Daily Papers: 7 papers, 2 in the top 3 of their day, 144 upvotes.
- Discrete Diffusion VLA: Bringing Discrete Diffusion to Action Decoding in Vision-Language-Action Policies 28 upvotes, #5 of 2025-08-28
- Aligning Latent Spaces with Flow Priors 25 upvotes, #9 of 2025-06-06
- AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation 22 upvotes, #15 of 2025-06-04
- Divot: Diffusion Powers Video Tokenizer for Comprehension and Generation 13 upvotes, #7 of 2024-12-10
- Moto: Latent Motion Token as the Bridging Language for Robot Manipulation 20 upvotes, #7 of 2024-12-09
- InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation 26 upvotes, #3 of 2023-07-14
- VideoChat: Chat-Centric Video Understanding 3 upvotes, #3 of 2023-05-11
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.