Yiyuan Zhang
Yiyuan Zhang on Hugging Face Daily Papers: 12 papers, 3 in the top 3 of their day, 448 upvotes.
- GenFirst: Generation Before Reconstruction for Stable End-to-End Latent Generative Modeling 67 upvotes, #4 of 2026-09-01
- OneThinker: All-in-one Reasoning Model for Image and Video 30 upvotes, #4 of 2025-12-04
- Transition Models: Rethinking the Generative Learning Objective 28 upvotes, #6 of 2025-09-05
- Native-Resolution Image Synthesis 18 upvotes, #18 of 2025-06-04
- Seed1.5-VL Technical Report 136 upvotes, #1 of 2025-05-13
- Scaling Up Your Kernels: Large Kernel Design in ConvNets towards Universal Representations 8 upvotes, #16 of 2024-10-11
- Explore the Limits of Omni-modal Pretraining at Scale 10 upvotes, #16 of 2024-06-14
- InteractiveVideo: User-Centric Controllable Video Generation with Synergistic Multimodal Instructions 19 upvotes, #8 of 2024-02-06
- Multimodal Pathway: Improve Transformers with Irrelevant Data from Other Modalities 13 upvotes, #8 of 2024-01-26
- Text-to-3D Generation with Bidirectional Diffusion using both 2D and 3D priors 17 upvotes, #3 of 2023-12-11
- OneLLM: One Framework to Align All Modalities with Language 23 upvotes, #6 of 2023-12-06
- Meta-Transformer: A Unified Framework for Multimodal Learning 45 upvotes, #2 of 2023-07-21
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.