Chen Chen
Chen Chen on Hugging Face Daily Papers: 8 papers, 3 in the top 3 of their day, 284 upvotes.
- CAR-Flow: Condition-Aware Reparameterization Aligns Source and Target for Better Flow Matching 4 upvotes, #15 of 2025-09-24
- MANZANO: A Simple and Scalable Unified Multimodal Model with a Hybrid Vision Tokenizer 49 upvotes, #2 of 2025-09-22
- AToken: A Unified Tokenizer for Vision 30 upvotes, #5 of 2025-09-19
- GIE-Bench: Towards Grounded Evaluation for Text-Guided Image Editing 3 upvotes, #15 of 2025-05-19
- DiT-Air: Revisiting the Efficiency of Diffusion Model Architecture Design in Text to Image Generation 17 upvotes, #17 of 2025-03-14
- STIV: Scalable Text and Image Conditioned Video Generation 67 upvotes, #1 of 2024-12-11
- Revisit Large-Scale Image-Caption Data in Pre-training Multimodal Foundation Models 51 upvotes, #1 of 2024-10-04
- Ferret-v2: An Improved Baseline for Referring and Grounding with Large Language Models 26 upvotes, #6 of 2024-04-12
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.