QinghaoYe
QinghaoYe on Hugging Face Daily Papers: 7 papers, 4 in the top 3 of their day, 290 upvotes.
- Seed1.5-VL Technical Report 136 upvotes, #1 of 2025-05-13
- Painting with Words: Elevating Detailed Image Captioning with Benchmark and Alignment Learning 4 upvotes, #40 of 2025-03-21
- LLaVA-Critic: Learning to Evaluate Multimodal Models 31 upvotes, #6 of 2024-10-04
- MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation? 49 upvotes, #1 of 2024-07-09
- mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration 22 upvotes, #2 of 2023-11-09
- mPLUG-DocOwl: Modularized Multimodal Large Language Model for Document Understanding 16 upvotes, #3 of 2023-07-07
- Youku-mPLUG: A 10 Million Large-scale Chinese Video-Language Dataset for Pre-training and Benchmarks 2 upvotes, #9 of 2023-06-08
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.