QinghaoYe

QinghaoYe on Hugging Face Daily Papers: 7 papers, 4 in the top 3 of their day, 290 upvotes.

  1. Seed1.5-VL Technical Report 136 upvotes, #1 of 2025-05-13
  2. Painting with Words: Elevating Detailed Image Captioning with Benchmark and Alignment Learning 4 upvotes, #40 of 2025-03-21
  3. LLaVA-Critic: Learning to Evaluate Multimodal Models 31 upvotes, #6 of 2024-10-04
  4. MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation? 49 upvotes, #1 of 2024-07-09
  5. mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration 22 upvotes, #2 of 2023-11-09
  6. mPLUG-DocOwl: Modularized Multimodal Large Language Model for Document Understanding 16 upvotes, #3 of 2023-07-07
  7. Youku-mPLUG: A 10 Million Large-scale Chinese Video-Language Dataset for Pre-training and Benchmarks 2 upvotes, #9 of 2023-06-08

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.