Fuxiao Liu

Fuxiao Liu on Hugging Face Daily Papers: 8 papers, 5 in the top 3 of their day, 396 upvotes.

  1. MM-Zero: Self-Evolving Multi-Model Vision Language Models From Zero Data 51 upvotes, #3 of 2026-03-11
  2. First Frame Is the Place to Go for Video Content Customization 51 upvotes, #4 of 2025-11-21
  3. NVIDIA Nemotron Nano V2 VL 25 upvotes, #5 of 2025-11-07
  4. Self-Rewarding Vision-Language Model via Reasoning Decomposition 77 upvotes, #2 of 2025-08-28
  5. ColorBench: Can VLMs See and Understand the Colorful World? A Comprehensive Benchmark for Color Perception, Reasoning, and Robustness 45 upvotes, #3 of 2025-04-17
  6. Eagle: Exploring The Design Space for Multimodal LLMs with Mixture of Encoders 76 upvotes, #1 of 2024-08-29
  7. HallusionBench: You See What You Think? Or You Think What You See? An Image-Context Reasoning Benchmark Challenging for GPT-4V(ision), LLaVA-1.5, and Other Multi-modality Models 27 upvotes, #2 of 2023-10-24
  8. Aligning Large Multi-Modal Model with Robust Instruction Tuning 6 upvotes, #11 of 2023-06-27

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.