Fuxiao Liu
Fuxiao Liu on Hugging Face Daily Papers: 8 papers, 5 in the top 3 of their day, 396 upvotes.
- MM-Zero: Self-Evolving Multi-Model Vision Language Models From Zero Data 51 upvotes, #3 of 2026-03-11
- First Frame Is the Place to Go for Video Content Customization 51 upvotes, #4 of 2025-11-21
- NVIDIA Nemotron Nano V2 VL 25 upvotes, #5 of 2025-11-07
- Self-Rewarding Vision-Language Model via Reasoning Decomposition 77 upvotes, #2 of 2025-08-28
- ColorBench: Can VLMs See and Understand the Colorful World? A Comprehensive Benchmark for Color Perception, Reasoning, and Robustness 45 upvotes, #3 of 2025-04-17
- Eagle: Exploring The Design Space for Multimodal LLMs with Mixture of Encoders 76 upvotes, #1 of 2024-08-29
- HallusionBench: You See What You Think? Or You Think What You See? An Image-Context Reasoning Benchmark Challenging for GPT-4V(ision), LLaVA-1.5, and Other Multi-modality Models 27 upvotes, #2 of 2023-10-24
- Aligning Large Multi-Modal Model with Robust Instruction Tuning 6 upvotes, #11 of 2023-06-27
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.