Haoning Wu, Teo
Haoning Wu, Teo on Hugging Face Daily Papers: 10 papers, 4 in the top 3 of their day, 677 upvotes.
- Kimi K2.5: Visual Agentic Intelligence 219 upvotes, #2 of 2026-02-03
- VideoReasonBench: Can MLLMs Perform Vision-Centric Complex Video Reasoning? 39 upvotes, #6 of 2025-05-30
- Kimi-VL Technical Report 113 upvotes, #1 of 2025-04-11
- VideoAutoArena: An Automated Arena for Evaluating Large Multimodal Models in Video Analysis through User Simulation 17 upvotes, #3 of 2024-11-21
- Aria: An Open Multimodal Native Mixture-of-Experts Model 102 upvotes, #1 of 2024-10-10
- LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding 18 upvotes, #5 of 2024-07-23
- Towards Open-ended Visual Quality Comparison 17 upvotes, #9 of 2024-02-27
- Towards A Better Metric for Text-to-Video Generation 15 upvotes, #7 of 2024-01-17
- Q-Refine: A Perceptual Quality Refiner for AI-Generated Image 8 upvotes, #11 of 2024-01-03
- Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models 27 upvotes, #4 of 2023-11-14
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.