Haoning Wu, Teo

Haoning Wu, Teo on Hugging Face Daily Papers: 10 papers, 4 in the top 3 of their day, 677 upvotes.

  1. Kimi K2.5: Visual Agentic Intelligence 219 upvotes, #2 of 2026-02-03
  2. VideoReasonBench: Can MLLMs Perform Vision-Centric Complex Video Reasoning? 39 upvotes, #6 of 2025-05-30
  3. Kimi-VL Technical Report 113 upvotes, #1 of 2025-04-11
  4. VideoAutoArena: An Automated Arena for Evaluating Large Multimodal Models in Video Analysis through User Simulation 17 upvotes, #3 of 2024-11-21
  5. Aria: An Open Multimodal Native Mixture-of-Experts Model 102 upvotes, #1 of 2024-10-10
  6. LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding 18 upvotes, #5 of 2024-07-23
  7. Towards Open-ended Visual Quality Comparison 17 upvotes, #9 of 2024-02-27
  8. Towards A Better Metric for Text-to-Video Generation 15 upvotes, #7 of 2024-01-17
  9. Q-Refine: A Perceptual Quality Refiner for AI-Generated Image 8 upvotes, #11 of 2024-01-03
  10. Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models 27 upvotes, #4 of 2023-11-14

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.