Xinlong Chen

Xinlong Chen on Hugging Face Daily Papers: 4 papers, 0 in the top 3 of their day, 143 upvotes.

  1. AVoCaDO: An Audiovisual Video Captioner Driven by Temporal Orchestration 28 upvotes, #10 of 2025-10-14
  2. RealUnify: Do Unified Models Truly Benefit from Unification? A Comprehensive Benchmark 44 upvotes, #6 of 2025-09-30
  3. MME-VideoOCR: Evaluating OCR-Based Capabilities of Multimodal LLMs in Video Scenarios 38 upvotes, #12 of 2025-05-28
  4. Mavors: Multi-granularity Video Representation for Multimodal Large Language Model 30 upvotes, #7 of 2025-04-15

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.