Yuxuan Wang

Yuxuan Wang on Hugging Face Daily Papers: 12 papers, 2 in the top 3 of their day, 516 upvotes.

  1. Qwen3-VL Technical Report 120 upvotes, #1 of 2025-12-04
  2. OmniVideoBench: Towards Audio-Visual Understanding Evaluation for Omni MLLMs 45 upvotes, #4 of 2025-10-14
  3. Qwen3-Omni Technical Report 121 upvotes, #1 of 2025-09-23
  4. Discrete Markov Bridge 17 upvotes, #21 of 2025-05-27
  5. Seek in the Dark: Reasoning via Test-Time Instance-Level Policy Gradient in Latent Space 26 upvotes, #11 of 2025-05-20
  6. OmniMMI: A Comprehensive Multi-modal Interaction Benchmark in Streaming Video Contexts 18 upvotes, #15 of 2025-04-02
  7. From Hours to Minutes: Lossless Acceleration of Ultra Long Sequence Generation up to 100K Tokens 24 upvotes, #6 of 2025-03-04
  8. Friends-MMC: A Dataset for Multi-modal Multi-party Conversation Understanding 8 upvotes, #14 of 2024-12-24
  9. VideoLLM Knows When to Speak: Enhancing Time-Sensitive Video Comprehension with Video-Text Duet Interaction Format 5 upvotes, #18 of 2024-11-28
  10. VideoLLaMB: Long-context Video Understanding with Recurrent Memory Bridges 26 upvotes, #8 of 2024-09-04
  11. ExoViP: Step-by-step Verification and Exploration with Exoskeleton Modules for Compositional Visual Reasoning 6 upvotes, #11 of 2024-08-06
  12. VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models 22 upvotes, #6 of 2024-06-25

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.