Tianxiang Jiang

Tianxiang Jiang on Hugging Face Daily Papers: 8 papers, 2 in the top 3 of their day, 187 upvotes.

  1. TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs 165 upvotes, #2 of 2026-07-21
  2. InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning 22 upvotes, #12 of 2026-06-11
  3. Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction 6 upvotes, #25 of 2026-06-05
  4. RIVER: A Real-Time Interaction Benchmark for Video LLMs 5 upvotes, #14 of 2026-03-05
  5. LaViT: Aligning Latent Visual Thoughts for Multi-modal Reasoning 11 upvotes, #24 of 2026-01-16
  6. ExpVid: A Benchmark for Experiment Video Understanding & Reasoning 3 upvotes, #32 of 2025-10-15
  7. Make Your Training Flexible: Towards Deployment-Efficient Video Models 5 upvotes, #38 of 2025-03-21
  8. InternVideo2: Scaling Video Foundation Models for Multimodal Video Understanding 14 upvotes, #3 of 2024-03-25

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.