Tianxiang Jiang
Tianxiang Jiang on Hugging Face Daily Papers: 8 papers, 2 in the top 3 of their day, 187 upvotes.
- TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs 165 upvotes, #2 of 2026-07-21
- InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning 22 upvotes, #12 of 2026-06-11
- Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction 6 upvotes, #25 of 2026-06-05
- RIVER: A Real-Time Interaction Benchmark for Video LLMs 5 upvotes, #14 of 2026-03-05
- LaViT: Aligning Latent Visual Thoughts for Multi-modal Reasoning 11 upvotes, #24 of 2026-01-16
- ExpVid: A Benchmark for Experiment Video Understanding & Reasoning 3 upvotes, #32 of 2025-10-15
- Make Your Training Flexible: Towards Deployment-Efficient Video Models 5 upvotes, #38 of 2025-03-21
- InternVideo2: Scaling Video Foundation Models for Multimodal Video Understanding 14 upvotes, #3 of 2024-03-25
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.