Xiaoqian Shen
Xiaoqian Shen on Hugging Face Daily Papers: 5 papers, 3 in the top 3 of their day, 94 upvotes.
- LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding 18 upvotes, #2 of 2024-10-24
- Openstory++: A Large-scale Dataset and Benchmark for Instance-aware Open-domain Visual Storytelling 10 upvotes, #6 of 2024-08-08
- Goldfish: Vision-Language Understanding of Arbitrarily Long Videos 6 upvotes, #10 of 2024-07-18
- MiniGPT4-Video: Advancing Multimodal LLMs for Video Understanding with Interleaved Visual-Textual Tokens 21 upvotes, #3 of 2024-04-05
- MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning 21 upvotes, #3 of 2023-10-17
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.