Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Bin Lin, Bin Zhu, Yang Ye, Munan Ning, Peng Jin, Li Yuan
Video-LLaVA: Learning United Visual Representation by Alignment Before Projection: 28 upvotes on Hugging Face Daily Papers, #1 of 11 papers on 2023-11-20. Day-by-day upvote history.
Paper page on Hugging Face · arXiv
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.