Yonatan
Yonatan on Hugging Face Daily Papers: 9 papers, 5 in the top 3 of their day, 313 upvotes.
- PaliGemma 2: A Family of Versatile VLMs for Transfer 109 upvotes, #1 of 2024-12-05
- NL-Eye: Abductive NLI for Images 22 upvotes, #3 of 2024-10-07
- Visual Riddles: a Commonsense and World Knowledge Challenge for Large Vision and Language Models 22 upvotes, #11 of 2024-07-30
- Video-STaR: Self-Training Enables Video Instruction Tuning with Any Supervision 24 upvotes, #3 of 2024-07-10
- DataComp-LM: In search of the next generation of training sets for language models 33 upvotes, #2 of 2024-06-18
- DOCCI: Descriptions of Connected and Contrasting Images 5 upvotes, #13 of 2024-05-01
- VisIT-Bench: A Benchmark for Vision-Language Instruction Following Inspired by Real-World Use 6 upvotes, #9 of 2023-08-15
- OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models 34 upvotes, #3 of 2023-08-04
- What You See is What You Read? Improving Text-Image Alignment Evaluation 2 upvotes, #7 of 2023-05-18
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.