taesiri
taesiri on Hugging Face Daily Papers: 11 papers, 2 in the top 3 of their day, 362 upvotes.
- AutoWorldModel-Bench: A State-Centric Benchmark for Automated World-Model Research 13 upvotes, #10 of 2026-08-13
- SketchVLM: Vision language models can annotate images to explain thoughts and guide users 11 upvotes, #10 of 2026-04-28
- A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality 22 upvotes, #10 of 2025-07-11
- Vision Language Models are Biased 20 upvotes, #12 of 2025-06-02
- B-score: Detecting biases in large language models using response history 30 upvotes, #10 of 2025-05-27
- VideoGameQA-Bench: Evaluating Vision-Language Models for Video Game Quality Assurance 19 upvotes, #18 of 2025-05-23
- Understanding Generative AI Capabilities in Everyday Image Editing Tasks 23 upvotes, #12 of 2025-05-23
- HoT: Highlighted Chain of Thought for Referencing Supporting Facts from Inputs 42 upvotes, #2 of 2025-03-06
- ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models 38 upvotes, #5 of 2025-02-17
- VideoGameBunny: Towards vision assistants for video games 18 upvotes, #5 of 2024-07-23
- Vision language models are blind 71 upvotes, #1 of 2024-07-10
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.