Daily Papers of 2025-07-07

  1. How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks 33 upvotes, #1 of 2025-07-07
  2. Lost in Latent Space: An Empirical Study of Latent Diffusion Models for Physics Emulation 20 upvotes, #2 of 2025-07-07
  3. Eka-Eval : A Comprehensive Evaluation Framework for Large Language Models in Indian Languages 11 upvotes, #3 of 2025-07-07
  4. LitBench: A Benchmark and Dataset for Reliable Evaluation of Creative Writing 4 upvotes, #4 of 2025-07-07

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.