Jaewoo Ahn

Jaewoo Ahn on Hugging Face Daily Papers: 6 papers, 0 in the top 3 of their day, 81 upvotes.

  1. Science or Slop?: Benchmarking and Mitigating Scientific Slop in AI-Generated Papers 26 upvotes, #13 of 2026-10-05
  2. Lies We Can See: Joint Verbal and Non-Verbal Deception by VLM Agents in Embodied Social Interactions 14 upvotes, #16 of 2026-09-01
  3. FlashAdventure: A Benchmark for GUI Agents Solving Full Story Arcs in Diverse Adventure Games 19 upvotes, #17 of 2025-09-03
  4. ChartCap: Mitigating Hallucination of Dense Chart Captioning 5 upvotes, #13 of 2025-08-06
  5. Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games 9 upvotes, #20 of 2025-06-05
  6. Can LLMs Deceive CLIP? Benchmarking Adversarial Compositionality of Pre-trained Multimodal Representation via Text Updates 4 upvotes, #49 of 2025-05-30

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.