DasolChoi

DasolChoi on Hugging Face Daily Papers: 7 papers, 0 in the top 3 of their day, 112 upvotes.

  1. NOLLI: A Difficulty-Calibrated Puzzle Benchmark for Diagnosing the English-Korean Performance Gap 23 upvotes, #12 of 2026-08-06
  2. XL-SafetyBench: A Country-Grounded Cross-Cultural Benchmark for LLM Safety and Cultural Sensitivity 11 upvotes, #13 of 2026-05-07
  3. What Users Leave Unsaid: Under-Specified Queries Limit Vision-Language Models 15 upvotes, #15 of 2026-01-13
  4. COMPASS: A Framework for Evaluating Organization-Specific Policy Alignment in LLMs 6 upvotes, #17 of 2026-01-06
  5. Global PIQA: Evaluating Physical Commonsense Reasoning Across 100+ Languages and Cultures 9 upvotes, #26 of 2025-10-29
  6. Pushing on Multilingual Reasoning Models with Language-Mixed Chain-of-Thought 24 upvotes, #11 of 2025-10-09
  7. When Good Sounds Go Adversarial: Jailbreaking Audio-Language Models with Benign Inputs 2 upvotes, #28 of 2025-08-12

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.