DasolChoi
DasolChoi on Hugging Face Daily Papers: 7 papers, 0 in the top 3 of their day, 112 upvotes.
- NOLLI: A Difficulty-Calibrated Puzzle Benchmark for Diagnosing the English-Korean Performance Gap 23 upvotes, #12 of 2026-08-06
- XL-SafetyBench: A Country-Grounded Cross-Cultural Benchmark for LLM Safety and Cultural Sensitivity 11 upvotes, #13 of 2026-05-07
- What Users Leave Unsaid: Under-Specified Queries Limit Vision-Language Models 15 upvotes, #15 of 2026-01-13
- COMPASS: A Framework for Evaluating Organization-Specific Policy Alignment in LLMs 6 upvotes, #17 of 2026-01-06
- Global PIQA: Evaluating Physical Commonsense Reasoning Across 100+ Languages and Cultures 9 upvotes, #26 of 2025-10-29
- Pushing on Multilingual Reasoning Models with Language-Mixed Chain-of-Thought 24 upvotes, #11 of 2025-10-09
- When Good Sounds Go Adversarial: Jailbreaking Audio-Language Models with Benign Inputs 2 upvotes, #28 of 2025-08-12
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.