Hyunwoo Ko

Hyunwoo Ko on Hugging Face Daily Papers: 7 papers, 1 in the top 3 of their day, 231 upvotes.

  1. ResearchMath-14K: Scaling Research-Level Mathematics via Agents 49 upvotes, #6 of 2026-05-28
  2. Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs 77 upvotes, #2 of 2026-05-12
  3. Judging What We Cannot Solve: A Consequence-Based Approach for Oracle-Free Evaluation of Research-Level Math 22 upvotes, #10 of 2026-02-09
  4. What Users Leave Unsaid: Under-Specified Queries Limit Vision-Language Models 15 upvotes, #15 of 2026-01-13
  5. Pushing on Multilingual Reasoning Models with Language-Mixed Chain-of-Thought 24 upvotes, #11 of 2025-10-09
  6. When AI Co-Scientists Fail: SPOT-a Benchmark for Automated Verification of Scientific Research 9 upvotes, #23 of 2025-05-20
  7. Linguistic Generalizability of Test-Time Scaling in Mathematical Reasoning 24 upvotes, #7 of 2025-02-25

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.