Hyunwoo Ko
Hyunwoo Ko on Hugging Face Daily Papers: 7 papers, 1 in the top 3 of their day, 231 upvotes.
- ResearchMath-14K: Scaling Research-Level Mathematics via Agents 49 upvotes, #6 of 2026-05-28
- Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs 77 upvotes, #2 of 2026-05-12
- Judging What We Cannot Solve: A Consequence-Based Approach for Oracle-Free Evaluation of Research-Level Math 22 upvotes, #10 of 2026-02-09
- What Users Leave Unsaid: Under-Specified Queries Limit Vision-Language Models 15 upvotes, #15 of 2026-01-13
- Pushing on Multilingual Reasoning Models with Language-Mixed Chain-of-Thought 24 upvotes, #11 of 2025-10-09
- When AI Co-Scientists Fail: SPOT-a Benchmark for Automated Verification of Scientific Research 9 upvotes, #23 of 2025-05-20
- Linguistic Generalizability of Test-Time Scaling in Mathematical Reasoning 24 upvotes, #7 of 2025-02-25
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.