Wenhao Huang

Wenhao Huang on Hugging Face Daily Papers: 7 papers, 2 in the top 3 of their day, 260 upvotes.

  1. FormalMATH: Benchmarking Formal Mathematical Reasoning of Large Language Models 27 upvotes, #7 of 2025-05-06
  2. IV-Bench: A Benchmark for Image-Grounded Video Perception and Reasoning in Multimodal LLMs 22 upvotes, #8 of 2025-04-23
  3. COIG-P: A High-Quality and Large-Scale Chinese Preference Dataset for Alignment with Human Values 41 upvotes, #5 of 2025-04-09
  4. YuE: Scaling Open Foundation Models for Long-Form Music Generation 57 upvotes, #3 of 2025-03-12
  5. CodeCriticBench: A Holistic Code Critique Benchmark for Large Language Models 23 upvotes, #8 of 2025-02-25
  6. PopAlign: Diversifying Contrasting Patterns for a More Comprehensive Alignment 18 upvotes, #14 of 2024-10-18
  7. AutoCrawler: A Progressive Understanding Web Agent for Web Crawler Generation 33 upvotes, #1 of 2024-04-22

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.