Dawei Li

Dawei Li on Hugging Face Daily Papers: 6 papers, 2 in the top 3 of their day, 381 upvotes.

  1. ToolPRMBench: Evaluating and Advancing Process Reward Models for Tool-using Agents 17 upvotes, #13 of 2026-01-21
  2. Who's Your Judge? On the Detectability of LLM-Generated Judgments 27 upvotes, #11 of 2025-10-01
  3. Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens 204 upvotes, #1 of 2025-08-07
  4. The Quest for Efficient Reasoning: A Data-Centric Benchmark to CoT Distillation 12 upvotes, #31 of 2025-05-27
  5. Preference Leakage: A Contamination Problem in LLM-as-a-judge 34 upvotes, #4 of 2025-02-04
  6. From Generation to Judgment: Opportunities and Challenges of LLM-as-a-judge 35 upvotes, #2 of 2024-11-26

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.