Mario Sanz

Mario Sanz on Hugging Face Daily Papers: 5 papers, 0 in the top 3 of their day, 22 upvotes.

  1. Dating the Model: Hidden Dates in System Prompts Affect LLM Evaluation 3 upvotes, #66 of 2026-10-01
  2. Calibration as a First-Class Criterion in LLM Evaluation 4 upvotes, #27 of 2026-09-24
  3. Large Language Models Are Overconfident in Their Own Responses 3 upvotes, #29 of 2026-06-11
  4. Mitigating Label Length Bias in Large Language Models 6 upvotes, #14 of 2025-11-19
  5. Mind the Gap: A Closer Look at Tokenization for Multiple-Choice Question Answering with LLMs 4 upvotes, #16 of 2025-09-19

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.