Mario Sanz
Mario Sanz on Hugging Face Daily Papers: 5 papers, 0 in the top 3 of their day, 22 upvotes.
- Dating the Model: Hidden Dates in System Prompts Affect LLM Evaluation 3 upvotes, #66 of 2026-10-01
- Calibration as a First-Class Criterion in LLM Evaluation 4 upvotes, #27 of 2026-09-24
- Large Language Models Are Overconfident in Their Own Responses 3 upvotes, #29 of 2026-06-11
- Mitigating Label Length Bias in Large Language Models 6 upvotes, #14 of 2025-11-19
- Mind the Gap: A Closer Look at Tokenization for Multiple-Choice Question Answering with LLMs 4 upvotes, #16 of 2025-09-19
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.