Benjamin Feuer
Benjamin Feuer on Hugging Face Daily Papers: 10 papers, 0 in the top 3 of their day, 223 upvotes.
- OpenThoughts-Agent: Data Recipes for Agentic Models 46 upvotes, #4 of 2026-06-24
- When Judgment Becomes Noise: How Design Failures in LLM Judge Benchmarks Silently Undermine Validity 6 upvotes, #22 of 2025-09-26
- MARVIS: Modality Adaptive Reasoning over VISualizations 11 upvotes, #8 of 2025-07-03
- OpenThoughts: Data Recipes for Reasoning Models 39 upvotes, #4 of 2025-06-05
- WILDCHAT-50M: A Deep Dive Into the Role of Synthetic Data in Post-Training 17 upvotes, #7 of 2025-01-31
- Hidden in the Noise: Two-Stage Robust Watermarking for Images 28 upvotes, #5 of 2024-12-11
- SELECT: A Large-Scale Benchmark of Data Curation Strategies for Image Classification 7 upvotes, #17 of 2024-10-08
- Style over Substance: Failure Modes of LLM Judges in Alignment Benchmarking 11 upvotes, #7 of 2024-09-24
- Arboretum: A Large Multimodal Dataset Enabling AI for Biodiversity 7 upvotes, #7 of 2024-07-01
- LiveBench: A Challenging, Contamination-Free LLM Benchmark 12 upvotes, #9 of 2024-06-28
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.