Benjamin Feuer

Benjamin Feuer on Hugging Face Daily Papers: 10 papers, 0 in the top 3 of their day, 223 upvotes.

  1. OpenThoughts-Agent: Data Recipes for Agentic Models 46 upvotes, #4 of 2026-06-24
  2. When Judgment Becomes Noise: How Design Failures in LLM Judge Benchmarks Silently Undermine Validity 6 upvotes, #22 of 2025-09-26
  3. MARVIS: Modality Adaptive Reasoning over VISualizations 11 upvotes, #8 of 2025-07-03
  4. OpenThoughts: Data Recipes for Reasoning Models 39 upvotes, #4 of 2025-06-05
  5. WILDCHAT-50M: A Deep Dive Into the Role of Synthetic Data in Post-Training 17 upvotes, #7 of 2025-01-31
  6. Hidden in the Noise: Two-Stage Robust Watermarking for Images 28 upvotes, #5 of 2024-12-11
  7. SELECT: A Large-Scale Benchmark of Data Curation Strategies for Image Classification 7 upvotes, #17 of 2024-10-08
  8. Style over Substance: Failure Modes of LLM Judges in Alignment Benchmarking 11 upvotes, #7 of 2024-09-24
  9. Arboretum: A Large Multimodal Dataset Enabling AI for Biodiversity 7 upvotes, #7 of 2024-07-01
  10. LiveBench: A Challenging, Contamination-Free LLM Benchmark 12 upvotes, #9 of 2024-06-28

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.