Lj V. Miranda

Lj V. Miranda on Hugging Face Daily Papers: 12 papers, 2 in the top 3 of their day, 372 upvotes.

  1. Polyglot Teachers: Evaluating Language Models for Multilingual Synthetic Data Generation 2 upvotes, #40 of 2026-04-14
  2. Olmo 3 22 upvotes, #11 of 2025-12-17
  3. R3: Robust Rubric-Agnostic Reward Models 11 upvotes, #22 of 2025-05-20
  4. Crowdsource, Crawl, or Generate? Creating SEA-VL, a Multicultural Vision-Language Dataset for Southeast Asia 92 upvotes, #1 of 2025-03-12
  5. MMTEB: Massive Multilingual Text Embedding Benchmark 31 upvotes, #5 of 2025-02-20
  6. Bridging the Data Provenance Gap Across Text, Speech and Video 6 upvotes, #10 of 2024-12-25
  7. TÜLU 3: Pushing Frontiers in Open Language Model Post-Training 55 upvotes, #1 of 2024-11-25
  8. Hybrid Preferences: Learning to Route Instances for Human vs. AI Feedback 10 upvotes, #9 of 2024-10-28
  9. M-RewardBench: Evaluating Reward Models in Multilingual Settings 10 upvotes, #7 of 2024-10-24
  10. Consent in Crisis: The Rapid Decline of the AI Data Commons 9 upvotes, #11 of 2024-07-23
  11. SEACrowd: A Multilingual Multimodal Data Hub and Benchmark Suite for Southeast Asian Languages 24 upvotes, #7 of 2024-06-17
  12. RewardBench: Evaluating Reward Models for Language Modeling 14 upvotes, #9 of 2024-03-21

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.