Lj V. Miranda
Lj V. Miranda on Hugging Face Daily Papers: 12 papers, 2 in the top 3 of their day, 372 upvotes.
- Polyglot Teachers: Evaluating Language Models for Multilingual Synthetic Data Generation 2 upvotes, #40 of 2026-04-14
- Olmo 3 22 upvotes, #11 of 2025-12-17
- R3: Robust Rubric-Agnostic Reward Models 11 upvotes, #22 of 2025-05-20
- Crowdsource, Crawl, or Generate? Creating SEA-VL, a Multicultural Vision-Language Dataset for Southeast Asia 92 upvotes, #1 of 2025-03-12
- MMTEB: Massive Multilingual Text Embedding Benchmark 31 upvotes, #5 of 2025-02-20
- Bridging the Data Provenance Gap Across Text, Speech and Video 6 upvotes, #10 of 2024-12-25
- TÜLU 3: Pushing Frontiers in Open Language Model Post-Training 55 upvotes, #1 of 2024-11-25
- Hybrid Preferences: Learning to Route Instances for Human vs. AI Feedback 10 upvotes, #9 of 2024-10-28
- M-RewardBench: Evaluating Reward Models in Multilingual Settings 10 upvotes, #7 of 2024-10-24
- Consent in Crisis: The Rapid Decline of the AI Data Commons 9 upvotes, #11 of 2024-07-23
- SEACrowd: A Multilingual Multimodal Data Hub and Benchmark Suite for Southeast Asian Languages 24 upvotes, #7 of 2024-06-17
- RewardBench: Evaluating Reward Models for Language Modeling 14 upvotes, #9 of 2024-03-21
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.