mzr1996
mzr1996 on Hugging Face Daily Papers: 7 papers, 3 in the top 3 of their day, 264 upvotes.
- SWE-Bench Pro Verified: A Reliable Benchmark for Software Engineering Agents 14 upvotes, #15 of 2026-09-10
- TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration 13 upvotes, #14 of 2026-04-16
- Intern-S1-Pro: Scientific Multimodal Foundation Model at Trillion Scale 125 upvotes, #1 of 2026-03-27
- DataChef: Cooking Up Optimal Data Recipes for LLM Adaptation via Reinforcement Learning 18 upvotes, #12 of 2026-02-12
- MIG: Automatic Data Selection for Instruction Tuning by Maximizing Information Gain in Semantic Space 36 upvotes, #3 of 2025-04-21
- GTA: A Benchmark for General Tool Agents 8 upvotes, #15 of 2024-07-12
- InternLM2 Technical Report 22 upvotes, #3 of 2024-03-27
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.