Yuansheng Ni
Yuansheng Ni on Hugging Face Daily Papers: 11 papers, 2 in the top 3 of their day, 379 upvotes.
- Towards Retrieving Interaction Spaces for Agentic Search 4 upvotes, #27 of 2026-06-08
- VisCoder2: Building Multi-Language Visualization Coding Agents 20 upvotes, #14 of 2025-10-29
- VisCoder: Fine-Tuning LLMs for Executable Python Visualization Code Generation 22 upvotes, #10 of 2025-06-05
- PhyX: Does Your Model Have the "Wits" for Physical Reasoning? 47 upvotes, #7 of 2025-05-26
- SuperGPQA: Scaling LLM Evaluation across 285 Graduate Disciplines 92 upvotes, #3 of 2025-02-21
- MEGA-Bench: Scaling Multimodal Evaluation to over 500 Real-World Tasks 34 upvotes, #5 of 2024-10-15
- MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark 27 upvotes, #4 of 2024-09-05
- MantisScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation 13 upvotes, #8 of 2024-06-24
- GenAI Arena: An Open Evaluation Platform for Generative Models 18 upvotes, #4 of 2024-06-10
- MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark 35 upvotes, #1 of 2024-06-04
- A Comprehensive Study of Knowledge Editing for Large Language Models 19 upvotes, #6 of 2024-01-03
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.