Yuan Wu
Yuan Wu on Hugging Face Daily Papers: 7 papers, 1 in the top 3 of their day, 164 upvotes.
- More Choices, Fewer Decisions: Ordinal-Scale Bias in JEV-like Direct-Decision Models 54 upvotes, #16 of 2026-10-01
- CaSKG: Counterfactual-Causal Skill Graphs for Scalable Agent Skill Retrieval 5 upvotes, #18 of 2026-08-28
- Can Large Multimodal Models Actively Recognize Faulty Inputs? A Systematic Evaluation Framework of Their Input Scrutiny Ability 10 upvotes, #11 of 2025-08-08
- StructFlowBench: A Structured Flow Benchmark for Multi-turn Instruction Following 13 upvotes, #10 of 2025-02-24
- Large Language Model Evaluation via Matrix Nuclear-Norm 18 upvotes, #7 of 2024-10-17
- Rethinking Data Selection at Scale: Random Selection is Almost All You Need 14 upvotes, #11 of 2024-10-15
- A Survey on Evaluation of Large Language Models 43 upvotes, #2 of 2023-07-07
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.