li haodong
li haodong on Hugging Face Daily Papers: 12 papers, 3 in the top 3 of their day, 397 upvotes.
- PerceptionRubrics: Calibrating Multimodal Evaluation to Human Perception 41 upvotes, #1 of 2026-07-02
- Auto-Rubric as Reward: From Implicit Preferences to Explicit Multimodal Generative Criteria 23 upvotes, #13 of 2026-05-12
- SpatialEvo: Self-Evolving Spatial Intelligence via Deterministic Geometric Environments 62 upvotes, #3 of 2026-04-16
- PEARL: Personalized Streaming Video Understanding Model 40 upvotes, #6 of 2026-03-25
- WebVR: Benchmarking Multimodal LLMs for WebPage Recreation from Videos via Human-Aligned Visual Rubrics 19 upvotes, #14 of 2026-03-17
- CoCo: Code as CoT for Text-to-Image Preview and Rare Concept Generation 36 upvotes, #6 of 2026-03-10
- Proact-VL: A Proactive VideoLLM for Real-Time AI Companions 31 upvotes, #4 of 2026-03-05
- GENIUS: Generative Fluid Intelligence Evaluation Suite 53 upvotes, #3 of 2026-02-12
- Chain of Mindset: Reasoning with Adaptive Cognitive Modes 70 upvotes, #4 of 2026-02-11
- GEBench: Benchmarking Image Generation Models as GUI Environments 38 upvotes, #14 of 2026-02-10
- How Well Do Models Follow Visual Instructions? VIBE: A Systematic Benchmark for Visual Instruction-Driven Image Editing 16 upvotes, #27 of 2026-02-03
- DraCo: Draft as CoT for Text-to-Image Preview and Rare Concept Generation 11 upvotes, #18 of 2025-12-05
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.