Chuanhao
Chuanhao on Hugging Face Daily Papers: 26 papers, 9 in the top 3 of their day, 1,003 upvotes.
- AlayaVista: Streaming World Modeling from Panoramic States to Perspective Video 26 upvotes, #11 of 2026-09-15
- Marionette: Predicting World States, Rendering Geometry, Painting Appearance 33 upvotes, #6 of 2026-08-17
- Alaya-EVOKE: From Linear-Scaling Supervision to Endless World 132 upvotes, #1 of 2026-08-14
- AlayaWorld: Interactive Long-Horizon World Modeling -- Full Technical Report 57 upvotes, #6 of 2026-07-22
- From Pixels to States: Rethinking Interactive World Models as Game Engines 35 upvotes, #7 of 2026-07-17
- AlayaWorld: Long-Horizon and Playable Video World Generation 87 upvotes, #2 of 2026-07-08
- AgenticSTS: A Bounded-Memory Testbed for Long-Horizon LLM Agents 60 upvotes, #2 of 2026-07-03
- JAMER: Project-Level Code Framework Dataset and Benchmark on Professional Game Engines 3 upvotes, #30 of 2026-06-19
- PackForcing: Short Video Training Suffices for Long Video Sampling and Long Context Inference 50 upvotes, #3 of 2026-03-30
- WildWorld: A Large-Scale Dataset for Dynamic World Modeling with Actions and Explicit State toward Generative ARPG 90 upvotes, #2 of 2026-03-25
- LongCLI-Bench: A Preliminary Benchmark and Study for Long-horizon Agentic Programming in Command-Line Interfaces 12 upvotes, #10 of 2026-02-25
- World Craft: Agentic Framework to Create Visualizable Worlds via Text 20 upvotes, #9 of 2026-01-28
- MeepleLM: A Virtual Playtester Simulating Diverse Subjective Experiences 11 upvotes, #12 of 2026-01-26
- Yume-1.5: A Text-Controlled Interactive World Generation Model 57 upvotes, #3 of 2025-12-30
- SVBench: Evaluation of Video Generation Models on Social Reasoning 7 upvotes, #13 of 2025-12-29
- InMind: Evaluating LLMs in Capturing and Applying Individual Human Reasoning Styles 2 upvotes, #14 of 2025-08-25
- MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents 16 upvotes, #5 of 2025-08-20
- Yume: An Interactive World Generation Model 77 upvotes, #1 of 2025-07-24
- Sekai: A Video Dataset towards World Exploration 60 upvotes, #1 of 2025-06-19
- A High-Quality Dataset and Reliable Evaluation for Interleaved Image-Text Generation 7 upvotes, #14 of 2025-06-16
- SridBench: Benchmark of Scientific Research Illustration Drawing of Image Generation Model 4 upvotes, #49 of 2025-05-30
- IA-T2I: Internet-Augmented Text-to-Image Generation 15 upvotes, #16 of 2025-05-22
- MDK12-Bench: A Multi-Discipline Benchmark for Evaluating Reasoning in Multimodal Large Language Models 4 upvotes, #24 of 2025-04-15
- ARMOR v0.1: Empowering Autoregressive Multimodal Understanding Model with Interleaved Multimodal Generation via Asymmetric Synergy 8 upvotes, #14 of 2025-03-17
- GATE OpenING: A Comprehensive Benchmark for Judging Open-ended Interleaved Image-Text Generation 17 upvotes, #9 of 2024-12-03
- MMIU: Multimodal Multi-image Understanding for Evaluating Large Vision-Language Models 56 upvotes, #1 of 2024-08-07
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.