Chuanhao

Chuanhao on Hugging Face Daily Papers: 26 papers, 9 in the top 3 of their day, 1,003 upvotes.

  1. AlayaVista: Streaming World Modeling from Panoramic States to Perspective Video 26 upvotes, #11 of 2026-09-15
  2. Marionette: Predicting World States, Rendering Geometry, Painting Appearance 33 upvotes, #6 of 2026-08-17
  3. Alaya-EVOKE: From Linear-Scaling Supervision to Endless World 132 upvotes, #1 of 2026-08-14
  4. AlayaWorld: Interactive Long-Horizon World Modeling -- Full Technical Report 57 upvotes, #6 of 2026-07-22
  5. From Pixels to States: Rethinking Interactive World Models as Game Engines 35 upvotes, #7 of 2026-07-17
  6. AlayaWorld: Long-Horizon and Playable Video World Generation 87 upvotes, #2 of 2026-07-08
  7. AgenticSTS: A Bounded-Memory Testbed for Long-Horizon LLM Agents 60 upvotes, #2 of 2026-07-03
  8. JAMER: Project-Level Code Framework Dataset and Benchmark on Professional Game Engines 3 upvotes, #30 of 2026-06-19
  9. PackForcing: Short Video Training Suffices for Long Video Sampling and Long Context Inference 50 upvotes, #3 of 2026-03-30
  10. WildWorld: A Large-Scale Dataset for Dynamic World Modeling with Actions and Explicit State toward Generative ARPG 90 upvotes, #2 of 2026-03-25
  11. LongCLI-Bench: A Preliminary Benchmark and Study for Long-horizon Agentic Programming in Command-Line Interfaces 12 upvotes, #10 of 2026-02-25
  12. World Craft: Agentic Framework to Create Visualizable Worlds via Text 20 upvotes, #9 of 2026-01-28
  13. MeepleLM: A Virtual Playtester Simulating Diverse Subjective Experiences 11 upvotes, #12 of 2026-01-26
  14. Yume-1.5: A Text-Controlled Interactive World Generation Model 57 upvotes, #3 of 2025-12-30
  15. SVBench: Evaluation of Video Generation Models on Social Reasoning 7 upvotes, #13 of 2025-12-29
  16. InMind: Evaluating LLMs in Capturing and Applying Individual Human Reasoning Styles 2 upvotes, #14 of 2025-08-25
  17. MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents 16 upvotes, #5 of 2025-08-20
  18. Yume: An Interactive World Generation Model 77 upvotes, #1 of 2025-07-24
  19. Sekai: A Video Dataset towards World Exploration 60 upvotes, #1 of 2025-06-19
  20. A High-Quality Dataset and Reliable Evaluation for Interleaved Image-Text Generation 7 upvotes, #14 of 2025-06-16
  21. SridBench: Benchmark of Scientific Research Illustration Drawing of Image Generation Model 4 upvotes, #49 of 2025-05-30
  22. IA-T2I: Internet-Augmented Text-to-Image Generation 15 upvotes, #16 of 2025-05-22
  23. MDK12-Bench: A Multi-Discipline Benchmark for Evaluating Reasoning in Multimodal Large Language Models 4 upvotes, #24 of 2025-04-15
  24. ARMOR v0.1: Empowering Autoregressive Multimodal Understanding Model with Interleaved Multimodal Generation via Asymmetric Synergy 8 upvotes, #14 of 2025-03-17
  25. GATE OpenING: A Comprehensive Benchmark for Judging Open-ended Interleaved Image-Text Generation 17 upvotes, #9 of 2024-12-03
  26. MMIU: Multimodal Multi-image Understanding for Evaluating Large Vision-Language Models 56 upvotes, #1 of 2024-08-07

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.