The University of Hong Kong
The University of Hong Kong on Hugging Face Daily Papers: 32 papers, 1 in the top 3 of their day, 0 paper of the day.
- Understanding On-Policy Distillation: A Mechanistic Interpretability Perspective via Sparse Crosscoders 5 upvotes, #74 of 2026-09-30
- TempCloze: Can Video-LLMs Identify the Missing Middle? 30 upvotes, #15 of 2026-09-11
- SceneMosaic: Efficient and Diverse Simulation-Ready Scene Generation via Hybrid Agentic Layout Evolution 47 upvotes, #11 of 2026-09-09
- PlayWorld: Benchmarking World Models with Agent Players over Long-Horizon Objectives 45 upvotes, #9 of 2026-08-14
- UniClawBench: A Universal Benchmark for Proactive Agents on Real-World Tasks 34 upvotes, #6 of 2026-07-10
- Decoupled Residual Denoising Diffusion Models for Unified and Data Efficient Image-to-Image Translation 14 upvotes, #19 of 2026-06-03
- OScaR: The Occam's Razor for Extreme KV Cache Quantization in LLMs and Beyond 39 upvotes, #8 of 2026-05-21
- DexHoldem: Playing Texas Hold'em with Dexterous Embodied System 7 upvotes, #30 of 2026-05-19
- From Pixels to Concepts: Do Segmentation Models Understand What They Segment? 2 upvotes, #42 of 2026-05-14
- SAVOIR: Learning Social Savoir-Faire via Shapley-based Reward Attribution 4 upvotes, #21 of 2026-04-23
- Meta-learning In-Context Enables Training-Free Cross Subject Brain Decoding 9 upvotes, #18 of 2026-04-21
- MultiWorld: Scalable Multi-Agent Multi-View Video World Models 43 upvotes, #5 of 2026-04-21
- Stratagem: Learning Transferable Reasoning via Trajectory-Modulated Game Self-Play 6 upvotes, #23 of 2026-04-21
- ImplicitMemBench: Measuring Unconscious Behavioral Adaptation in Large Language Models 8 upvotes, #28 of 2026-04-10
- Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms 30 upvotes, #10 of 2026-03-31
- MACRO: Advancing Multi-Reference Image Generation with Structured Long-Context Data 32 upvotes, #7 of 2026-03-27
- FASTER: Rethinking Real-Time Flow VLAs 55 upvotes, #5 of 2026-03-20
- Cubic Discrete Diffusion: Discrete Visual Generation on High-Dimensional Representation Tokens 33 upvotes, #9 of 2026-03-20
- EVATok: Adaptive Length Video Tokenization for Efficient Visual Autoregressive Generation 13 upvotes, #15 of 2026-03-13
- Surgical Post-Training: Cutting Errors, Keeping Knowledge 11 upvotes, #13 of 2026-03-04
- The Art of Efficient Reasoning: Data, Reward, and Optimization 6 upvotes, #15 of 2026-02-25
- AssetFormer: Modular 3D Assets Generation with Autoregressive Transformer 2 upvotes, #20 of 2026-02-24
- Unveiling Implicit Advantage Symmetry: Why GRPO Struggles with Exploration and Difficulty Adaptation 12 upvotes, #19 of 2026-02-13
- χ_{0}: Resource-Aware Robust Manipulation via Taming Distributional Inconsistencies 25 upvotes, #13 of 2026-02-13
- Dream-VL & Dream-VLA: Open Vision-Language and Vision-Language-Action Models with Diffusion Language Model Backbone 43 upvotes, #6 of 2025-12-30
- Learning to Reason in 4D: Dynamic Spatial Understanding for Vision Language Models 48 upvotes, #2 of 2025-12-25
- VideoSSM: Autoregressive Long Video Generation with Hybrid State-Space Memory 3 upvotes, #16 of 2025-12-11
- OmniX: From Unified Panoramic Generation and Perception to Graphics-Ready 3D Scenes 21 upvotes, #12 of 2025-10-31
- Revisiting Model Interpolation for Efficient Reasoning 8 upvotes, #22 of 2025-10-16
- SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Models 19 upvotes, #14 of 2025-10-15
- CodePlot-CoT: Mathematical Visual Reasoning by Thinking with Code-Driven Images 13 upvotes, #24 of 2025-10-14
- Compose Your Policies! Improving Diffusion-based or Flow-based Robot Policies via Test-time Distribution-level Composition 19 upvotes, #7 of 2025-10-06
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.