Han Zhao

Han Zhao on Hugging Face Daily Papers: 8 papers, 3 in the top 3 of their day, 488 upvotes.

  1. FRAPPE: Infusing World Modeling into Generalist Policies via Multiple Future Representation Alignment 5 upvotes, #16 of 2026-02-20
  2. VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation 13 upvotes, #18 of 2025-10-17
  3. Spatial Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model 139 upvotes, #1 of 2025-10-15
  4. VLA-Adapter: An Effective Paradigm for Tiny-Scale Vision-Language-Action Model 189 upvotes, #1 of 2025-09-12
  5. SSR: Enhancing Depth Perception in Vision-Language Models via Rationale-Guided Spatial Reasoning 10 upvotes, #20 of 2025-05-21
  6. OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation 8 upvotes, #12 of 2025-05-08
  7. PiTe: Pixel-Temporal Alignment for Large Video-Language Model 11 upvotes, #7 of 2024-09-13
  8. Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference 30 upvotes, #3 of 2024-03-22

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.