Han Zhao
Han Zhao on Hugging Face Daily Papers: 8 papers, 3 in the top 3 of their day, 488 upvotes.
- FRAPPE: Infusing World Modeling into Generalist Policies via Multiple Future Representation Alignment 5 upvotes, #16 of 2026-02-20
- VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation 13 upvotes, #18 of 2025-10-17
- Spatial Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model 139 upvotes, #1 of 2025-10-15
- VLA-Adapter: An Effective Paradigm for Tiny-Scale Vision-Language-Action Model 189 upvotes, #1 of 2025-09-12
- SSR: Enhancing Depth Perception in Vision-Language Models via Rationale-Guided Spatial Reasoning 10 upvotes, #20 of 2025-05-21
- OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation 8 upvotes, #12 of 2025-05-08
- PiTe: Pixel-Temporal Alignment for Large Video-Language Model 11 upvotes, #7 of 2024-09-13
- Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference 30 upvotes, #3 of 2024-03-22
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.