Wenbo Hu

Wenbo Hu on Hugging Face Daily Papers: 5 papers, 0 in the top 3 of their day, 104 upvotes.

  1. OpenVLThinkerV2: A Generalist Multimodal Reasoning Model for Multi-domain Visual Tasks 48 upvotes, #11 of 2026-04-10
  2. MMSI-Video-Bench: A Holistic Benchmark for Video-Based Spatial Intelligence 21 upvotes, #8 of 2025-12-18
  3. G^2VLM: Geometry Grounded Vision Language Model with Unified 3D Reconstruction and Spatial Reasoning 8 upvotes, #10 of 2025-11-27
  4. ARES: Multimodal Adaptive Reasoning via Difficulty-Aware Token-Level Entropy Shaping 12 upvotes, #16 of 2025-10-13
  5. Interleaving Reasoning for Better Text-to-Image Generation 13 upvotes, #10 of 2025-09-09

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.