Jingkang Yang

Jingkang Yang on Hugging Face Daily Papers: 12 papers, 3 in the top 3 of their day, 391 upvotes.

  1. A Simple Baseline for Streaming Video Understanding 72 upvotes, #4 of 2026-04-06
  2. HippoCamp: Benchmarking Contextual Agents on Personal Computers 27 upvotes, #9 of 2026-04-02
  3. Ego-R1: Chain-of-Tool-Thought for Ultra-Long Egocentric Video Reasoning 42 upvotes, #7 of 2025-06-17
  4. EgoLife: Towards Egocentric Life Assistant 35 upvotes, #4 of 2025-03-07
  5. Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models 19 upvotes, #7 of 2024-11-22
  6. Generalized Out-of-Distribution Detection and Beyond in Vision Language Model Era: A Survey 3 upvotes, #14 of 2024-08-02
  7. LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models 28 upvotes, #5 of 2024-07-18
  8. Long Context Transfer from Language to Vision 30 upvotes, #4 of 2024-06-25
  9. Unsolvable Problem Detection: Evaluating Trustworthiness of Vision Language Models 14 upvotes, #5 of 2024-04-01
  10. OtterHD: A High-Resolution Multi-modality Model 34 upvotes, #1 of 2023-11-08
  11. Octopus: Embodied Vision-Language Programmer from Environmental Feedback 37 upvotes, #2 of 2023-10-13
  12. MIMIC-IT: Multi-Modal In-Context Instruction Tuning 12 upvotes, #2 of 2023-06-09

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.