Jingkang Yang
Jingkang Yang on Hugging Face Daily Papers: 12 papers, 3 in the top 3 of their day, 391 upvotes.
- A Simple Baseline for Streaming Video Understanding 72 upvotes, #4 of 2026-04-06
- HippoCamp: Benchmarking Contextual Agents on Personal Computers 27 upvotes, #9 of 2026-04-02
- Ego-R1: Chain-of-Tool-Thought for Ultra-Long Egocentric Video Reasoning 42 upvotes, #7 of 2025-06-17
- EgoLife: Towards Egocentric Life Assistant 35 upvotes, #4 of 2025-03-07
- Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models 19 upvotes, #7 of 2024-11-22
- Generalized Out-of-Distribution Detection and Beyond in Vision Language Model Era: A Survey 3 upvotes, #14 of 2024-08-02
- LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models 28 upvotes, #5 of 2024-07-18
- Long Context Transfer from Language to Vision 30 upvotes, #4 of 2024-06-25
- Unsolvable Problem Detection: Evaluating Trustworthiness of Vision Language Models 14 upvotes, #5 of 2024-04-01
- OtterHD: A High-Resolution Multi-modality Model 34 upvotes, #1 of 2023-11-08
- Octopus: Embodied Vision-Language Programmer from Environmental Feedback 37 upvotes, #2 of 2023-10-13
- MIMIC-IT: Multi-Modal In-Context Instruction Tuning 12 upvotes, #2 of 2023-06-09
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.