Song XiXuan

Song XiXuan on Hugging Face Daily Papers: 3 papers, 1 in the top 3 of their day, 101 upvotes.

  1. CogVLM2: Visual Language Models for Image and Video Understanding 55 upvotes, #2 of 2024-08-30
  2. VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents 12 upvotes, #7 of 2024-08-13
  3. CogVLM: Visual Expert for Pretrained Language Models 27 upvotes, #4 of 2023-11-07

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.