Guo Chen
Guo Chen on Hugging Face Daily Papers: 7 papers, 4 in the top 3 of their day, 261 upvotes.
- Bridging Perspectives: A Survey on Cross-view Collaborative Intelligence with Egocentric-Exocentric Vision 8 upvotes, #13 of 2025-06-09
- AV-Reasoner: Improving and Benchmarking Clue-Grounded Audio-Visual Counting for MLLMs 20 upvotes, #13 of 2025-06-06
- Eagle 2.5: Boosting Long-Context Post-Training for Frontier Vision-Language Models 65 upvotes, #2 of 2025-04-22
- Token-Efficient Long Video Understanding for Multimodal LLMs 79 upvotes, #2 of 2025-03-07
- InternVideo2: Scaling Video Foundation Models for Multimodal Video Understanding 14 upvotes, #3 of 2024-03-25
- Video Mamba Suite: State Space Model as a Versatile Alternative for Video Understanding 10 upvotes, #9 of 2024-03-15
- InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks 22 upvotes, #2 of 2023-12-26
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.