SII - Wenqi Shao
SII - Wenqi Shao on Hugging Face Daily Papers: 22 papers, 6 in the top 3 of their day, 990 upvotes.
- RetroAgent: From Solving to Evolving via Retrospective Dual Intrinsic Feedback 12 upvotes, #12 of 2026-03-12
- MM-PRM: Enhancing Multimodal Mathematical Reasoning with Scalable Step-Level Supervision 25 upvotes, #12 of 2025-05-20
- CPGD: Toward Stable Rule-based Reinforcement Learning for Language Models 23 upvotes, #13 of 2025-05-20
- InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models 239 upvotes, #1 of 2025-04-15
- MPBench: A Comprehensive Multimodal Reasoning Benchmark for Process Errors Identification 9 upvotes, #16 of 2025-03-19
- PEBench: A Fictitious Dataset to Benchmark Machine Unlearning for Multimodal Large Language Models 5 upvotes, #22 of 2025-03-19
- MM-Eureka: Exploring Visual Aha Moment with Rule-based Large-scale Reinforcement Learning 53 upvotes, #3 of 2025-03-11
- Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling 103 upvotes, #1 of 2024-12-09
- GATE OpenING: A Comprehensive Benchmark for Judging Open-ended Interleaved Image-Text Generation 17 upvotes, #9 of 2024-12-03
- TP-Eval: Tap Multimodal LLMs' Potential in Evaluation by Customizing Prompts 6 upvotes, #9 of 2024-10-24
- PrefixQuant: Static Quantization Beats Dynamic through Prefixed Outliers in LLMs 29 upvotes, #4 of 2024-10-11
- Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation 42 upvotes, #5 of 2024-10-10
- MMIU: Multimodal Multi-image Understanding for Evaluating Large Vision-Language Models 56 upvotes, #1 of 2024-08-07
- Diffree: Text-Guided Shape Free Object Inpainting with Diffusion Model 39 upvotes, #1 of 2024-07-26
- GUI Odyssey: A Comprehensive Dataset for Cross-App GUI Navigation on Mobile Devices 21 upvotes, #8 of 2024-06-17
- Needle In A Multimodal Haystack 51 upvotes, #4 of 2024-06-17
- Adapting LLaMA Decoder to Vision Transformer 12 upvotes, #6 of 2024-04-11
- SPHINX-X: Scaling Data and Parameters for a Family of Multi-modal Large Language Models 17 upvotes, #8 of 2024-02-09
- Cached Transformers: Improving Transformers with Differentiable Memory Cache 13 upvotes, #7 of 2023-12-21
- ImageBind-LLM: Multi-modality Instruction Tuning 17 upvotes, #7 of 2023-09-08
- OmniQuant: Omnidirectionally Calibrated Quantization for Large Language Models 20 upvotes, #2 of 2023-08-28
- Tiny LVLM-eHub: Early Multimodal Experiments with Bard 11 upvotes, #10 of 2023-08-08
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.