Umberto Cappellazzo

Umberto Cappellazzo on Hugging Face Daily Papers: 7 papers, 0 in the top 3 of their day, 26 upvotes.

  1. Listening Forward: Next Patch Embedding Prediction Enables Scalable Audio Learners 6 upvotes, #25 of 2026-08-21
  2. Dr. SHAP-AV: Decoding Relative Modality Contributions via Shapley Attribution in Audio-Visual Speech Recognition 3 upvotes, #35 of 2026-03-13
  3. Omni-AVSR: Towards Unified Multimodal Speech Recognition with Large Language Models 2 upvotes, #31 of 2025-11-11
  4. Mitigating Attention Sinks and Massive Activations in Audio-Visual Speech Recognition with LLMS 2 upvotes, #32 of 2025-10-28
  5. MoME: Mixture of Matryoshka Experts for Audio-Visual Speech Recognition 3 upvotes, #27 of 2025-10-07
  6. Scaling and Enhancing LLM-based AVSR: A Sparse Mixture of Projectors Approach 3 upvotes, #38 of 2025-05-22
  7. Adaptive Audio-Visual Speech Recognition via Matryoshka-Based Multimodal LLMs 2 upvotes, #38 of 2025-03-11

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.