Huck Yang

Huck Yang on Hugging Face Daily Papers: 13 papers, 2 in the top 3 of their day, 251 upvotes.

  1. Voice Memory for Agentic Speech Recognition 11 upvotes, #16 of 2026-07-30
  2. Nemotron 3 Nano Omni: Efficient and Open Multimodal Intelligence 17 upvotes, #15 of 2026-05-01
  3. Audio Flamingo Next: Next-Generation Open Audio-Language Models for Speech, Sound, and Music 28 upvotes, #12 of 2026-04-14
  4. How Auditory Knowledge in LLM Backbones Shapes Audio Language Models: A Holistic Evaluation 4 upvotes, #36 of 2026-04-01
  5. PRiSM: Benchmarking Phone Realization in Speech Models 6 upvotes, #20 of 2026-01-21
  6. Long Grounded Thoughts: Distilling Compositional Visual Reasoning Chains at Scale 6 upvotes, #20 of 2025-11-11
  7. SAKE: Towards Editing Auditory Attribute Knowledge of Large Audio-Language Models 19 upvotes, #7 of 2025-10-24
  8. Investigating Safety Vulnerabilities of Large Audio-Language Models Under Speaker Emotional Variations 17 upvotes, #9 of 2025-10-24
  9. OmniVinci: Enhancing Architecture and Data for Omni-Modal Understanding LLM 79 upvotes, #2 of 2025-10-20
  10. Estimating Time Series Foundation Model Transferability via In-Context Learning 1 upvotes, #50 of 2025-10-01
  11. Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models 9 upvotes, #12 of 2025-07-14
  12. NeKo: Toward Post Recognition Generative Correction Large Language Models with Task-Oriented Experts 4 upvotes, #13 of 2024-11-12
  13. Low-rank Adaptation of Large Language Model Rescoring for Parameter-Efficient Speech Recognition 23 upvotes, #3 of 2023-09-28

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.