Bryan Catanzaro

Bryan Catanzaro on Hugging Face Daily Papers: 19 papers, 8 in the top 3 of their day, 590 upvotes.

  1. MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos 13 upvotes, #18 of 2026-03-17
  2. NEMOTRON-CROSSTHINK: Scaling Self-Learning beyond Math Reasoning 6 upvotes, #21 of 2025-04-22
  3. Eagle 2.5: Boosting Long-Context Post-Training for Frontier Vision-Language Models 65 upvotes, #2 of 2025-04-22
  4. Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities 22 upvotes, #6 of 2025-03-07
  5. AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling 12 upvotes, #9 of 2024-12-20
  6. Synthio: Augmenting Small-Scale Audio Classification Datasets with Synthetic Data 4 upvotes, #25 of 2024-10-04
  7. PHI-S: Distribution Balancing for Label-Free Multi-Teacher Distillation 31 upvotes, #2 of 2024-10-03
  8. NVLM: Open Frontier-Class Multimodal LLMs 54 upvotes, #2 of 2024-09-18
  9. Eagle: Exploring The Design Space for Multimodal LLMs with Mixture of Encoders 76 upvotes, #1 of 2024-08-29
  10. LLM Pruning and Distillation in Practice: The Minitron Approach 48 upvotes, #2 of 2024-08-22
  11. ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities 20 upvotes, #4 of 2024-07-22
  12. NV-Embed: Improved Techniques for Training LLMs as Generalist Embedding Models 13 upvotes, #6 of 2024-05-28
  13. Audio Dialogues: Dialogues dataset for audio and music understanding 12 upvotes, #11 of 2024-04-12
  14. Nemotron-4 15B Technical Report 46 upvotes, #2 of 2024-02-27
  15. ODIN: Disentangled Reward Mitigates Hacking in RLHF 14 upvotes, #9 of 2024-02-13
  16. ChatQA: Building GPT-4 Level Conversational QA Models 35 upvotes, #3 of 2024-01-19
  17. ChipNeMo: Domain-Adapted LLMs for Chip Design 9 upvotes, #8 of 2023-11-02
  18. RAVEN: In-Context Learning with Retrieval Augmented Encoder-Decoder Language Models 19 upvotes, #3 of 2023-08-16
  19. Preserve Your Own Correlation: A Noise Prior for Video Diffusion Models 1 upvotes, #16 of 2023-05-19

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.