Daily Papers of 2023-11-02

  1. Distil-Whisper: Robust Knowledge Distillation via Large-Scale Pseudo Labelling 56 upvotes, #1 of 2023-11-02
  2. LLaVA-Interactive: An All-in-One Demo for Image Chat, Segmentation, Generation and Editing 42 upvotes, #2 of 2023-11-02
  3. Controllable Music Production with Diffusion Models and Guidance Gradients 24 upvotes, #3 of 2023-11-02
  4. De-Diffusion Makes Text a Strong Cross-Modal Interface 21 upvotes, #4 of 2023-11-02
  5. The Generative AI Paradox: "What It Can Create, It May Not Understand" 18 upvotes, #5 of 2023-11-02
  6. Text Rendering Strategies for Pixel Language Models 11 upvotes, #6 of 2023-11-02
  7. ChatCoder: Chat-based Refine Requirement Improves LLMs' Code Generation 10 upvotes, #7 of 2023-11-02
  8. Grounding Visual Illusions in Language: Do Vision-Language Models Perceive Illusions Like Humans? 9 upvotes, #8 of 2023-11-02
  9. ChipNeMo: Domain-Adapted LLMs for Chip Design 9 upvotes, #8 of 2023-11-02
  10. AMSP: Super-Scaling LLM Training via Advanced Model States Partitioning 9 upvotes, #8 of 2023-11-02

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.