Daily Papers of 2023-11-02
- Distil-Whisper: Robust Knowledge Distillation via Large-Scale Pseudo Labelling 56 upvotes, #1 of 2023-11-02
- LLaVA-Interactive: An All-in-One Demo for Image Chat, Segmentation, Generation and Editing 42 upvotes, #2 of 2023-11-02
- Controllable Music Production with Diffusion Models and Guidance Gradients 24 upvotes, #3 of 2023-11-02
- De-Diffusion Makes Text a Strong Cross-Modal Interface 21 upvotes, #4 of 2023-11-02
- The Generative AI Paradox: "What It Can Create, It May Not Understand" 18 upvotes, #5 of 2023-11-02
- Text Rendering Strategies for Pixel Language Models 11 upvotes, #6 of 2023-11-02
- ChatCoder: Chat-based Refine Requirement Improves LLMs' Code Generation 10 upvotes, #7 of 2023-11-02
- Grounding Visual Illusions in Language: Do Vision-Language Models Perceive Illusions Like Humans? 9 upvotes, #8 of 2023-11-02
- ChipNeMo: Domain-Adapted LLMs for Chip Design 9 upvotes, #8 of 2023-11-02
- AMSP: Super-Scaling LLM Training via Advanced Model States Partitioning 9 upvotes, #8 of 2023-11-02
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.