Daily Papers of 2023-06-02

  1. SQL-PaLM: Improved Large Language ModelAdaptation for Text-to-SQL 21 upvotes, #1 of 2023-06-02
  2. SnapFusion: Text-to-Image Diffusion Model on Mobile Devices within Two Seconds 16 upvotes, #2 of 2023-06-02
  3. LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day 14 upvotes, #3 of 2023-06-02
  4. Wuerstchen: Efficient Pretraining of Text-to-Image Models 13 upvotes, #4 of 2023-06-02
  5. Example-based Motion Synthesis via Generative Motion Matching 8 upvotes, #5 of 2023-06-02
  6. StyleDrop: Text-to-Image Generation in Any Style 7 upvotes, #6 of 2023-06-02
  7. MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training 6 upvotes, #7 of 2023-06-02
  8. Bytes Are All You Need: Transformers Operating Directly On File Bytes 6 upvotes, #7 of 2023-06-02
  9. Make-Your-Video: Customized Video Generation Using Textual and Structural Guidance 6 upvotes, #7 of 2023-06-02
  10. The Hidden Language of Diffusion Models 5 upvotes, #10 of 2023-06-02
  11. ViCo: Detail-Preserving Visual Condition for Personalized Text-to-Image Generation 4 upvotes, #11 of 2023-06-02
  12. StableRep: Synthetic Images from Text-to-Image Models Make Strong Visual Representation Learners 4 upvotes, #11 of 2023-06-02
  13. Inserting Anybody in Diffusion Models via Celeb Basis 3 upvotes, #13 of 2023-06-02
  14. CodeTF: One-stop Transformer Library for State-of-the-art Code LLM 2 upvotes, #14 of 2023-06-02
  15. MuseCoco: Generating Symbolic Music from Text 2 upvotes, #14 of 2023-06-02
  16. ReviewerGPT? An Exploratory Study on Using Large Language Models for Paper Reviewing 2 upvotes, #14 of 2023-06-02
  17. Birth of a Transformer: A Memory Viewpoint 2 upvotes, #14 of 2023-06-02
  18. Diffusion Self-Guidance for Controllable Image Generation 2 upvotes, #14 of 2023-06-02
  19. Brainformers: Trading Simplicity for Efficiency 1 upvotes, #19 of 2023-06-02
  20. SafeDiffuser: Safe Planning with Diffusion Probabilistic Models 1 upvotes, #19 of 2023-06-02
  21. The ObjectFolder Benchmark: Multisensory Learning with Neural and Real Objects 1 upvotes, #19 of 2023-06-02
  22. Cocktail: Mixing Multi-Modality Controls for Text-Conditional Image Generation 1 upvotes, #19 of 2023-06-02

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.