Daily Papers of 2023-06-16

  1. TryOnDiffusion: A Tale of Two UNets 75 upvotes, #1 of 2023-06-16
  2. WizardCoder: Empowering Code Large Language Models with Evol-Instruct 34 upvotes, #2 of 2023-06-16
  3. Seeing the World through Your Eyes 34 upvotes, #2 of 2023-06-16
  4. AssistGPT: A General Multi-modal Assistant that can Plan, Execute, Inspect, and Learn 27 upvotes, #4 of 2023-06-16
  5. Knowledge Distillation of Large Language Models 24 upvotes, #5 of 2023-06-16
  6. KoLA: Carefully Benchmarking World Knowledge of Large Language Models 20 upvotes, #6 of 2023-06-16
  7. h2oGPT: Democratizing Large Language Models 19 upvotes, #7 of 2023-06-16
  8. DreamHuman: Animatable 3D Avatars from Text 17 upvotes, #8 of 2023-06-16
  9. Macaw-LLM: Multi-Modal Language Modeling with Image, Audio, Video, and Text Integration 16 upvotes, #9 of 2023-06-16
  10. Language to Rewards for Robotic Skill Synthesis 13 upvotes, #10 of 2023-06-16
  11. ChessGPT: Bridging Policy Learning and Language Modeling 12 upvotes, #11 of 2023-06-16
  12. Anticipatory Music Transformer 10 upvotes, #12 of 2023-06-16
  13. Exploring the MIT Mathematics and EECS Curriculum Using Large Language Models 10 upvotes, #12 of 2023-06-16
  14. Diffusion Models for Zero-Shot Open-Vocabulary Segmentation 10 upvotes, #12 of 2023-06-16
  15. Language-Guided Music Recommendation for Video via Prompt Analogies 10 upvotes, #12 of 2023-06-16
  16. Agile Catching with Whole-Body MPC and Blackbox Policy Learning 9 upvotes, #16 of 2023-06-16
  17. LOVM: Language-Only Vision Model Selection 7 upvotes, #17 of 2023-06-16
  18. DORSal: Diffusion for Object-centric Representations of Scenes et al. 6 upvotes, #18 of 2023-06-16
  19. VidEdit: Zero-Shot and Spatially Aware Text-Driven Video Editing 6 upvotes, #18 of 2023-06-16
  20. AVIS: Autonomous Visual Information Seeking with Large Language Models 5 upvotes, #20 of 2023-06-16
  21. UrbanIR: Large-Scale Urban Scene Inverse Rendering from a Single Video 5 upvotes, #20 of 2023-06-16
  22. Large-scale Language Model Rescoring on Long-form Data 4 upvotes, #22 of 2023-06-16
  23. NAVI: Category-Agnostic Image Collections with High-Quality 3D Shape and Pose Annotations 4 upvotes, #22 of 2023-06-16
  24. Tune As You Scale: Hyperparameter Optimization For Compute Efficient Training 3 upvotes, #24 of 2023-06-16
  25. Toward Grounded Social Reasoning 3 upvotes, #24 of 2023-06-16
  26. Neural Relighting with Subsurface Scattering by Learning the Radiance Transfer Gradient 3 upvotes, #24 of 2023-06-16

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.