Daily Papers of 2024-09-23

  1. Imagine yourself: Tuning-Free Personalized Image Generation 64 upvotes, #1 of 2024-09-23
  2. YesBut: A High-Quality Annotated Multimodal Dataset for evaluating Satire Comprehension capability of Vision-Language Models 45 upvotes, #2 of 2024-09-23
  3. Prithvi WxC: Foundation Model for Weather and Climate 32 upvotes, #3 of 2024-09-23
  4. MuCodec: Ultra Low-Bitrate Music Codec 20 upvotes, #4 of 2024-09-23
  5. Fact, Fetch, and Reason: A Unified Evaluation of Retrieval-Augmented Generation 18 upvotes, #5 of 2024-09-23
  6. Portrait Video Editing Empowered by Multimodal Generative Priors 15 upvotes, #6 of 2024-09-23
  7. Colorful Diffuse Intrinsic Image Decomposition in the Wild 12 upvotes, #7 of 2024-09-23
  8. V^3: Viewing Volumetric Videos on Mobiles via Streamable 2D Dynamic Gaussians 9 upvotes, #8 of 2024-09-23
  9. Minstrel: Structural Prompt Generation with Multi-Agents Coordination for Non-AI Experts 7 upvotes, #9 of 2024-09-23
  10. Temporally Aligned Audio for Video with Autoregression 7 upvotes, #9 of 2024-09-23
  11. Hackphyr: A Local Fine-Tuned LLM Agent for Network Security Environments 6 upvotes, #11 of 2024-09-23
  12. LLM-Agent-UMF: LLM-based Agent Unified Modeling Framework for Seamless Integration of Multi Active/Passive Core-Agents 2 upvotes, #12 of 2024-09-23

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.