Daily Papers of 2024-09-23
- Imagine yourself: Tuning-Free Personalized Image Generation 64 upvotes, #1 of 2024-09-23
- YesBut: A High-Quality Annotated Multimodal Dataset for evaluating Satire Comprehension capability of Vision-Language Models 45 upvotes, #2 of 2024-09-23
- Prithvi WxC: Foundation Model for Weather and Climate 32 upvotes, #3 of 2024-09-23
- MuCodec: Ultra Low-Bitrate Music Codec 20 upvotes, #4 of 2024-09-23
- Fact, Fetch, and Reason: A Unified Evaluation of Retrieval-Augmented Generation 18 upvotes, #5 of 2024-09-23
- Portrait Video Editing Empowered by Multimodal Generative Priors 15 upvotes, #6 of 2024-09-23
- Colorful Diffuse Intrinsic Image Decomposition in the Wild 12 upvotes, #7 of 2024-09-23
- V^3: Viewing Volumetric Videos on Mobiles via Streamable 2D Dynamic Gaussians 9 upvotes, #8 of 2024-09-23
- Minstrel: Structural Prompt Generation with Multi-Agents Coordination for Non-AI Experts 7 upvotes, #9 of 2024-09-23
- Temporally Aligned Audio for Video with Autoregression 7 upvotes, #9 of 2024-09-23
- Hackphyr: A Local Fine-Tuned LLM Agent for Network Security Environments 6 upvotes, #11 of 2024-09-23
- LLM-Agent-UMF: LLM-based Agent Unified Modeling Framework for Seamless Integration of Multi Active/Passive Core-Agents 2 upvotes, #12 of 2024-09-23
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.