Daily Papers of 2024-09-18
- OmniGen: Unified Image Generation 76 upvotes, #1 of 2024-09-18
- NVLM: Open Frontier-Class Multimodal LLMs 54 upvotes, #2 of 2024-09-18
- Fine-Tuning Image-Conditional Diffusion Models is Easier than You Think 25 upvotes, #3 of 2024-09-18
- Phidias: A Generative Model for Creating 3D Content from Text, Image, and 3D Conditions with Reference-Augmented Diffusion 22 upvotes, #4 of 2024-09-18
- Promptriever: Instruction-Trained Retrievers Can Be Prompted Like Language Models 20 upvotes, #5 of 2024-09-18
- EzAudio: Enhancing Text-to-Audio Generation with Efficient Diffusion Transformer 15 upvotes, #6 of 2024-09-18
- A Comprehensive Evaluation of Quantized Instruction-Tuned Large Language Models: An Experimental Analysis up to 405B 15 upvotes, #6 of 2024-09-18
- On the limits of agency in agent-based models 12 upvotes, #8 of 2024-09-18
- OSV: One Step is Enough for High-Quality Image to Video Generation 12 upvotes, #8 of 2024-09-18
- Agile Continuous Jumping in Discontinuous Terrains 10 upvotes, #10 of 2024-09-18
- SplatFields: Neural Gaussian Splats for Sparse 3D and 4D Reconstruction 7 upvotes, #11 of 2024-09-18
- Implicit Neural Representations with Fourier Kolmogorov-Arnold Networks 4 upvotes, #12 of 2024-09-18
- Single-Layer Learnable Activation for Implicit Neural Representation (SL^{2}A-INR) 4 upvotes, #12 of 2024-09-18
- Measuring and Enhancing Trustworthiness of LLMs in RAG through Grounded Attributions and Learning to Refuse 4 upvotes, #12 of 2024-09-18
- Human-like Affective Cognition in Foundation Models 4 upvotes, #12 of 2024-09-18
- PDMX: A Large-Scale Public Domain MusicXML Dataset for Symbolic Music Processing 3 upvotes, #16 of 2024-09-18
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.