Daily Papers of 2024-09-18

  1. OmniGen: Unified Image Generation 76 upvotes, #1 of 2024-09-18
  2. NVLM: Open Frontier-Class Multimodal LLMs 54 upvotes, #2 of 2024-09-18
  3. Fine-Tuning Image-Conditional Diffusion Models is Easier than You Think 25 upvotes, #3 of 2024-09-18
  4. Phidias: A Generative Model for Creating 3D Content from Text, Image, and 3D Conditions with Reference-Augmented Diffusion 22 upvotes, #4 of 2024-09-18
  5. Promptriever: Instruction-Trained Retrievers Can Be Prompted Like Language Models 20 upvotes, #5 of 2024-09-18
  6. EzAudio: Enhancing Text-to-Audio Generation with Efficient Diffusion Transformer 15 upvotes, #6 of 2024-09-18
  7. A Comprehensive Evaluation of Quantized Instruction-Tuned Large Language Models: An Experimental Analysis up to 405B 15 upvotes, #6 of 2024-09-18
  8. On the limits of agency in agent-based models 12 upvotes, #8 of 2024-09-18
  9. OSV: One Step is Enough for High-Quality Image to Video Generation 12 upvotes, #8 of 2024-09-18
  10. Agile Continuous Jumping in Discontinuous Terrains 10 upvotes, #10 of 2024-09-18
  11. SplatFields: Neural Gaussian Splats for Sparse 3D and 4D Reconstruction 7 upvotes, #11 of 2024-09-18
  12. Implicit Neural Representations with Fourier Kolmogorov-Arnold Networks 4 upvotes, #12 of 2024-09-18
  13. Single-Layer Learnable Activation for Implicit Neural Representation (SL^{2}A-INR) 4 upvotes, #12 of 2024-09-18
  14. Measuring and Enhancing Trustworthiness of LLMs in RAG through Grounded Attributions and Learning to Refuse 4 upvotes, #12 of 2024-09-18
  15. Human-like Affective Cognition in Foundation Models 4 upvotes, #12 of 2024-09-18
  16. PDMX: A Large-Scale Public Domain MusicXML Dataset for Symbolic Music Processing 3 upvotes, #16 of 2024-09-18

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.