Daily Papers of 2024-07-04

  1. InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output 84 upvotes, #1 of 2024-07-04
  2. TabReD: A Benchmark of Tabular Machine Learning in-the-Wild 44 upvotes, #2 of 2024-07-04
  3. TokenPacker: Efficient Visual Projector for Multimodal LLM 18 upvotes, #3 of 2024-07-04
  4. No Training, No Problem: Rethinking Classifier-Free Guidance for Diffusion Models 18 upvotes, #3 of 2024-07-04
  5. PicoAudio: Enabling Precise Timestamp and Frequency Controllability of Audio Events in Text-to-audio Generation 14 upvotes, #5 of 2024-07-04
  6. DisCo-Diff: Enhancing Continuous Diffusion Models with Discrete Latents 10 upvotes, #6 of 2024-07-04
  7. Investigating Decoder-only Large Language Models for Speech-to-text Translation 9 upvotes, #7 of 2024-07-04
  8. A False Sense of Safety: Unsafe Information Leakage in 'Safe' AI Responses 7 upvotes, #8 of 2024-07-04
  9. Eliminating Position Bias of Language Models: A Mechanistic Approach 6 upvotes, #9 of 2024-07-04

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.