Daily Papers of 2024-06-19

  1. DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence 53 upvotes, #1 of 2024-06-19
  2. Depth Anywhere: Enhancing 360 Monocular Depth Estimation via Perspective Distillation and Unlabeled Data Augmentation 45 upvotes, #2 of 2024-06-19
  3. Bootstrapping Language Models with DPO Implicit Rewards 34 upvotes, #3 of 2024-06-19
  4. TroL: Traversal of Layers for Large Language and Vision Models 32 upvotes, #4 of 2024-06-19
  5. VoCo-LLaMA: Towards Vision Compression with Large Language Models 28 upvotes, #5 of 2024-06-19
  6. ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools 26 upvotes, #6 of 2024-06-19
  7. AgileCoder: Dynamic Collaborative Agents for Software Development based on Agile Methodology 25 upvotes, #7 of 2024-06-19
  8. From RAGs to rich parameters: Probing how language models utilize external knowledge over parametric information for factual queries 19 upvotes, #8 of 2024-06-19
  9. Safety Arithmetic: A Framework for Test-time Safety Alignment of Language Models by Steering Parameters and Activations 15 upvotes, #9 of 2024-06-19
  10. RepLiQA: A Question-Answering Dataset for Benchmarking LLMs on Unseen Reference Content 14 upvotes, #10 of 2024-06-19
  11. OlympicArena: Benchmarking Multi-discipline Cognitive Reasoning for Superintelligent AI 14 upvotes, #10 of 2024-06-19
  12. Tokenization Falling Short: The Curse of Tokenization 13 upvotes, #12 of 2024-06-19
  13. SafeInfer: Context Adaptive Decoding Time Safety Alignment for Large Language Models 13 upvotes, #12 of 2024-06-19
  14. Benchmarking Multi-Image Understanding in Vision and Language Models: Perception, Knowledge, Reasoning, and Multi-Hop Reasoning 13 upvotes, #12 of 2024-06-19
  15. HumanSplat: Generalizable Single-Image Human Gaussian Splatting with Structure Priors 11 upvotes, #15 of 2024-06-19
  16. Learn Beyond The Answer: Training Language Models with Reflection for Mathematical Reasoning 10 upvotes, #16 of 2024-06-19
  17. Not All Prompts Are Made Equal: Prompt-based Pruning of Text-to-Image Diffusion Models 7 upvotes, #17 of 2024-06-19
  18. Language Models are Surprisingly Fragile to Drug Names in Biomedical Benchmarks 7 upvotes, #17 of 2024-06-19
  19. BPO: Supercharging Online Preference Learning by Adhering to the Proximity of Behavior LLM 7 upvotes, #17 of 2024-06-19
  20. Mixture of Scales: Memory-Efficient Token-Adaptive Binarization for Large Language Models 7 upvotes, #17 of 2024-06-19
  21. Large Scale Transfer Learning for Tabular Data via Language Modeling 6 upvotes, #21 of 2024-06-19
  22. Estimating Knowledge in Large Language Models Without Generating a Single Token 6 upvotes, #21 of 2024-06-19
  23. From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline 5 upvotes, #23 of 2024-06-19
  24. Super(ficial)-alignment: Strong Models May Deceive Weak Models in Weak-to-Strong Generalization 4 upvotes, #24 of 2024-06-19
  25. JEN-1 DreamStyler: Customized Musical Concept Learning via Pivotal Parameters Tuning 4 upvotes, #24 of 2024-06-19
  26. Immiscible Diffusion: Accelerating Diffusion Training with Noise Assignment 4 upvotes, #24 of 2024-06-19
  27. Hierarchical Prompting Taxonomy: A Universal Evaluation Framework for Large Language Models 4 upvotes, #24 of 2024-06-19
  28. Adversarial Attacks on Multimodal Agents 4 upvotes, #24 of 2024-06-19
  29. VIA: A Spatiotemporal Video Adaptation Framework for Global and Local Video Editing 4 upvotes, #24 of 2024-06-19
  30. Mixture-of-Subspaces in Low-Rank Adaptation 3 upvotes, #30 of 2024-06-19

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.