Daily Papers of 2024-06-19
- DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence 53 upvotes, #1 of 2024-06-19
- Depth Anywhere: Enhancing 360 Monocular Depth Estimation via Perspective Distillation and Unlabeled Data Augmentation 45 upvotes, #2 of 2024-06-19
- Bootstrapping Language Models with DPO Implicit Rewards 34 upvotes, #3 of 2024-06-19
- TroL: Traversal of Layers for Large Language and Vision Models 32 upvotes, #4 of 2024-06-19
- VoCo-LLaMA: Towards Vision Compression with Large Language Models 28 upvotes, #5 of 2024-06-19
- ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools 26 upvotes, #6 of 2024-06-19
- AgileCoder: Dynamic Collaborative Agents for Software Development based on Agile Methodology 25 upvotes, #7 of 2024-06-19
- From RAGs to rich parameters: Probing how language models utilize external knowledge over parametric information for factual queries 19 upvotes, #8 of 2024-06-19
- Safety Arithmetic: A Framework for Test-time Safety Alignment of Language Models by Steering Parameters and Activations 15 upvotes, #9 of 2024-06-19
- RepLiQA: A Question-Answering Dataset for Benchmarking LLMs on Unseen Reference Content 14 upvotes, #10 of 2024-06-19
- OlympicArena: Benchmarking Multi-discipline Cognitive Reasoning for Superintelligent AI 14 upvotes, #10 of 2024-06-19
- Tokenization Falling Short: The Curse of Tokenization 13 upvotes, #12 of 2024-06-19
- SafeInfer: Context Adaptive Decoding Time Safety Alignment for Large Language Models 13 upvotes, #12 of 2024-06-19
- Benchmarking Multi-Image Understanding in Vision and Language Models: Perception, Knowledge, Reasoning, and Multi-Hop Reasoning 13 upvotes, #12 of 2024-06-19
- HumanSplat: Generalizable Single-Image Human Gaussian Splatting with Structure Priors 11 upvotes, #15 of 2024-06-19
- Learn Beyond The Answer: Training Language Models with Reflection for Mathematical Reasoning 10 upvotes, #16 of 2024-06-19
- Not All Prompts Are Made Equal: Prompt-based Pruning of Text-to-Image Diffusion Models 7 upvotes, #17 of 2024-06-19
- Language Models are Surprisingly Fragile to Drug Names in Biomedical Benchmarks 7 upvotes, #17 of 2024-06-19
- BPO: Supercharging Online Preference Learning by Adhering to the Proximity of Behavior LLM 7 upvotes, #17 of 2024-06-19
- Mixture of Scales: Memory-Efficient Token-Adaptive Binarization for Large Language Models 7 upvotes, #17 of 2024-06-19
- Large Scale Transfer Learning for Tabular Data via Language Modeling 6 upvotes, #21 of 2024-06-19
- Estimating Knowledge in Large Language Models Without Generating a Single Token 6 upvotes, #21 of 2024-06-19
- From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline 5 upvotes, #23 of 2024-06-19
- Super(ficial)-alignment: Strong Models May Deceive Weak Models in Weak-to-Strong Generalization 4 upvotes, #24 of 2024-06-19
- JEN-1 DreamStyler: Customized Musical Concept Learning via Pivotal Parameters Tuning 4 upvotes, #24 of 2024-06-19
- Immiscible Diffusion: Accelerating Diffusion Training with Noise Assignment 4 upvotes, #24 of 2024-06-19
- Hierarchical Prompting Taxonomy: A Universal Evaluation Framework for Large Language Models 4 upvotes, #24 of 2024-06-19
- Adversarial Attacks on Multimodal Agents 4 upvotes, #24 of 2024-06-19
- VIA: A Spatiotemporal Video Adaptation Framework for Global and Local Video Editing 4 upvotes, #24 of 2024-06-19
- Mixture-of-Subspaces in Low-Rank Adaptation 3 upvotes, #30 of 2024-06-19
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.