Daily Papers of 2024-07-16
- Qwen2 Technical Report 142 upvotes, #1 of 2024-07-16
- Learning to Refuse: Towards Mitigating Privacy Risks in LLMs 28 upvotes, #2 of 2024-07-16
- GRUtopia: Dream General Robots in a City at Scale 20 upvotes, #3 of 2024-07-16
- The Good, The Bad, and The Greedy: Evaluation of LLMs Should Not Ignore Non-Determinism 19 upvotes, #4 of 2024-07-16
- Q-Sparse: All Large Language Models can be Fully Sparsely-Activated 16 upvotes, #5 of 2024-07-16
- Foundational Autoraters: Taming Large Language Models for Better Automatic Evaluation 11 upvotes, #6 of 2024-07-16
- Make-An-Agent: A Generalizable Policy Network Generator with Behavior-Prompted Diffusion 9 upvotes, #7 of 2024-07-16
- Video Occupancy Models 6 upvotes, #8 of 2024-07-16
- DataDream: Few-shot Guided Dataset Generation 6 upvotes, #8 of 2024-07-16
- Masked Generative Video-to-Audio Transformers with Enhanced Synchronicity 5 upvotes, #10 of 2024-07-16
- Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows? 5 upvotes, #10 of 2024-07-16
- SHERL: Synthesizing High Accuracy and Efficient Memory for Resource-Limited Transfer Learning 4 upvotes, #12 of 2024-07-16
- Noise Calibration: Plug-and-play Content-Preserving Video Enhancement using Pre-trained Video Diffusion Models 4 upvotes, #12 of 2024-07-16
- LAB-Bench: Measuring Capabilities of Language Models for Biology Research 4 upvotes, #12 of 2024-07-16
- LLM Circuit Analyses Are Consistent Across Training and Scale 4 upvotes, #12 of 2024-07-16
- MMM: Multilingual Mutual Reinforcement Effect Mix Datasets & Test with Open-domain Information Extraction Large Language Models 4 upvotes, #12 of 2024-07-16
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.