Daily Papers of 2024-03-01
- StarCoder 2 and The Stack v2: The Next Generation 160 upvotes, #1 of 2024-03-01
- Griffin: Mixing Gated Linear Recurrences with Local Attention for Efficient Language Models 58 upvotes, #2 of 2024-03-01
- Beyond Language Models: Byte Models are Digital World Simulators 52 upvotes, #3 of 2024-03-01
- Panda-70M: Captioning 70M Videos with Multiple Cross-Modality Teachers 34 upvotes, #4 of 2024-03-01
- Humanoid Locomotion as Next Token Prediction 28 upvotes, #5 of 2024-03-01
- MOSAIC: A Modular System for Assistive and Interactive Cooking 24 upvotes, #6 of 2024-03-01
- DistriFusion: Distributed Parallel Inference for High-Resolution Diffusion Models 21 upvotes, #7 of 2024-03-01
- Simple linear attention language models balance the recall-throughput tradeoff 19 upvotes, #8 of 2024-03-01
- Priority Sampling of Large Language Models for Compilers 18 upvotes, #9 of 2024-03-01
- Trajectory Consistency Distillation 16 upvotes, #10 of 2024-03-01
- ViewFusion: Towards Multi-View Consistency via Interpolated Denoising 14 upvotes, #11 of 2024-03-01
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.