Daily Papers of 2024-03-01

  1. StarCoder 2 and The Stack v2: The Next Generation 160 upvotes, #1 of 2024-03-01
  2. Griffin: Mixing Gated Linear Recurrences with Local Attention for Efficient Language Models 58 upvotes, #2 of 2024-03-01
  3. Beyond Language Models: Byte Models are Digital World Simulators 52 upvotes, #3 of 2024-03-01
  4. Panda-70M: Captioning 70M Videos with Multiple Cross-Modality Teachers 34 upvotes, #4 of 2024-03-01
  5. Humanoid Locomotion as Next Token Prediction 28 upvotes, #5 of 2024-03-01
  6. MOSAIC: A Modular System for Assistive and Interactive Cooking 24 upvotes, #6 of 2024-03-01
  7. DistriFusion: Distributed Parallel Inference for High-Resolution Diffusion Models 21 upvotes, #7 of 2024-03-01
  8. Simple linear attention language models balance the recall-throughput tradeoff 19 upvotes, #8 of 2024-03-01
  9. Priority Sampling of Large Language Models for Compilers 18 upvotes, #9 of 2024-03-01
  10. Trajectory Consistency Distillation 16 upvotes, #10 of 2024-03-01
  11. ViewFusion: Towards Multi-View Consistency via Interpolated Denoising 14 upvotes, #11 of 2024-03-01

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.