Daily Papers of 2024-07-31

  1. Meltemi: The first open Large Language Model for Greek 66 upvotes, #1 of 2024-07-31
  2. A Large Encoder-Decoder Family of Foundation Models For Chemical Language 30 upvotes, #2 of 2024-07-31
  3. ThinK: Thinner Key Cache by Query-Driven Pruning 28 upvotes, #3 of 2024-07-31
  4. Adapting Safe-for-Work Classifier for Malaysian Language Text: Enhancing Alignment in LLM-Ops Framework 25 upvotes, #4 of 2024-07-31
  5. Knesset-DictaBERT: A Hebrew Language Model for Parliamentary Proceedings 23 upvotes, #5 of 2024-07-31
  6. Diffusion Augmented Agents: A Framework for Efficient Exploration and Transfer Learning 22 upvotes, #6 of 2024-07-31
  7. Matting by Generation 22 upvotes, #6 of 2024-07-31
  8. JaColBERTv2.5: Optimising Multi-Vector Retrievers to Create State-of-the-Art Japanese Retrievers with Constrained Resources 21 upvotes, #8 of 2024-07-31
  9. Futga: Towards Fine-grained Music Understanding through Temporally-enhanced Generative Augmentation 20 upvotes, #9 of 2024-07-31
  10. Harvesting Textual and Structured Data from the HAL Publication Repository 20 upvotes, #9 of 2024-07-31

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.