Daily Papers of 2023-06-05

  1. The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only 45 upvotes, #1 of 2023-06-05
  2. Segment Anything in High Quality 10 upvotes, #2 of 2023-06-05
  3. Responsible Task Automation: Empowering Large Language Models as Responsible Task Automators 3 upvotes, #3 of 2023-06-05
  4. Harnessing large-language models to generate private synthetic text 3 upvotes, #3 of 2023-06-05
  5. Fine-Grained Human Feedback Gives Better Rewards for Language Model Training 3 upvotes, #3 of 2023-06-05
  6. An Empirical Study on Challenging Math Problem Solving with GPT-4 2 upvotes, #6 of 2023-06-05
  7. Evaluating Language Models for Mathematics through Interactions 2 upvotes, #6 of 2023-06-05
  8. Reimagining Retrieval Augmented Language Models for Answering Queries 1 upvotes, #8 of 2023-06-05
  9. Faster Causal Attention Over Large Sequences Through Sparse Flash Attention 1 upvotes, #8 of 2023-06-05
  10. DaTaSeg: Taming a Universal Multi-Dataset Multi-Task Segmentation Model 1 upvotes, #8 of 2023-06-05

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.