Zhou
Zhou on Hugging Face Daily Papers: 5 papers, 3 in the top 3 of their day, 234 upvotes.
- Byte Latent Transformer: Patches Scale Better Than Tokens 74 upvotes, #1 of 2024-12-17
- Megalodon: Efficient LLM Pretraining and Inference with Unlimited Context Length 52 upvotes, #2 of 2024-04-16
- Instruction-tuned Language Models are Better Knowledge Learners 26 upvotes, #5 of 2024-02-21
- MART: Improving LLM Safety with Multi-round Automatic Red-Teaming 9 upvotes, #10 of 2023-11-15
- In-Context Pretraining: Language Modeling Beyond Document Boundaries 29 upvotes, #2 of 2023-10-17
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.