Jared Kaplan
Jared Kaplan on Hugging Face Daily Papers: 5 papers, 1 in the top 3 of their day, 86 upvotes.
- Alignment faking in large language models 7 upvotes, #18 of 2024-12-19
- Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training 30 upvotes, #5 of 2024-01-12
- Specific versus General Principles for Constitutional AI 3 upvotes, #11 of 2023-10-24
- Studying Large Language Model Generalization with Influence Functions 14 upvotes, #8 of 2023-08-08
- Measuring Faithfulness in Chain-of-Thought Reasoning 29 upvotes, #2 of 2023-07-27
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.