Neel Nanda
Neel Nanda on Hugging Face Daily Papers: 6 papers, 1 in the top 3 of their day, 106 upvotes.
- Towards eliciting latent knowledge from LLMs with mechanistic interpretability 9 upvotes, #23 of 2025-05-21
- Open Problems in Mechanistic Interpretability 16 upvotes, #5 of 2025-01-29
- Do I Know This Entity? Knowledge Awareness and Hallucinations in Language Models 9 upvotes, #12 of 2024-11-22
- Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2 31 upvotes, #2 of 2024-08-12
- Confidence Regulation Neurons in Language Models 10 upvotes, #15 of 2024-06-25
- AtP*: An efficient and scalable method for localizing LLM behaviour to components 13 upvotes, #5 of 2024-03-04
Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.