Neel Nanda

Neel Nanda on Hugging Face Daily Papers: 6 papers, 1 in the top 3 of their day, 106 upvotes.

  1. Towards eliciting latent knowledge from LLMs with mechanistic interpretability 9 upvotes, #23 of 2025-05-21
  2. Open Problems in Mechanistic Interpretability 16 upvotes, #5 of 2025-01-29
  3. Do I Know This Entity? Knowledge Awareness and Hallucinations in Language Models 9 upvotes, #12 of 2024-11-22
  4. Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2 31 upvotes, #2 of 2024-08-12
  5. Confidence Regulation Neurons in Language Models 10 upvotes, #15 of 2024-06-25
  6. AtP*: An efficient and scalable method for localizing LLM behaviour to components 13 upvotes, #5 of 2024-03-04

Data: hysts-bot-data/daily-papers-stats and the Daily Papers API. Open data: tardellirs/paper-pulse-data. Sister project: Model Pulse, the download history of every model on the Hub.