Mechanistic interpretability project that localizes entity-selective neurons (“entity cells”) in language models and uses causal interventions on PopQA-style factual question answering to study how these neurons mediate entity-centric factual recall.
Endorsements made here support MentaLeap.
Mechanistic interpretability project that localizes entity-selective neurons (“entity cells”) in language models and uses causal interventions on PopQA-style factual question answering to study how these neurons mediate entity-centric factual recall.
Endorsements made here support MentaLeap.
People
Updated 05/18/26 · By grantmaking.aiCo-author
Discussion
Sign in to comment
No comments yet. Be the first to share your thoughts.