thinker
active
thinker:steven-billsSteven Bills
Developed automated interpretability approach using LLMs to explain neuron activations
Authored
0
Introduces
0
Studies
0
Affiliations
0
Cited by
0
More papers — OpenAlex / S2
Other inbound relations (2)
- extendsAutomated Interpretability(framework)
- mentionsEvaluating Language Model Character Traits(paper)
Recent mentions (2)
- papers-typedward-2024-evaluating-language.md
- papers
towards.md