Interpretability researcher at Anthropic known for work on the biology of large language models — reverse-engineering the features and circuits inside models to see how they actually compute.
Interpretability researcher at Anthropic known for work on the biology of large language models — reverse-engineering the features and circuits inside models to see how they actually compute.