详细内容
Open & scalable technology for understanding AI systems
We are an independent research lab working toward responsible development and deployment in the public interest.


Highlights
\
\
research report\
\
Scalably Extracting Latent Representations of Users \
\
Constructing datasets and training decoders to extract user models from language models\
\
25 November 2025
\
\
research report\
\
Language Model Circuits Are Sparse in the Neuron Basis \
\
A new technique for tracing sparse and faithful circuits directly on a model's MLPs\
\
20 November 2025
\
\
technical demonstration\
\
Monitoring SWE-bench Agents \
\
Partnering with SWE-bench to enable reliable monitoring of AI coding agents\
\
19 November 2025
