other8.0 / 10
Language Model Circuits Are Sparse in the Neuron Basis
Abstract
A new technique for tracing sparse and faithful circuits directly on a model's MLPs
Research area
model robustness
Published
—
Source
other
Org
Transluce
Sign in to read and join the discussion.