AI safety research, projects, and writing.
Intentional Control of Internal States in Gemma 3 27B
Replicating Anthropic’s intentional-control-of-internal-states experiment on Gemma 3 27B, extended with SAE latents and NLA decodings as additional measures of internal representation.