Achieving long-lasting, comprehensive human-AI alignment.
Achieving equitable economic prosperity through automation.
Current AI systems are aligned at the level of behavior — trained to produce the right outputs and trusted to have the right interior. We work one level down, on the activations themselves: reading and shaping what happens inside the model during inference. Two lines of work so far — an emotional memory that improves decisions, and a direction that detects and corrects deception.
All experimental code, evaluation protocols, and echo construction tools will be released publicly.
We welcome research collaborations and are actively seeking grant partnerships.