Marc Carauleanu
Founder & Director
They are an AI alignment researcher focused on empathy, cooperation and deception in AI systems. Their experience includes research supported by the Long-Term Future Fund and a research fellowship at the Stanford Existential Risks Initiative. At the Center for AI Safety, their research on self–other overlap in cooperative reinforcement-learning agents became their thesis for which they were awarded the Brian Clarke Memorial Prize for “the most outstanding contribution to computing and mathematics”. At AE Studio, they led research on self–other overlap in state-of-the-art ML models, resulting in their first-authored paper “Towards Safe and Honest AI Agents With Neural Self-Other Overlap”.