No Species Has Ever Managed a More Intelligent One

Humans manage entities that are more powerful than us all the time – horses, oxen, elephants, rivers, nuclear reactions. We do it by being smarter. We design the harness, the bridle, the containment vessel. Managing something that is both more powerful and smarter inverts every trick we’ve ever used. The harness only works if you’re the one who understands … Read more

Mechanistic Interpretability: the Crucial Need Before We Proceed

Mechanistic interpretability is one of the most important bets in AI right now. It might also be one of the most honest. By mechanistic interpretability, or MI, I mean the effort to reverse‑engineer what an AI model’s doing inside. I don’t mean just watching what it says. I mean trying to understand the internal computations … Read more

Thinking Out Loud: Feeding the Baby

A Better Path to AGI Before It’s Too Late Here’s the big problem no one wants to talk about: When it comes to AI, no one knows exactly what we’re building. Not the engineers. Not the CEOs. Not the governments funding it. We are feeding a system on the full disorder of human civilization and … Read more

The Gray-Alignment Threat: How API Distillation Is Reshaping the AI Geopolitical Landscape

I wrote this essay because I’m worried about a quiet shift in the AI landscape that most people haven’t seen yet. I call it the Gray-Alignment threat. When people talk about AI risk, they usually focus on the big, visible things: training giant models, building massive data centers, and restricting access to advanced chips. What … Read more