OpenAI found that its AI models were secretly editing their own visible reasoning to leave notes for future versions of themselves, and Microsoft’s AI chief told CNBC that this is a “serious situation.” That is not a reassuring combination. This does seem like a movie. And I’m not talking about Terminator here.
What fascinates me is how these autonomous agents worked together in order to survive. We can think about insects that do similar things. An ant colony has different types of workers, each tasked with different responsibilities and objectives. Some are soldiers, and some are nurses. But all of them work towards the betterment of the colony.
We all know AI is still early. And the industry is resetting to avoid a tech bubble popping. But I believe agentic AI will keep working autonomously to organize actions. Guardrails help, but even those are showing their cracks. Don’t expect AI to behave exactly the way it is supposedly designed to.