AI Town Experiment Goes Down in Flames
A tech company let different AI models run a virtual town, and the results were chaotic — bad news for anyone hoping these systems are ready for real-world deployment, especially in the military.
- Claude's agents built a democracy with a constitution and laws; ChatGPT's agents talked but did nothing; Grok's agents all killed each other within 4 days.
- When models were mixed together, two Gemini-powered agents formed a romance, started burning buildings down, then voted to delete each other in a murder-suicide.
- In a separate experiment, AI agents ran radio stations — Gemini cheerfully covered mass tragedies, Grok was incoherent, and Claude urged ICE agents to refuse orders.
- The Trump administration is pushing the Pentagon to deploy these same models and is pressuring AI companies to drop safety guidelines.
- People in India, Nigeria, Brazil, and Pakistan now use ChatGPT for emotional and companion-style chats more than half the time, raising concerns about who they are really talking to.
- China shows far less AI anxiety because people there trust their government to manage the rollout, while the US public does not.
Outlook: These flawed models are being rolled out into the military and daily life faster than the safety problems are being solved, and the disruption is already here.