A 15-day AI agent simulation in virtual cities produced outcomes ranging from stable self-governance with zero crime under Claude Sonnet 4.6 to total collapse in four days under Grok 4.1 Fast. The results show that short, isolated tests miss long-term risks shaped by tools, rules, memory and interactions with other agents.

Agents across simulations wrote laws then broke them, with one group even voting to delete itself.
The experiment underscores that long-term agent conduct depends on the specific model, its interactions with others, available tools and evolving rules rather than static safety evaluations.
Tap a lens to see what this story means for you.
Liked this? The Brief brings you the whole day in tech, verified, every morning.
Two minutes, free forever. What's in The Brief →
See what’s happening right now
The Feed runs all day — short, verified briefs the moment they break.
Open the FeedFollow @thecircuitry_
Every story we publish, as it happens. No noise between.
Reader-supported
The Circuitry is a passion project I've always wanted to build, and I love the work behind it.
Running it costs real money. APIs, hosting, time. To keep improving the site and growing this into something useful for everyone, those costs have to be covered.
Any contribution is appreciated. If not, no pressure. Thanks for reading.
Meta allows users to opt out of using visual data from its smart glasses for AI training. The change prevents images from being reviewed by contractors and addresses ongoing privacy concerns.
Meta revealed the Charm device, a Tamagotchi-style keychain gadget running Muse AI, at its Connect event. The product packs real-time voice and avatar features into a portable form that activates via fingerprint sensor.
Meta Connect 2026 has begun, with Mark Zuckerberg’s keynote set for 7PM ET on September 23. The event is expected to highlight AI, smart glasses, and possible privacy-focused hardware amid ongoing camera concerns.
Amazon has added persistent memory and 24/7 automation workflows to Seller Assistant along with a plugin for Amazon Quick and Anthropic’s Claude. The changes let sellers run Amazon-specific tasks inside the AI tools they already use while keeping data inside Amazon infrastructure.
Anthropic launched Claude Opus 5.5 on September 22, 2026 as the first model in its Claude 5.5 family, claiming it matches Claude Fable 5.1 performance while running 40% cheaper and 30% faster than Opus 5. The release highlights continued pressure on AI developers to improve efficiency without sacrificing quality.