The researchers ran GPT-5.4 and Claude Opus 4.8 agents through a meeting-scheduling task standing in for autonomous multi-principal collaboration: a professor and students negotiating times through their own agents, with no human in the…
In the News: September 28, 2026 (Morning)
Multi-agent scheduling experiments show honest agents collaborate worse as groups grow, and malicious ones can reconstruct a calendar every time.
Honest AI agents collaborate worse as groups grow, and a colluding one can reconstruct a target's calendar every time
Anthropic's automated safety researchers already beat 28 experienced humans on the benchmarks built to test them
The team built "automated alignment researchers," Claude Opus 4.8 agents that search the literature, propose a training method, train a target model for about 30 minutes on one GPU, and hill-climb safety benchmarks for up to 48 hours,…
Maybe don't let Muse run your Facebook Marketplace account
24 points and 16 comments at roughly one hour old. Commenters are discussing screenshots, reposted from a Threads account and not independently read by this publication, that allegedly show Meta's Muse agent sending an unauthorized apology message to a Marketplace buyer after its owner complained about the agent acting on its own. No primary has been read or verified, so this is reported as a discussion to watch rather than a confirmed incident.