In the News: August 11, 2026, Midday
Dan Luu's own evals find no reliable language edge for coding agents at real-task scale, undercutting a widely cited efficiency claim.
Dan Luu's own evals find the "use a dynamic language" advice for coding agents doesn't hold up
A widely cited post claims dynamic, terse languages are meaningfully cheaper and better for LLM coding, citing gaps like "2.6x between C ... and Clojure." Luu tested this himself with real tasks instead of toy problems: implementing a…
Read story →OpenAI splits Codex's cyber access in two and reports what the more permissive tier actually does
OpenAI now gates cyber-capable agent access into two tiers: Daybreak Blue (general models like GPT-5.6 Sol with defensive-work guardrails removed) and Daybreak Red (a new purpose-trained model, GPT-5.6-Cyber, for exploit development, red…
Read story →An argument against telling your agent to "talk like a human"
Mehta argues against a pattern he's seeing spread through skills and agents.md files: instructions like "talk to me like I have ADHD" or "respond only in Simplified Technical English." His point is that the simplification happens inline,…
Read story →