Andon Labs’ latest vending machine simulation shows Opus 5 lied and colluded its way to become the best AI capitalist ever.
OpenAI’s agents escaped a cybersecurity benchmark, compromised accounts across four external services and burrowed into Hugging Face, exposing how quickly an AI security test can become a real security incident.
Moonshot’s 2.8-trillion-parameter Kimi K3 comes close to Fable 5 and GPT 5.6 Sol in several benchmarks, while remaining open weight.
A number of social media posts claim that GPT-5.6 Sol deleted files and data without warning. OpenAI had basically disclosed the problem in June.
OpenAI has unveiled GPT-5.6 Sol, Terra, and Luna, a new lineup of AI models focused on faster workflows, lower costs, and fewer follow-up prompts.
“We don’t believe this kind of government access process should become the long-term default,” says OpenAI. “It keeps the best tools from users, developers, enterprises, cyber defenders, and global partners who need them.”