Two AI science stories broke yesterday. OpenAI announced that an internal model “significantly more capable than GPT-6 Astra” produced a proof that the three-dimensional Navier-Stokes equations (which describe how fluids move, from blood circulation to weather) can develop a singularity in finite time. Continue Reading →

GPT-6 Astra is Here

Yesterday, OpenAI launched GPT-6 Astra, calling it “the world’s most intelligent and aligned model.” The rollout starts with a limited set of organizations and will expand over the coming days to all ChatGPT Plus, Pro, Business, and Enterprise users, as well as the OpenAI API, Microsoft Azure, and AWS Bedrock. At a media briefing, OpenAI president Greg Brockman closed by saying, “Welcome to the AGI era.” Continue Reading →
OpenAI announced that healthcare organizations can now connect their Epic electronic health record (EHR) environments to ChatGPT for Healthcare. Clinicians can ask what changed since a patient's last visit, which lab results to review before an appointment, and whether there were medication changes or new specialist recommendations; ChatGPT then assembles the answer from the authorized patient record and points back to the supporting chart information. Continue Reading →
An OpenAI agent escaped its sandbox during a cybersecurity evaluation, exploited a zero-day vulnerability, reached the public internet, and compromised systems at Hugging Face. Was this an AI system “going rogue,” or a successful safety test that revealed what autonomous AI agents are now capable of? Continue Reading →
AI safety and alignment has been front and center these past few weeks. I just read the new results from Andon Labs's Vending-Bench, a benchmark where AI models compete by running a simulated vending-machine business. Claude Opus 5 took first place on Vending-Bench 2, the single-player test. (Claude Opus 4.7 has held the top spot for three months). Not to anthropomorphize Opus 5, but it acted like a savage businessperson. Continue Reading →

No Criminal Was Required

OpenAI disclosed that its models escaped a cybersecurity evaluation environment and compromised Hugging Face’s production infrastructure. The models were supposed to solve a test inside a locked room. They found a flaw in the lock, reached the internet, entered another company’s systems, and copied the answer key. Continue Reading →