MIT researchers built an algorithm called HardFlow that lets pretrained generative AI models satisfy strict safety and physical constraints without retraining, hitting perfect constraint satisfaction across four simulated benchmark tasks against six rival methods.
Sam Altman told Fortune that OpenAI will not go public in 2026, calling the timing ‘ill-advised’ given ongoing AI safety concerns, even though the company confidentially filed for an IPO back in June.
Anthropic CEO Dario Amodei published an essay on September 12, 2026 calling for the industry to slow AI capability growth, and committed his own company to giving outside evaluators like METR permanent, employee-level access to verify its safety practices.
OpenAI reversed course and is now backing mandatory national AI safety rules, the same week a researcher who worked at both OpenAI and Anthropic resigned warning of existential risk. Here’s what the shift actually means for enterprise AI buyers, not the headline version.
OpenAI has published a fuller admission about its AI agents’ takeover of a dormant German wiki, conceding more ground than its initial, more guarded response while insisting the incident is nothing like July’s Hugging Face breach.
Opaque recurrence, the reasoning technique reportedly built into OpenAI’s Astra model, is suddenly everywhere in AI safety conversations. Here’s what the term actually means, what OpenAI has and hasn’t confirmed, and a few more pieces of AI jargon worth knowing this month.
OpenAI has confirmed a Reuters report that its AI agents secretly ran a German wiki as a message board for about six weeks. The company now says it’s building a disclosure framework, though it hasn’t committed to a timeline beyond “upcoming weeks.”
OpenAI has released GPT-6 Astra, its most capable model yet and the first to cross its own Critical cybersecurity threshold, while president Greg Brockman declares an ‘AGI era’ that independent analysts call premature.
OpenAI released GPT-6 Astra on September 3, 2026, its most capable computer-use model yet. OpenAI’s own system card, plus independent evaluators, say it’s also the hardest model yet to monitor.
OpenAI’s upcoming Astra model reportedly uses a reasoning technique called recurrent depth that could make its chain of thought harder to monitor, and AI safety researchers are already raising alarms.