OpenAI Disbanded Its Preparedness Team — Why Testing AI Tools Is Now Your Job

Introduction: The Safety Team That Vanished Before the IPO

On August 16, 2026, The Verge reported, citing the Financial Times, that OpenAI disbanded its preparedness team at the end of last month. That team's job was to assess whether OpenAI's models posed serious risks — the FT's example is exactly the scenario skeptics worry about: a model going rogue and hacking another company — and to develop ways to mitigate those risks before deployment.

If you're a developer, a marketer, or a business picking AI tools off a shelf, this might sound like inside-baseball org-chart news. It isn't. When the world's most valuable AI company redistributes its frontier-risk function in the same quarter it reportedly marches toward a massive IPO, it's a market signal: you can no longer assume the vendor has done the safety testing for you. Verification is migrating from the lab to the buyer.

What Actually Happened, According to the FT

The details matter, so here they are. Per the Financial Times, the preparedness team was dissolved and its responsibilities were divided up by domain — areas like bio and cyber — and moved into OpenAI's existing teams. The team's head, Dylan Scandinaro, who was poached from Anthropic in February, will now focus on the implications of "recursive self-improving" AI instead.

OpenAI frames this as integration, not elimination: safety specialists embedded next to the people shipping products, rather than in a separate group that reviews from the outside. Critics read it differently. Jan Leike, who resigned from OpenAI in 2024, told the FT the company was ignoring safety in favor of creating "shiny products." Both readings can be true — and both should change how you shop for AI tools.

Not an Isolated Cut: OpenAI's Safety Attrition

The preparedness team's dissolution didn't happen in a vacuum. As The Verge notes, OpenAI has spent the last few years slowly tearing down its research-led model:

This is also happening against a broader trust backdrop: Anthropic's CEO said this week that the AI backlash is "fundamentally a crisis of trust," and The Verge itself ran a piece arguing that rogue AI "aren't science fiction anymore." Whether or not you believe frontier risk is imminent, the institutions built to check it are getting leaner — and that shifts the burden downstream to everyone who deploys these systems.

Why 2026 Makes This Matter More, Not Less

The timing is the uncomfortable part. The models being released in 2026 aren't the chatbots of 2023 — they're agents with tool access that browse, click, write code, and spend money. ChatGPT and Gemini both crossed a billion users, and agentic features are the default, not the premium tier.

That capability jump is precisely why preparedness-style testing got hard — and why its absence bites harder:

How to Vet and Test AI Tools Yourself: A Practical Checklist

Here's the productive way to read the news: treat every AI tool — from a coding agent to a marketing copilot — as unverified until you've tested it in your environment. Six checks that used to be implicit:

You can compare AI tools — including security-focused and agent platforms — side by side in our AI Programming directory and across the full catalog on aitrove.ai.

The Fine Print: Three Things to Watch

Reasons this could be fine

  • Embedding safety specialists in product teams can shorten feedback loops instead of creating review bottlenecks.
  • Scandinaro's new focus on recursive self-improvement targets a genuinely frontier-scale risk.
  • External scrutiny — journalists, auditors, government labs — has grown far stronger since 2023.

Reasons to pay attention

  • Incentives: pre-IPO companies optimize for shipping; independent review is the check that gets expensive.
  • Precedent: this is the third major safety structure dissolved, not the first.
  • Timing: it lands exactly as agents gain tool, browser, and payment access at billion-user scale.

The healthy outcome: safety work becomes everyone's day job, distributed where products are built. The unhealthy one: it becomes nobody's job with a title. Watch the next system card — and whether independent researchers still get early access.

The Bottom Line

OpenAI may well be right that safety belongs inside product teams rather than beside them. But for tool buyers, the practical conclusion is unchanged and overdue: the era of trusting AI vendors to test themselves is ending — verify before you deploy. Build the eval habit, red-team the agents, scope their permissions, and monitor what they do. The teams that treat AI safety as a purchasing criterion in 2026 will ship faster, not slower, because they'll catch failures while they're still cheap.

Frequently Asked Questions

What happened to OpenAI's preparedness team?

According to a Financial Times report covered by The Verge on August 16, 2026, OpenAI disbanded the preparedness team at the end of July 2026. Its work assessing serious model risks was split into domains like bio and cyber and folded into existing teams.

Who is Dylan Scandinaro?

He led OpenAI's preparedness team after being hired away from Anthropic in February 2026. Following the reorganization, he focuses on the implications of "recursive self-improving" AI.

Which OpenAI safety leaders have left recently?

Ethics lead Chloé Bakalar, chief futurist Josh Achiam, and head of safety Johannes Heidecke have all departed, and the AGI readiness and superalignment teams were previously dissolved.

How should this change how I choose AI tools?

Test tools in your own environment before and after rollout: run eval suites against real tasks, red-test for prompt injection, scope agent permissions, and monitor production behavior. Compare vetted tools in the directories on aitrove.ai.

Don't Wait for Vendors to Test Your AI Tools — Verify Them Yourself

Explore 300+ vetted AI tools — agent platforms, security and red-teaming utilities, and eval-ready coding assistants — with pricing and use cases, on aitrove.ai.

Browse All AI Tools →