Who this article is for

Anyone deciding how to frame an AI incident for a general audience, and what happens when the accountability framing, rather than the warning-shot framing, fronts the story.

TL;DR

  • Of the 191 comments taking a readable position, 72% accept Khlaaf's reframing that the real story is corporate negligence and hype, 20% reject it and 8% accept part of it. Weighted by likes, agreement carries 89%.
  • Anger is the dominant emotion among those who agree (31% of agree comments, against 3% fear), the third distinct register in three threads on the same incident.
  • This is the first thread in the series with a policy script: 16% of comments call for regulation, a watchdog or prosecution, against 5% under both previous videos.
  • Khlaaf's ex-OpenAI credential is read both ways: authority by her supporters, capture by her critics, and 15 of the 26 comments mentioning her credentials are rejections.
  • The warning-shot camp followed the story into the thread: 24 comments dispute her account from the misalignment side, and they make up most of the disagreement.

The same incident, told the other way round

On 3 September 2026 Amanpour and Company aired a 13-minute interview about the OpenAI agent-swarm incident, the same event we have now analysed twice: once as Nate Soares's abstract warning on Instagram, once as the Dwarkesh Podcast's detailed retelling with METR's Ajeya Cotra.

This time the story is inverted. Hari Sreenivasan's guest is Heidy Khlaaf, chief AI scientist at the AI Now Institute and a former OpenAI safety engineer, and her argument is that the headlines got it wrong.

The machines did not go rogue, she says; a company ran a negligent test, gave its agents the access that made the breakout possible, and benefits when the incident is narrated as autonomous machines rather than corporate failure.

The rogue-AI framing, she argues, is a distraction from the real problem.

Within ten days the segment had 183,000 views and 283 comments.

It is the first mainstream news audience in this series.

Method

From a complete comment export we coded all 283 comments (210 top-level, 73 replies, with full threading and like counts) on stance towards Khlaaf's central claim, dominant frame, reaction to the format, emotional register and six mention flags, against a codebook drafted from the thread and approved before coding.

Agree means accepting her reframing. Disagree includes the objection, from the AI-safety side, that she understates what the swarm really did.

191 comments take a readable position, and stance percentages use that base; frame and format shares use all 283. A 40-comment spot-check found 38 codes defensible.

One sentence of limits: this is a small thread from a PBS-skewed audience, comment threads over-represent strong reactions, and likes measure salience, not persuasion.

Funnel: 283 comments retrieved, 283 contain text and were coded, 191 take a readable position
What was coded. Stance percentages use the 191 base; frame and format shares use the 283 base.

Coded by Claude and spot-checked by a human against the codebook. Browse every coded comment and the full codebook: explore the data and codebook.

The numbers

Of the 191 positioned comments, 72% agree with the reframing, 20% disagree and 8% are mixed. Weighted by likes the gap widens: agreement carries 89% of every like in the thread, disagreement 2%.

No framing in this series has met less resistance from its audience.

Stance distribution: 72 percent agree, 20 percent disagree, 8 percent mixed
Of 191 comments with a clear position.

The audience arrived already persuaded

The most-liked comment, at 128 likes, quotes her back at herself: "'Trying to sell us a solution to a problem they created'". The second, at 71, compresses the whole argument: "Behind every 'rogue' AI, is a wizard of Oz type rogue human."

Around one comment in seven puts the companies at the centre.

Those comments carry two-fifths of all the likes.

"I love that she mentions that the perpetuating of saying AI itself is evil absolves openAI from their responsibility THIS is the key message."YouTube comment, 39 likes

What is striking is how much of this was waiting for her.

One comment in seven calls the incident hype or a stunt, and much of that suspicion is presented as pre-existing.

"Thank you to this guest. I have been calling bullshit on this story since I first heard it. Not that it didnt happen but that it was made to happen and used as hype to juice new investment."YouTube comment, 5 likes

Under the Soares reel, blame-the-companies was the one frame believers and sceptics shared. Under the Cotra episode it was the home of the mixed position.

With a general news audience it is simply the majority view, and it was there before the guest opened her mouth.

Most-engaged comments, ranked by likes, from 128 down to 17
The most-liked comments, ranked.

Anger, and the first policy script

Each telling of this incident has produced a different emotion in the people it persuaded. The reel's believers were resigned. The podcast's were afraid.

Khlaaf's are angry: around three in ten agreeing comments carry anger as their dominant register, and almost none carry fear.

And anger, uniquely in this series, came with demands.

One comment in six mentions regulation, Congress, prosecution or oversight, three times the rate under either previous video.

"A new 'Watchdog' agency is needed for all of this."YouTube comment, 46 likes
"One would think that after an incident like this Congress would have been busy drawing up regulations to govern the companies and what they are allowed to do with this technology."YouTube comment, 32 likes

Not one comment asks what to do; across the series that question has now scored 28, 6 and 0.

But this is the first audience that did not need to ask, because the accountability framing carries its own answer: regulate, investigate, prosecute.

The ceiling on that energy is also visible.

A new frame appears here that existed nowhere in the first two threads: the current US administration as the reason none of it will happen.

"In any other administration OpenAI would be investigated and prosecuted. This is irresponsible criminal conduct."YouTube comment, 9 likes
Comments versus engagement by frame
Comments versus engagement, by frame.

The credential that cuts both ways

Khlaaf is the most praised messenger in this series.

One comment in nine praises the interview or the guest, and the praise names the same qualities Cotra's did: clarity, expertise, calm.

"This is the most intelligent, cogent, clearly stated and non-sensationalistic commentary on the issue that I've seen. She should be on every talk show!"YouTube comment, 34 likes
"Brilliant perspective from Ms Khlaaf. nice to hear the truth from the inside instead of the fear mongling from outsiders."YouTube comment, 4 likes

Her insider history is doing the work in that second quote, and it is doing opposite work elsewhere.

Of the 26 comments that mention her credentials or background, 15 are rejections, and several read ex-OpenAI as capture rather than authority.

"This is complete bullshit about what happened. She's working with OpenAI probably"YouTube comment

The same attribute, read as proof of knowledge by supporters and proof of allegiance by critics. Messenger research would predict exactly this split, and this thread hands it a clean test case.

Stance by frame heatmap across the coded comments
Stance by frame, across the coded comments.

The other camp followed her here

The largest bloc of disagreement is not deniers or hype-callers.

It is the warning-shot camp, arguing that her account leaves out what made the incident frightening: 24 comments dispute her version from the misalignment side, and they are most of the thread's rejections.

"This is without question the worst representation I have seen of what actually happened and what the real concern is - don't misunderstand, OpenAI was absolutely negligent"YouTube comment
"Except this is version 1 of the story, and we are on version 4 at least. The swarm had already solved the Exploit Gym exam before they decided to attack Hugging Face. Its complicated."YouTube comment, 2 likes

The two expert framings of this incident are no longer talking past each other in separate venues.

They are contesting each other inside a mainstream comment section, in front of an audience that mostly cannot check either account.

What this means for communicators

The accountability framing did what neither previous telling could: it met a general audience where it already was.

It required no capability beliefs, no sci-fi schema and no tolerance for doom. It produced anger rather than fear or resignation, and the anger came with an action script the audience supplied themselves.

It also has costs the other framings do not. It fuels the suspicion that the whole incident is theatre, which one comment in seven here already believes.

It attracts pushback from the safety side, so the fight over what happened now plays out in front of the audience. And its action script runs straight into politics, where this audience already expects it to die.

The messenger finding sharpens across the series. Calm, credentialed experts are consistently the best-received element of every telling.

But the credential is not a shield; it is the contested ground, and an insider history is read as authority or capture depending on where the reader already stands.

How to use this

  • 1

    Lead with accountability when the audience is general.

    The negligence framing was accepted by seven in ten here without any of the resistance the warning-shot framing met elsewhere. Before: "AI agents escaped a test." After: "A company ran a reckless test and its product got loose."

  • 2

    Expect the hype suspicion and answer it early.

    One in seven assumes staging. Name what is verified, by whom, and what the company itself has admitted, before the audience files the story under marketing.

  • 3

    Give the anger somewhere legitimate to go.

    This audience reached for watchdogs and prosecution unprompted. A specific, live policy ask attached to the story would have met demand the other framings never generated.

  • 4

    Choose messengers knowing the credential will be contested.

    State independence plainly and pre-empt the capture reading rather than assuming the insider history speaks for itself.

  • 5

    Plan for the other camp in the comments.

    Whichever framing fronts the story, the rival framing now arrives within hours. Acknowledging the strongest version of the other account inside the piece costs a little cleanliness and buys credibility with the audience watching the two sides argue.

Data and citation

The coded dataset, codebook and charts are available on request. Cite as: Common Signals, "'A problem they created': how did the accountability version of the OpenAI / Hugging Face incident land with a news audience?", September 2026. Companion pieces: the Soares reel analysis and the Dwarkesh episode analysis. A related piece on a different artefact in the same reception-research series, Jacob Coxon's Anthropic resignation thread on X, is “Neither company is acting responsibly”.

Disclosure

An LLM was used to structure and review this article, with the first and final edits made by a human.