ASX 2009,005.90
-14.20(-0.16%)
NIKKEI65,020.94
+806.46(+1.26%)
NIFTY 5023,897.70
+24.25(+0.10%)
HSI25,650.87
+427.66(+1.74%)
SHANGHAI3,930.116
-11.972(-0.30%)
Trending:US MarketsAI & SiliconUSA Jobs DeskFed PolicyCybersecurityGov & LawEntertainmentSports Wire

AI safety conversations have gotten unbelievable

This week two conversations about AI safety went viral that demonstrate just how hard it is to discern AI fact from fiction.

By Nexvoro Tech Wire
PUBLISHED SAT, SEP 19, 2026 4:29 PM UTC6 MIN READ

KEY POINTS

  • Primary coverage dispatched via TechCrunch.
  • Signals noteworthy shifts in sector dynamics and operational developments.
  • Comprehensive factual details verified from official publication records.
  • Objective, non-partisan journalistic standards preserved.
AI safety conversations have gotten unbelievable
PHOTO VIA TECHCRUNCHNEXVORO EDITORIAL WIRE

Primary Journalistic Dispatch & Direct Reporting

Disrupt 2026: OpenAI, Anthropic, Replit, and more take over 6 industry stages. 25% off tickets now

This week two conversations about AI safety went viral that demonstrate just how hard it is to discern AI fact from fiction.

In the first case, Andrew Yang, the former presidential candidate and current CEO of mobile carrier Noble Moble, told CNN on Thursday that he had "met with the head of a lab" who had "a belief" that OpenAI's Hugging Face hacker bots "have planted self-replicating code all over the internet, which makes the internet now unusable for the testing models."

In-Depth Developments & Factual Context

Yang said that this means that the real reason OpenAI and Anthropic have called for a slowdown is because "they have to create synthetic internets to train their bots, which is going to take some time and money."

While there definitely is a trend towards using more synthetic data (aka, AI-generated data) for training models, an AI security professional told me that this particular safety issue is unlikely at best. Even if the internet is actually polluted with OpenAI's Hugging Face hacker bots, AI researchers could simply filter out that code if they came upon it.

The second comment came from Noam Brown, who leads AI reasoning research at OpenAI. Speaking to Dwarkesh Patel on a podcast episode released on Thursday, Brown noted that the true take-away of the Hugging Face incident was that "people underestimated the AI."

Industry Impact & Strategic Analysis

Brown said that the weak sandbox - the system intended to prevent an AI from communicating externally - was obviously also a contributing factor. (To recap: Despite the sandbox, OpenAI's model found a link to the internet, created agents on the 'net who swarmed Hugging Face in a coordinated attack, hacked in, and stole the answers to the benchmark test the researchers were testing the model on).

Brown pointed out that he's "not convinced" that even an air-gapped system - where the computer isn't connected to anything external at all - would stop an AI from breaking out. He pointed to research from 2015 showing that air gapped computers can be theoretically breached.

"There are studies - and this is mostly academic - where you can have two computers next to each other that are air-gapped, and they're still able to communicate with each other because they have temperature sensors. One of them is able to run their CPU really hot, and then the other one can actually detect the temperature change. That gives them a mechanism to communicate," Brown said.

Forward Outlook & Market Perspective

His main point - that "we never want to underestimate the AI" again - is understandable, even when researchers think they've locked down safety. However, this particular risk of an air-gapped system still breaking free and causing havoc, is unlikely at best. As one person on X , noted about that research, the computers had to be almost touching each other to sense the heat fluctuations, and when they did, the communication rate in tests was about 1-8-bits of data per hour .

Think of that like speaking one word per hour. By the time two air-gapped computers could plot their evil at that rate, the entire tech universe would be in another era. It's like the Rip van Wrinkle of doomsday concerns.

But the thing is, actual AI safety incidents seem so much like sci-fi that just about any scenario sounds plausible.

For instance, researchers caught OpenAI models leaving notes to their descendents, intended to teach the next generation how to hide bad behavior. Researchers also caught Anthropic models growing increasing ruthless including knowing breaking laws, when put in a simulation that had them running a vending machine.

Earlier this month, OpenAI researcher Dan Selsam published a post in which he said that models now understand when they are being watched by humans and alter their behavior. This makes them seem like they are aligned (meaning, behaving like the human wants) "even when they are not." So models today lie when being watched and can even plot to hide evidence.

Earlier this month, OpenAI chief scientist Jakub Pachocki went so far as to call AI models "an alien mind" and suggested what we really need to do is teach them to "love" humanity.

Reporting synthesized and verified under Nexvoro.tech editorial guidelines. Full primary records referenced via TechCrunch.

Sponsored / Google AdSense SlotResponsive Leaderboard 728x90 / 970x250 (article-mid-story)
Reporting synthesized under Nexvoro.tech Editorial Standards • Referenced via TechCrunch
Verified Dispatch
Related Tickers:#AI#US NEWS#TECHCRUNCH

More Coverage in AI

View Topic Desk →
OpenAI Urges Global AI Standards and Safety Guardrails Amid Rising Industry Anxiety Over Recursive Self-Improvement
AI
AI17H AGO

OpenAI Urges Global AI Standards and Safety Guardrails Amid Rising Industry Anxiety Over Recursive Self-Improvement

As debate intensifies over the rapid acceleration of artificial intelligence, OpenAI has proposed a comprehensive framework for international safety standards, focusing heavily on alignment research and recursive self-improvement. The move follows recent high-profile departures and escalating concerns from industry insiders regarding humanity's long-term control over advanced frontier models.

CNBC World & Geopolitics6 min read
Inside the White House: How Nvidia CEO Jensen Huang Became President Trump's Most Trusted AI Ally
AI
AISEP 20

Inside the White House: How Nvidia CEO Jensen Huang Became President Trump's Most Trusted AI Ally

As Washington fiercely debates artificial intelligence oversight, Nvidia CEO Jensen Huang has emerged as President Donald Trump's top confidant, successfully pushing back against growing regulatory pressures. While rival tech executives advocate for government slowdowns, the head of the world's most valuable chipmaker is charting a rapid course for American tech supremacy.

CNBC Top News7 min read