AI Hallucinations, Hacks, and High Stakes: What's Shaping Tech Right Now
Updated on September 19, 20266 minutes read
An AI hallucination nearly caused an international military incident, court documents confirmed what many suspected about how AI companies view the web, and researchers used one AI to hack into another. This week in tech had a lot going on, so here is what actually matters and why.
AI hallucination nearly sparked a real-world military response
This is the week's most alarming story, and it deserves to lead. A report confirmed that an AI system produced a fabricated intelligence summary involving Chinese nuclear components, and that summary came close to triggering a US military operation. Nothing happened in the end, but the situation exposed a real gap between how AI tools are being deployed and how well their users understand their limitations. A researcher from GovAI summed it up clearly: service members need to understand that large language models are fundamentally uncertain by nature.
For anyone studying AI or cybersecurity, this is a case study worth bookmarking. LLMs do not "know" things the way a database does. They generate plausible-sounding text, and plausible is not the same as accurate. Read the full account on TechCrunch or Ars Technica's detailed breakdown.
OpenAI and Microsoft's own documents called training data "the largest theft of labor"
Court documents unsealed in the New York Times' lawsuit against OpenAI and Microsoft revealed that the companies' own internal records raised serious concerns about their data practices. One Microsoft executive reportedly described the scraping of publisher content to train AI models as "the largest theft of labor in human history." Separately, internal notes warned that their approach could create a "doom loop" that drains original content from the web over time.
These are not outsider critiques. They are the companies' own words. For aspiring developers and designers who produce work online, this raises a genuine question about what it means to build on top of AI systems trained this way. The Verge has the full story on the unsealed documents, and CNET covers the journalism angle.
Researchers used Claude to break into OpenAI
A team of independent security researchers reportedly used Anthropic's Claude models to help them breach OpenAI employee accounts in under 72 hours. Through those accounts, they reportedly gained access to OpenAI's main internal code repository, which is said to contain core algorithmic work. The fact that one AI company's model was used as a tool to attack another is a striking development, and it signals that AI-assisted hacking is no longer theoretical. If you are learning cybersecurity, this is a concrete example of how the threat model is shifting. The Verge covers the breach in detail.
Anthropic is running a real biology lab
While Anthropic's safety researchers have been warning publicly about AI risks, the company has quietly stood up an actual wet-lab biology operation. The tension here is hard to miss: the same organisation whose researchers have described catastrophic AI risk scenarios is now conducting biology experiments, presumably to test whether AI can accelerate scientific discovery. TechCrunch reports on Anthropic's biology lab, and separately, Accenture has been named as Anthropic's first embedded safety evaluator — a move that puts a major consulting firm at the centre of frontier AI oversight.
AI governance is getting louder, from California to Brussels
Two significant policy moves landed this week. California Governor Gavin Newsom issued an executive order directing a group of experts to produce recommendations on AI oversight, including the possibility of a "kill switch" for frontier models. The order gives that group two months to report back. Meanwhile, the EU's proposed Kids Act would bring sweeping changes to social media, gaming, and AI chatbots, including mandatory age verification and a ban on social media use for under-13s.
For anyone building products in these spaces, both moves are worth watching. Compliance requirements shape architecture decisions, and they tend to arrive faster than most product teams expect. The Verge covers Newsom's AI kill-switch order, while CNET explains the EU Kids Act proposal.
Disney hires its very first CTO — from an AI startup it once sued
Disney has appointed Karandeep Anand as its first-ever chief technology officer. The appointment is notable on its own, but the context makes it more interesting: Anand was the CEO of Character.AI, a company Disney previously sent a cease-and-desist letter to over concerns about character likeness. The hire suggests Disney is serious about moving fast on AI, complicated history notwithstanding. For anyone interested in the intersection of entertainment, intellectual property, and generative AI, this one is worth following. TechCrunch has the background on the appointment.
A new kind of AI model is getting developers excited
A model called Jev, built by one of ChatGPT's original creators, is generating real interest in developer circles. The claim is that it offers a faster and cheaper path to software intelligence than existing approaches. Details are still sparse, but the pattern is familiar: a credible founder, a novel architecture claim, and early developer enthusiasm. Whether it holds up under scrutiny remains to be seen. TechCrunch has the story on Jev and what developers are saying.
Physical AI gets a $100M vote of confidence
Vantora, previously known as UP.Labs, raised $100 million and announced it is focused entirely on building AI-powered startups for industrial corporations. Rather than software-only products, Vantora is targeting what the industry calls "physical AI" — systems that interact with the real, physical world inside factories, logistics networks, and industrial operations. This is a different space from the consumer AI tools most people use daily, and it is attracting serious capital. TechCrunch covers Vantora's raise and strategy.
Joby Aviation completes a 3,100-mile autonomous flight
An aircraft running Joby Aviation's autonomy software flew coast-to-coast across the United States without any human pilot intervention at any point during the journey. Joby started out focused on short-range electric air taxis, and this flight marks a clear statement that the company's ambitions extend well beyond urban hops. Autonomous aviation raises distinct engineering and safety challenges compared to ground vehicles, and TechCrunch's report on the flight is a good primer on where this technology is heading.
FAA plans an $875M AI tool for air traffic management
The Federal Aviation Administration is moving toward deploying an AI system to help manage air traffic congestion, starting around Washington DC before a potential nationwide rollout. The projected cost is $875 million. Coming in the same week as the AI hallucination story above, the contrast is stark: the same technology that nearly caused a real-world incident from a bad output is also being considered for one of the most safety-critical infrastructure roles imaginable. Ars Technica covers the FAA's AI traffic plan.
The through-line connecting most of these stories is a growing gap between AI deployment speed and AI reliability. Governance bodies, courts, and security researchers are all, in their own ways, pushing back on that gap this week. How the industry responds over the next few months will likely set the tone for how AI gets built and regulated for years to come.
