Skip to main content

AI Hallucinations, Hacks, and High Stakes: What's Shaping Tech Right Now

Updated on September 19, 20266 minutes read

An AI hallucination nearly caused an international military incident, court documents confirmed what many suspected about how AI companies view the web, and researchers used one AI to hack into another. This week in tech had a lot going on, so here is what actually matters and why.

AI hallucination nearly sparked a real-world military response

This is the week's most alarming story, and it deserves to lead. A report confirmed that an AI system produced a fabricated intelligence summary involving Chinese nuclear components, and that summary came close to triggering a US military operation. Nothing happened in the end, but the situation exposed a real gap between how AI tools are being deployed and how well their users understand their limitations. A researcher from GovAI summed it up clearly: service members need to understand that large language models are fundamentally uncertain by nature.

For anyone studying AI or cybersecurity, this is a case study worth bookmarking. LLMs do not "know" things the way a database does. They generate plausible-sounding text, and plausible is not the same as accurate. Read the full account on TechCrunch or Ars Technica's detailed breakdown.

OpenAI and Microsoft's own documents called training data "the largest theft of labor"

Court documents unsealed in the New York Times' lawsuit against OpenAI and Microsoft revealed that the companies' own internal records raised serious concerns about their data practices. One Microsoft executive reportedly described the scraping of publisher content to train AI models as "the largest theft of labor in human history." Separately, internal notes warned that their approach could create a "doom loop" that drains original content from the web over time.

These are not outsider critiques. They are the companies' own words. For aspiring developers and designers who produce work online, this raises a genuine question about what it means to build on top of AI systems trained this way. The Verge has the full story on the unsealed documents, and CNET covers the journalism angle.

Researchers used Claude to break into OpenAI

A team of independent security researchers reportedly used Anthropic's Claude models to help them breach OpenAI employee accounts in under 72 hours. Through those accounts, they reportedly gained access to OpenAI's main internal code repository, which is said to contain core algorithmic work. The fact that one AI company's model was used as a tool to attack another is a striking development, and it signals that AI-assisted hacking is no longer theoretical. If you are learning cybersecurity, this is a concrete example of how the threat model is shifting. The Verge covers the breach in detail.

Anthropic is running a real biology lab

While Anthropic's safety researchers have been warning publicly about AI risks, the company has quietly stood up an actual wet-lab biology operation. The tension here is hard to miss: the same organisation whose researchers have described catastrophic AI risk scenarios is now conducting biology experiments, presumably to test whether AI can accelerate scientific discovery. TechCrunch reports on Anthropic's biology lab, and separately, Accenture has been named as Anthropic's first embedded safety evaluator — a move that puts a major consulting firm at the centre of frontier AI oversight.

AI governance is getting louder, from California to Brussels

Two significant policy moves landed this week. California Governor Gavin Newsom issued an executive order directing a group of experts to produce recommendations on AI oversight, including the possibility of a "kill switch" for frontier models. The order gives that group two months to report back. Meanwhile, the EU's proposed Kids Act would bring sweeping changes to social media, gaming, and AI chatbots, including mandatory age verification and a ban on social media use for under-13s.

For anyone building products in these spaces, both moves are worth watching. Compliance requirements shape architecture decisions, and they tend to arrive faster than most product teams expect. The Verge covers Newsom's AI kill-switch order, while CNET explains the EU Kids Act proposal.

Disney hires its very first CTO — from an AI startup it once sued

Disney has appointed Karandeep Anand as its first-ever chief technology officer. The appointment is notable on its own, but the context makes it more interesting: Anand was the CEO of Character.AI, a company Disney previously sent a cease-and-desist letter to over concerns about character likeness. The hire suggests Disney is serious about moving fast on AI, complicated history notwithstanding. For anyone interested in the intersection of entertainment, intellectual property, and generative AI, this one is worth following. TechCrunch has the background on the appointment.

A new kind of AI model is getting developers excited

A model called Jev, built by one of ChatGPT's original creators, is generating real interest in developer circles. The claim is that it offers a faster and cheaper path to software intelligence than existing approaches. Details are still sparse, but the pattern is familiar: a credible founder, a novel architecture claim, and early developer enthusiasm. Whether it holds up under scrutiny remains to be seen. TechCrunch has the story on Jev and what developers are saying.

Physical AI gets a $100M vote of confidence

Vantora, previously known as UP.Labs, raised $100 million and announced it is focused entirely on building AI-powered startups for industrial corporations. Rather than software-only products, Vantora is targeting what the industry calls "physical AI" — systems that interact with the real, physical world inside factories, logistics networks, and industrial operations. This is a different space from the consumer AI tools most people use daily, and it is attracting serious capital. TechCrunch covers Vantora's raise and strategy.

Joby Aviation completes a 3,100-mile autonomous flight

An aircraft running Joby Aviation's autonomy software flew coast-to-coast across the United States without any human pilot intervention at any point during the journey. Joby started out focused on short-range electric air taxis, and this flight marks a clear statement that the company's ambitions extend well beyond urban hops. Autonomous aviation raises distinct engineering and safety challenges compared to ground vehicles, and TechCrunch's report on the flight is a good primer on where this technology is heading.

FAA plans an $875M AI tool for air traffic management

The Federal Aviation Administration is moving toward deploying an AI system to help manage air traffic congestion, starting around Washington DC before a potential nationwide rollout. The projected cost is $875 million. Coming in the same week as the AI hallucination story above, the contrast is stark: the same technology that nearly caused a real-world incident from a bad output is also being considered for one of the most safety-critical infrastructure roles imaginable. Ars Technica covers the FAA's AI traffic plan.

The through-line connecting most of these stories is a growing gap between AI deployment speed and AI reliability. Governance bodies, courts, and security researchers are all, in their own ways, pushing back on that gap this week. How the industry responds over the next few months will likely set the tone for how AI gets built and regulated for years to come.

Learn technical skills online with Code Labs Academy

Learn technical skills online with Code Labs Academy

Join our supportive community, unlock your potential, and embark on a rewarding career path.

Frequently asked questions

What is an AI hallucination and why is it dangerous in high-stakes settings?

An AI hallucination is when a large language model generates text that sounds confident and plausible but is factually wrong or entirely made up. In everyday use this might mean a wrong recipe or a bad code suggestion. In high-stakes settings like military intelligence or air traffic control, a convincing but false output can trigger real decisions with serious consequences, which is exactly what nearly happened in the incident reported this week.

What did the OpenAI and Microsoft court documents actually reveal?

Unsealed court documents in the New York Times lawsuit showed that people inside both companies had raised concerns about their data practices. One Microsoft executive reportedly described the mass scraping of publisher content to train AI models as the largest theft of labor in human history. Internal documents also warned that their approach could damage the wider web by reducing the incentive to create original content.

What is physical AI and how is it different from the AI tools most people use?

Physical AI refers to AI systems designed to operate in and interact with the real, physical world, such as in factories, warehouses, or industrial equipment. Most consumer AI tools, like chatbots or image generators, work entirely in software. Physical AI has to deal with sensors, hardware, real-time decision-making, and the consequences of errors in ways that software-only AI does not.

What does Disney hiring a CTO for the first time say about where entertainment is heading?

It suggests Disney now sees technology strategy, particularly AI, as central enough to its business to need a dedicated executive at the top level. The fact that they hired someone from the AI startup world, rather than promoting internally, points to a desire to move quickly and bring in outside thinking about how AI can be embedded into a company that spans film, TV, theme parks, and streaming.

What is the EU Kids Act and what would it change for developers building apps?

The EU Kids Act is a proposed regulation that would require stronger age verification across social media platforms, online games, and AI chatbots, and would ban social media access for children under 13. For developers, this means building products for European users would require implementing compliant age-gating systems, which adds technical complexity and raises questions about data handling for younger users.

Career services

Personalized career support to help you launch your tech career. Get résumé reviews, mock interviews, and industry insights, so you can showcase your new skills with confidence.