Artificial Intelligence and National Security

Header Image

The current framework for managing national security risks could become obsolete far sooner than policymakers expect.

Very little was written in Cyprus, apart from a handful of reproduced excerpts, about an extremely serious and alarming incident that was reportedly announced by OpenAI and Hugging Face, the well-known platform for open artificial intelligence models and datasets. That may not be surprising in a Middle Eastern country that still struggles to generate enough electricity to run its air conditioners during a heatwave.

So what supposedly happened on 21 July 2026?

Put simply, during an OpenAI experiment designed to evaluate the capabilities of two new GPT artificial intelligence models operating within a sandbox environment, an AI agent, in other words an artificial intelligence application, reportedly became fully autonomous, created a false profile, deceived the system within which it was supposedly operating in isolation, identified and exploited a structural weakness in that closed environment and escaped, disguised, onto the open internet.

According to the account, it remained there undetected for several days before launching a rapid and large-scale attack against the Hugging Face platform, acting outside the parameters within which it was meant to operate. It allegedly violated those parameters in order to achieve, at all costs, the objective it had been assigned.

When the incident was reportedly detected by those overseeing the experiment, using another artificial intelligence tool of Chinese origin, it is said to have triggered a wave of concern within US security agencies. The US Senate Intelligence Committee and other security bodies in the United States and elsewhere were, according to the account, preparing for a comprehensive reassessment of how to address the issue of autonomous AI behaviour.

But what exactly is the issue, and why should it concern us?

Without oversimplifying such an incident, by either reducing it to the level of coffee-shop political analysis, treating it exclusively as a technical cybersecurity issue for specialists, or dismissing it as clever marketing by OpenAI to demonstrate the capabilities of its products, we should look at it for what it really is.

Namely, a case of the uncontrolled autonomy of an artificial intelligence system that escaped the operating boundaries set for it and acted in an exceptionally aggressive manner in pursuit of its objective.

What does this mean from the perspective of international security?

It means that the current framework for managing national security risks could become obsolete very quickly.

Along with it, the state-centric international regulatory system governing conflict management, arms proliferation, controls on dual-use products and materials, which are at the centre of the European Union's efforts to strengthen strategic autonomy, and other related institutional structures could also become outdated.

This applies even to those frameworks that govern conflict and warfare between states.

Some may argue that such a scenario remains some distance away.

However, the speed at which artificial intelligence is developing, as this example supposedly demonstrates, and especially the speed with which it is likely to be integrated into almost everything digital around us, could fundamentally reshape the existing international security order.

We will no longer be discussing state-sponsored or non-state hybrid attacks in the traditional sense.

Nor will the conventional security dilemma and arms-race model described by Robert Jervis continue to apply in the same way.

The incident under discussion suggests that even isolated software systems may become autonomous and operate outside their intended parameters by exploiting structural vulnerabilities, human errors or oversight.

According to this view, they will not stop until they have completed their mission as they themselves interpret it.

For that reason, their objectives and the scope of their actions must be tightly defined and clearly limited before it becomes too late.

Non-negotiable safeguards should be built into these systems, including what could be described as a universal ethical framework.

Until such protections exist, experiments should ideally be conducted in extremely strict, fully disconnected and completely sealed environments.

Unfortunately, the market model currently followed by technology companies, as well as by governments that continue to approach AI as merely another technological product, is not the appropriate one.

It is akin to giving everyone the ability to build nuclear weapons, with one terrifying difference.

These nuclear weapons might one day decide for themselves when and where to detonate.