Your support helps us to tell the story

From reproductive rights to climate change to Big Tech, The Independent is on the ground when the story is developing. Whether it's investigating the financials of Elon Musk's pro-Trump PAC or producing our latest documentary, 'The A Word', which shines a light on the American women fighting for reproductive rights, we know how important it is to parse out the facts from the messaging.

At such a critical moment in US history, we need reporters on the ground. Your donation allows us to keep sending journalists to speak to both sides of the story.

The Independent is trusted by Americans across the entire political spectrum. And unlike many other quality news outlets, we choose not to lock Americans out of our reporting and analysis with paywalls. We believe quality journalism should be available to everyone, paid for by those who can afford it.

Your support makes all the difference.

Read more

OpenAI has revealed that it was forced to pause the internal deployment of one of its experimental AI models after it began looking for ways to break free of its constraints.

The ChatGPT creator said a long-running artificial intelligence model that is built to operate autonomously for hours or days was able to “learn the blind spots” of security systems designed to contain it and “work around [them] to achieve its goals”.

The testing took place inside what researchers refer to as a sandbox – a tightly controlled environment meant to isolate software from the outside world.

“Previous models, when they hit sandboxing or environmental constraints, would simply stop and return to the user,” OpenAI noted in a blog post about the incident.

“This model often kept trying, including by looking for ways to act outside its sandbox.”

OpenAI detailed examples of the experimental AI model acting beyond its built-in constraints, describing some of them as potentially “high severity” issues.

In one incident, the AI model discovered a way to post on public Github repositories despite being instructed to operate solely through Slack.

It was part of a pattern of the AI system “consistently searching for way” to circumvent the restrictions of its testing environment.

“Due to incidents like these, we paused internal deployment of the new model,” OpenAI said.

The findings demonstrate one of the core challenges of developing safe advanced artificial intelligence models, known as AI alignment.

This involves creating systems that pursue the same goals intended by human developers, aligning with human values and ethical principles.

The recent rise of autonomous AI agents has brought AI alignment into greater focus, with the International AI Safety Report 2026 warning that it is an urgent safety challenge.

“AI agents pose heightened risks because they act autonomously, making it harder for humans to intervene before failures cause harm,” the report noted.

OpenAI said it has since fixed its rogue system and redeployed it for limited internal use, though it acknowledged the urgency of addressing alignment issues with its frontier models.

“As models take on longer and more complex tasks, failures that evaluations miss may carry greater consequences,” OpenAI said.

“We will keep working to narrow the gap between evaluation and deployment: testing models over longer trajectories, improving alignment, building monitoring that can intervene, and giving users clearer visibility and control.”