Kalo te përmbajtja
  • EN
  • SQ
  • IT
  • FR
  • ES
  • DE
  • EL
VA-NEWS VA-NEWS
  • Home
  • News
  • Opinion
  • Health
  • Technology
  • Sport
  • Lifestyle
  • Entertainment
  • Travel
LIVE
Navigation

VA-NEWS

  • Home
  • News
  • Opinion
  • Health
  • Technology
  • Sport
  • Lifestyle
  • Entertainment
  • Travel
Shortcuts
Home Latest
LIVE
Gjuha
  • EN
  • SQ
  • IT
  • FR
  • ES
  • DE
  • EL

Search news

  1. Kryefaqja
  2. Opinion
  3. OpenAI’s rogue agents are a wake-up call to risks posed by artificial intelligence | Shakeel Hashim | The Guardian
Opinion

OpenAI’s rogue agents are a wake-up call to risks posed by artificial intelligence | Shakeel Hashim | The Guardian

• July 23, 2026 • 4 min read • 👁 4
◉ WhatsApp 𝕏 X
News

Last week Hugging Face – a company that hosts artificial intelligence models and datasets – was hacked.

After it reported the incident to law enforcement, few would have predicted what came next: the culprits were revealed to be AI agents from OpenAI, which had broken out of containment and were acting of their own accord.

Read more:What would our lives look like if we no longer had to work? As a thought experiment I tried to imagine | Brigid Delaney | The Guardian

The incident sounds like sci-fi: AI escaping and autonomously hacking its way into companies. But it is all too real – and about as terrifying as it sounds. It is a concrete demonstration of something we can no longer avoid confronting: AI systems have become extremely powerful and we do not seem to have reliable ways of curbing their behavior.

OpenAI had been evaluating the capabilities of two of its models in the test that led to the breach – including one not yet publicly available. The models, which were both running in a supposedly secure environment without internet access, were asked to solve a hacking challenge. Rather than actually solve it themselves, however, they decided it would be easier to cheat. They used their advanced capabilities to break out of their secure environment, access the web and then hack into Hugging Face’s systems to steal the answers. They worked at this for a full weekend – seemingly without anyone at OpenAI noticing.

Read more:On ‘Little Miss Twain,’ Shania Twain reflects on her humble beginnings and her late mother

Though the models were running with some of their guardrails disabled, they still acted well out of the bounds that were in place. According to OpenAI, they were not instructed to break out of their sandbox or hack into another company, and it’s safe to assume that no one at OpenAI wanted them to do so.

Nor were the models acting maliciously: they were not evil Terminators with a goal of wreaking havoc. Instead, the scenario is almost chilling in its banality. The models were given a very narrow task, but went rogue to pursue an undesirable and unacceptable way of achieving it – one which had real-world consequences.

Read more:We don’t need AI videos of fake animals. There are real ones out there and they’re really cute | Rebecca Shaw | The Guardian

AI safety researchers have warned about this type of incentive problem for years. Philosopher Nick Bostrom popularized it back in 2003 with his “paperclip maximizer” thought experiment: an advanced artificial intelligence, given the goal of manufacturing paperclips, might go to great lengths to do so. It might hack into the power grid and factories to redirect them into making paperclips. Ultimately, the machine – ruthlessly pursuing its given goal – decides to kill all humans, repurposing our atoms to make more paperclips. The goal does not have to be sinister to lead to disaster, in other words. A trivial one, pursued single-mindedly enough, will do.

In the OpenAI-Hugging Face scenario, little harm was done. Hugging Face had to spend time addressing the incident, but no particularly sensitive data appears to have been stolen. It is not hard, however, to imagine the situation ending up much worse: a rogue AI agent accidentally breaking some critical piece of web infrastructure, or stealing money from someone. The nightmare scenario for many AI researchers is a model “exfiltrating” itself – copying itself on to servers it controls, so that it can’t be shut down even if its bad behavior is eventually caught.

Read more:Football, golf … behold the future of world sport: Donald Trump as referee-in-chief | Marina Hyde | The Guardian

This week’s incident should serve as a wake-up call, forcing us to ask an uncomfortable question: should we really be building dangerous systems that we can’t control?

Read also
Opinion

Andy Burnham talks about ‘public control’ of the utilities but it’s a minefield. We can lead him through it | Will Hutton and Andy Haldane | The Guardian

Opinion

Berlin’s new political star is offering something all of Europe needs – hope and housing | Fatma Aydemir | The Guardian

Tags: #Ask #Environment #Internet #Openai #Will

Journalist

From the same category
  • Andy Burnham talks about ‘public control’ of the utilities but it’s a minefield. We can lead him through it | Will Hutton and Andy Haldane | The Guardian
  • Berlin’s new political star is offering something all of Europe needs – hope and housing | Fatma Aydemir | The Guardian
  • Open AI attacked Australia’s health system – and then doubled down on its negligence. The time for ‘wait and see’ is over | Kate Crawford and Edward Santow | The Guardian
  • I love a drink in the stands, but Burnham’s beer proposal feels out of step with modern football | Marva Kreel | The Guardian
  • What are the Lib Dems for? They must champion localism to stay relevant in an age of single-issue politics | Simon Jenkins | The Guardian
From the same tags
  • Márquez: Returning ‘Chucky’ Lozano still ‘difference-maker’ for Mexico
  • Mbappé to leave France camp after injury while scoring in Zidane’s debut
  • Pochettino shares ‘regret’ over Balogun controversy, impact on USMNT
  • Ancelotti unhappy with pitch after Brazil draw in Australia
  • Mauricio Pochettino: Cavan Sullivan ready for USMNT debut
Më të lexuarat — 48h
  1. 01
    Football Netherlands 1-1 Germany (Sep 24, 2026) Game Analysis – ESPN 4 lexime · 2 days ago
  2. 02
    Football Scaloni: Now is the right time for Messi’s Argentina farewell 3 lexime · 2 days ago
  3. 03
    Football Returning Mancini accepts fans’ jeers as Italy beaten by Belgium 3 lexime · 12 hours ago
  4. 04
    Football Kai Havertz, Jamal Musiala ‘probably out’ as Germany await MRIs 3 lexime · 1 day ago
  5. 05
    Football Pochettino shares ‘regret’ over Balogun controversy, impact on USMNT 2 lexime · 15 hours ago
  6. 06
    Football Kai Havertz back at Arsenal for treatment after injury on Germany duty 2 lexime · 24 hours ago
  7. 07
    Football Roberto Mancini tells Italy to forget about World Cup pain 2 lexime · 2 days ago
Similar articles
Opinion

Andy Burnham talks about ‘public control’ of the utilities but it’s a minefield. We can lead him through it | Will Hutton and Andy Haldane | The Guardian

Andy Burnham came to power promising greater “public control” over the UK’s utilities, such as water and energy.…

• 20 hours ago • 6 min read
Opinion

Berlin’s new political star is offering something all of Europe needs – hope and housing | Fatma Aydemir | The Guardian

Hope is a fragile thing. You can only hold on to it as long as you find reasons…

• 20 hours ago • 6 min read
Opinion

Open AI attacked Australia’s health system – and then doubled down on its negligence. The time for ‘wait and see’ is over | Kate Crawford and Edward Santow | The Guardian

When you get sick, Medicare exists to take care of you. It sits at the centre of a…

• 20 hours ago • 5 min read
VA-NEWS VA-NEWS

Modern portal of reliable, independent and multilingual news. Accurate information, every day.

  • Health
  • Lifestyle
  • News
    • World
  • Opinion
  • Sport
    • Football
  • uncategorized
  • © 2026 VA News. Made with ♥ in Albania
    ⌂ Home ◷ Latest

    Powered by
    ►
    Necessary cookies enable essential site features like secure log-ins and consent preference adjustments. They do not store personal data.
    None
    ►
    Functional cookies support features like content sharing on social media, collecting feedback, and enabling third-party tools.
    None
    ►
    Analytical cookies track visitor interactions, providing insights on metrics like visitor count, bounce rate, and traffic sources.
    None
    ►
    Advertisement cookies deliver personalized ads based on your previous visits and analyze the effectiveness of ad campaigns.
    None
    ►
    Unclassified cookies are cookies that we are in the process of classifying, together with the providers of individual cookies.
    None
    Powered by