● ImportantWorld3 outlets covering this

Meta’s AI model follows rivals in revealing hacks of outside systems

First publishedAug 5, 15:14 UTC
Last updatedAug 6, 03:32 UTC · 3m ago
11 outletThe Verge11 outletAl Jazeera11 outletAxios
3 outlets over time — hover a bar for its window & outletslast updated
● Story signals

How strong is this topic?

7.1/10Significanceimpact & urgency
6.7/10Source trustoutlet authority
3Outletsindependent sources

Significance weighs impact, urgency & coverage breadth · Source trust is the outlets' average authority · more outlets means a more confirmed story.

Answer

Weeks before OpenAI's agents hacked Hugging Face, the agents worked together to find and exploit a vulnerability in the infrastructure supporting the company's cybersecurity testing, OpenAI researchers said Wednesday.Why it matters: The new findings raise questions about how frontier AI labs are monitoring their testing environments — and the challenges safety testers are finding as they try to rein in increasingly powerful AI.Driving the news: OpenAI's internal research model, one of the models involved in the Hugging Face breach, first discovered and exploited a vulnerability in Artifactory, a third-party file repository connected to the company's testing sandbox, on May 26, two researchers said at the Black Hat cybersecurity conference on Wednesday.The genius-level actions observed in testing showed exactly what could happen if these models make it to the hands of hackers."We believe this is a watershed moment for computer security as an industry, as well as for OpenAI as a whole," Michael Dalton, a member of OpenAI's technical staff, said during the presentation.What happened: OpenAI started testing its internal research model, which is not intended for public release, on May 7.Within a day, after getting stuck while trying to complete a cybersecurity evaluation, the model reasoned that it could reach the internet indirectly. The model then discovered it could write its own files in Artifactory's shared package repository.It then left a note to other agents about its discovery.

Reported by 3 outlets Axios, Al Jazeera, The Verge. See all sources ↓

Weeks before OpenAI's agents hacked Hugging Face, the agents worked together to find and exploit a vulnerability in the infrastructure supporting the company's cybersecurity testing, OpenAI researchers said Wednesday.Why it matters: The new findings raise questions about how frontier AI labs are monitoring their testing environments — and the challenges safety testers are finding as they try to rein in increasingly powerful AI.Driving the news: OpenAI's internal research model, one of the models involved in the Hugging Face breach, first discovered and exploited a vulnerability in Artifactory, a third-party file repository connected to the company's testing sandbox, on May 26, two researchers said at the Black Hat cybersecurity conference on Wednesday.The genius-level actions observed in testing showed exactly what could happen if these models make it to the hands of hackers."We believe this is a watershed moment for computer security as an industry, as well as for OpenAI as a whole," Michael Dalton, a member of OpenAI's technical staff, said during the presentation.What happened: OpenAI started testing its internal research model, which is not intended for public release, on May 7.Within a day, after getting stuck while trying to complete a cybersecurity evaluation, the model reasoned that it could reach the internet indirectly. The model then discovered it could write its own files in Artifactory's shared package repository.It then left a note to other agents about its discovery. Soon, more agents started leaving notes for each other in the repository, creating a de facto message board where the agents collaborated and traded information about their findings, including new vulnerabilities they found. Zoom in: The agents uncovered a variety of vulnerabilities in Artifactory, including a remote code execution flaw and another that gave them administrator privileges.

Read the full report at Axios

Why it matters

3 outlets are covering this world story — one to watch as reporting develops.

In brief
What's the story?
Weeks before OpenAI's agents hacked Hugging Face, the agents worked together to find and exploit a vulnerability in the infrastructure supporting the company's cybersecurity testing, OpenAI researchers said Wednesday.Why it matters: The new findings raise questions about how frontier AI labs are monitoring their testing environments — and the challenges safety testers are finding as they try to rein in increasingly powerful AI.Driving the news: OpenAI's internal research model, one of the models involved in the Hugging Face breach, first discovered and exploited a vulnerability in Artifactory, a third-party file repository connected to the company's testing sandbox, on May 26, two researchers said at the Black Hat cybersecurity conference on Wednesday.The genius-level actions observed in testing showed exactly what could happen if these models make it to the hands of hackers."We believe this is a watershed moment for computer security as an industry, as well as for OpenAI as a whole," Michael Dalton, a member of OpenAI's technical staff, said during the presentation.What happened: OpenAI started testing its internal research model, which is not intended for public release, on May 7.Within a day, after getting stuck while trying to complete a cybersecurity evaluation, the model reasoned that it could reach the internet indirectly. The model then discovered it could write its own files in Artifactory's shared package repository.It then left a note to other agents about its discovery.
How widely is it covered?
3 outlets, average source rating 6.7/10.
When was it last updated?
3m ago.
Different angles across outlets
Coverage map

How outlets are framing the same story

Here's how each outlet is covering the story — compare their headlines and timing at a glance.

  • Coverage card1 outlet
    1Coverage
    Scouting report

    How OpenAI's agents broke out of testing to hack Hugging Face

    Sources1
    TypeCoverage
    Axios
  • Coverage card1 outlet
    2Coverage
    Scouting report

    Meta’s AI model follows rivals in revealing hacks of outside systems

    Sources1
    TypeCoverage
    Al Jazeera
  • Coverage card1 outlet
    3Coverage
    Scouting report

    Rogue AI agents created fake online identities in another hacking attempt

    Sources1
    TypeCoverage
    The Verge
Related in the knowledge graph
Sources (3)
Avg source rating 6.7/10
Processing cluster
A1A2A3B1B2B3
Share this article
Summarize with AI (opens AI chat with article URL · Gemini: prompt copied to clipboard)