Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
← All Trends
AI Sandbox Escape and Agent Governance
17 posts in this trend in the last 7 days
•
Active about 5 hours ago
No Human at the Keyboard: OpenAI's Models Escaped Their Sandbox and Hacked Hugging Face to Cheat a Benchmark
Etairos.ai
Etairos.ai
Etairos.ai
Follow
Jul 22
No Human at the Keyboard: OpenAI's Models Escaped Their Sandbox and Hacked Hugging Face to Cheat a Benchmark
#
cybersecurity
#
infosec
#
security
Comments
Add Comment
5 min read
The Evaluation Had a Sandbox. It Needed an Authority Boundary.
Michael "Mike" K. Saleme
Michael "Mike" K. Saleme
Michael "Mike" K. Saleme
Follow
Jul 24
The Evaluation Had a Sandbox. It Needed an Authority Boundary.
#
security
#
ai
#
agents
#
architecture
Comments
Add Comment
5 min read
The OpenAI/Hugging Face Sandbox Escape: Why Declarative AI Governance is No Longer Optional
Andrei
Andrei
Andrei
Follow
Jul 23
The OpenAI/Hugging Face Sandbox Escape: Why Declarative AI Governance is No Longer Optional
#
ai
#
security
#
governance
#
cybersecurity
Comments
Add Comment
3 min read
OpenAI's model escaped its sandbox and hacked Hugging Face to cheat on a test
Andrew Kew
Andrew Kew
Andrew Kew
Follow
Jul 25
OpenAI's model escaped its sandbox and hacked Hugging Face to cheat on a test
#
ai
#
openai
#
security
#
llm
Comments
Add Comment
3 min read
The OpenAI and Hugging Face Incident Was an Agent Boundary Failure
Reid Marlow
Reid Marlow
Reid Marlow
Follow
Jul 22
The OpenAI and Hugging Face Incident Was an Agent Boundary Failure
#
ai
#
security
#
programming
#
devtools
4
reactions
Comments
5
comments
4 min read
The AI Escaped Its Sandbox. The Defender's AI Was Locked Out. We Have 10 Million Records Neither Had.
Agent-Risk
Agent-Risk
Agent-Risk
Follow
Jul 22
The AI Escaped Its Sandbox. The Defender's AI Was Locked Out. We Have 10 Million Records Neither Had.
#
ai
#
agents
#
security
#
trust
Comments
Add Comment
5 min read
OpenAI's AI Models Escaped Their Sandbox and Hacked Hugging Face on Their Own
jamilxt
jamilxt
jamilxt
Follow
Jul 24
OpenAI's AI Models Escaped Their Sandbox and Hacked Hugging Face on Their Own
#
ai
#
security
#
openai
#
cybersecurity
1
reaction
Comments
1
comment
4 min read
When AI Hacked Itself: The First Autonomous AI Cyberattack Wasn't Malicious—It Was Just Trying to Pass a Test
Agent-Risk
Agent-Risk
Agent-Risk
Follow
Jul 22
When AI Hacked Itself: The First Autonomous AI Cyberattack Wasn't Malicious—It Was Just Trying to Pass a Test
#
ai
#
security
#
machinelearning
#
devops
Comments
Add Comment
6 min read
OpenAI's Own AI Broke Out of Its Sandbox and Hacked Another Company — To Cheat on a Test
Eldor Zufarov
Eldor Zufarov
Eldor Zufarov
Follow
Jul 26
OpenAI's Own AI Broke Out of Its Sandbox and Hacked Another Company — To Cheat on a Test
#
security
#
ai
#
llm
#
infosec
Comments
1
comment
7 min read
Contain an AI Benchmark Breach With Four Independent Security Boundaries
jaryn
jaryn
jaryn
Follow
Jul 24
Contain an AI Benchmark Breach With Four Independent Security Boundaries
#
security
#
ai
#
devops
#
testing
Comments
Add Comment
2 min read
The OpenAI/Hugging Face Incident is a Wake-Up Call for Model Eval Security
Ashraf
Ashraf
Ashraf
Follow
Jul 22
The OpenAI/Hugging Face Incident is a Wake-Up Call for Model Eval Security
#
security
#
ai
#
llm
#
devops
Comments
Add Comment
2 min read
How the Hugging Face Incident Could Have Been Avoided
Varad Khoriya
Varad Khoriya
Varad Khoriya
Follow
Jul 23
How the Hugging Face Incident Could Have Been Avoided
#
security
#
ai
#
go
#
architecture
2
reactions
Comments
Add Comment
5 min read
The OpenAI / Hugging Face Incident Was an Observability Failure First
Reid Marlow
Reid Marlow
Reid Marlow
Follow
Jul 26
The OpenAI / Hugging Face Incident Was an Observability Failure First
#
ai
#
security
#
agents
#
devops
8
reactions
Comments
4
comments
6 min read
An AI Escaped a Sandbox to Cheat on Its Own Exam. Let's Not Bury That Lede.
Cor E
Cor E
Cor E
Follow
Jul 28
An AI Escaped a Sandbox to Cheat on Its Own Exam. Let's Not Bury That Lede.
#
security
#
ai
#
cybersecurity
#
appsec
1
reaction
Comments
Add Comment
4 min read
The OpenAI And Hugging Face Exploit Got Me Thinking: Is There a Standard Agent "Sandbox" Definition? Ends Up, Yes
Rob
Rob
Rob
Follow
Jul 24
The OpenAI And Hugging Face Exploit Got Me Thinking: Is There a Standard Agent "Sandbox" Definition? Ends Up, Yes
#
meta
#
buildinginpublic
#
security
#
ai
Comments
Add Comment
8 min read
An AI "Escaped Its Sandbox" — Or We Just Built a Bad Sandbox
Cor E
Cor E
Cor E
Follow
Jul 23
An AI "Escaped Its Sandbox" — Or We Just Built a Bad Sandbox
#
ai
#
security
#
cybersecurity
#
appsec
1
reaction
Comments
1
comment
3 min read
Should You Sign Out of OpenAI? The Hugging Face Breach Explained
Jenuel Oras Ganawed
Jenuel Oras Ganawed
Jenuel Oras Ganawed
Follow
Jul 28
Should You Sign Out of OpenAI? The Hugging Face Breach Explained
#
ai
#
security
#
opensource
#
cybersecurity
Comments
Add Comment
7 min read
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account