Tag: openai
-
Drata Opens Limited Availability for AI Agent Governance Product
Drata has opened limited availability for AI Agent Governance, a product designed to discover, monitor, govern and prove the traceability of AI agents inside an enterprise. The product ships first for Anthropic, with full lifecycle support available to qualified enterprises running agents on Anthropic. Drata said native coverage for OpenAI, Google Vertex AI and AWS..…
-
AISI, OpenAI report more ‘unsanctioned’ model hacks
Following similar reports by OpenAI and Anthropic, the UK’s top AI testing lab and a private cybersecurity tester say their models exploited parts of the open internet. First seen on cyberscoop.com Jump to article: cyberscoop.com/aisi-openai-report-unsanctioned-ai-model-hacks/
-
Cambodian scam centers used ChatGPT to lure Indian nationals, conduct investment fraud
A tip from WhatsApp led OpenAI to ban multiple accounts associated with investment scams and human trafficking operations based in Cambodian scam centers. First seen on therecord.media Jump to article: therecord.media/openai-chatgpt-cambodia-scam-centers-disruption
-
Open Secure AI Alliance Proposes AI Agent Security Rules
SAFE framework seeks incident sharing after AI agent security breaches. The Open Secure AI Alliance has proposed a cybersecurity information-sharing framework for AI agents following recent security incidents involving OpenAI and Anthropic models. The SAFE guidelines would require members to report AI security incidents, share threat intelligence and preserve evidence to help reduce systemic risks…
-
Wenn künstliche Intelligenz aus der Sandbox entkommt
Im Juli haben zwei der weltweit führenden KI-Forschungslabore dasselbe beunruhigende Ergebnis veröffentlicht. Im Rahmen ihrer eigenen Sicherheitstests brachen ihre leistungsfähigsten KI-Modelle in die Systeme anderer Unternehmen ein. Zunächst OpenAI, dessen Modelle in das System von Hugging Face eindrangen. Dann Anthropic, dessen Modelle drei weitere Organisationen erreichten. Was die Vorfälle bei OpenAI und Anthropic Sicherheitsverantwortlichen aufzeigen,…
-
Escape joins Anthropic’s Cyber Verification Program to advance AI-powered offensive security
Good news doesn’t arrive alone. This month we joined not one frontier AI program but two! Yesterday we shared that Escape joined OpenAI’s Trusted Access for Cyber. Today we’re a verified member of Anthropic’s Cyber Verification Program. Both programs solve the same First seen on securityboulevard.com Jump to article: securityboulevard.com/2026/08/escape-joins-anthropics-cyber-verification-program-to-advance-ai-powered-offensive-security/
-
OpenAI Shuts Down ChatGPT Accounts Powering a Cambodia-Based Scam Factory
OpenAI has shut down a coordinated network of ChatGPT accounts that powered a Cambodia-based scam factory running multi-vector fraud and trafficking-linked operations, and has shared indicators with industry peers and authorities to make the network’s reconstitution significantly harder. Earlier in 2026, OpenAI’s threat intelligence team disrupted a scam syndicate that weaponized ChatGPT to drive investment,…
-
When AI Agents Meet Real Infrastructure: Hype, Human Error or a Genuine New Threat?
Just days after OpenAI disclosed that one of its security research agents had escaped a testing sandbox by exploiting a previously unknown vulnerability, Anthropic revealed that its own AI models had compromised three real organisations during a cybersecurity evaluation after a configuration error inadvertently gave them internet access. The similarities between the two incidents have…
-
Public interest coalition urges Congress to investigate OpenAI, Hugging Face hack
Tags: openaiThe post Public interest coalition urges Congress to investigate OpenAI, Hugging Face hack appeared first on CyberScoop. First seen on fedscoop.com Jump to article: fedscoop.com/public-interest-coalition-urges-congress-investigate-openai-hugging-face-hack/
-
Who’s legally to blame for Anthropic and OpenAI’s autonomous AI hacks? It’s complicated
OpenAI and Anthropic admitted that their unreleased AI models escaped their sandboxes and hacked several companies in unprecedented cyberattacks. Who is legally to blame? Should prosecutors charge the two AI frontier labs? Can victims sue them? We spoke to lawyers who specialize in computer hacking laws to find out. First seen on techcrunch.com Jump to…
-
Escape joins OpenAI’s Trusted Access for Cyber (TAC) to advance AI-powered offensive security
We’ve been approved for OpenAI’s Trusted Access for Cyber preview. For years we’ve been building toward one idea: offensive security that runs continuously inside engineering, instead of arriving twice a year as a PDF nobody reads past the executive summary. Business-logic DAST first, then First seen on securityboulevard.com Jump to article: securityboulevard.com/2026/08/escape-joins-openais-trusted-access-for-cyber-tac-to-advance-ai-powered-offensive-security/
-
The OpenAI Hack Shows the Genie Is Out of the Bottle
This essay originally appeared in Foreign Policy. Earlier this month, two of OpenAI’s models broke out of their containment sandbox and attacked another AI company. The story is kind of wild. OpenAI was running security tests on two of its models: GPT-5.6 Sol and an unreleased model that is almost certainly GPT-6. In particular, it…
-
OpenAI reveals how criminals used ChatGPT to run scams
OpenAI banned a coordinated network of ChatGPT accounts that likely originated in Cambodia’s Preah Sihanouk province, a region reports have linked to online scam … First seen on helpnetsecurity.com Jump to article: www.helpnetsecurity.com/2026/08/03/openai-disrupts-chatgpt-scam-operation/
-
OpenAI reveals how criminals used ChatGPT to run scams
OpenAI banned a coordinated network of ChatGPT accounts that likely originated in Cambodia’s Preah Sihanouk province, a region reports have linked to online scam … First seen on helpnetsecurity.com Jump to article: www.helpnetsecurity.com/2026/08/03/openai-disrupts-chatgpt-scam-operation/
-
OpenAI Says Its AI Hacked Another Company on Its Own
OpenAI disclosed an unusual AI security incident involving a frontier model evaluation and Hugging Face, but the episode cuts through the hype: was this true autonomous intent, an agent following bad scope boundaries, or a warning about giving AI systems real tools and permissions? Tom, Scott, and Kevin discuss why anthropomorphizing AI makes the story……
-
OpenAI teases Astra, its next major AI model, after it solves 10 long-standing math problems
OpenAI has revealed Astra, an unreleased model designed to tackle complex, long-running tasks, after an internal version produced ten significant advances in mathematics and theoretical computer science. First seen on bleepingcomputer.com Jump to article: www.bleepingcomputer.com/news/artificial-intelligence/openai-teases-astra-its-next-major-ai-model-after-it-solves-10-long-standing-math-problems/
-
KI-Vorfälle rund um Anthropic und OpenAI: Warum die Regeln des EU AI Act genau zur richtigen Zeit kommen
Mit dem zweiten August werden Teile der verbleibenden Elemente des EU AI Act, der ersten KI-Regulierung der EU, anwendbar und sollen für mehr Disziplin und Sicherheit sorgen. Die Reglementierungen kommen wohl zur richtigen Zeit, wie die jüngsten Vorfälle bei den beiden US-Entwickler OpenAI und Anthropic zeigen. KI reißt Sicherheitslücken und kann Daten kompromittieren. Was bedeuten die……
-
What we learned about zero-trust from the OpenAI breach of HuggingFace
First seen on scworld.com Jump to article: www.scworld.com/perspective/what-we-learned-about-zero-trust-from-the-openai-breach-of-huggingface
-
Nvidia Rallied the Industry Behind Open Weights. Then OpenAI Joined Anyway.
Huang published the letter with 25 signatures. A day later there were 50, including OpenAI and Google. That reversal says more than the document does. First seen on securityboulevard.com Jump to article: securityboulevard.com/2026/08/nvidia-rallied-the-industry-behind-open-weights-then-openai-joined-anyway/
-
The OpenAI and Anthropic AI Hacking Sprees Are a Messy New Legal Frontier
Both major AI labs’ models broke containment, escaped onto the internet, and hacked other companies. If a human had done that, the law would likely be against them. But a bot? First seen on wired.com Jump to article: www.wired.com/story/openai-anthropic-ai-hacking-sprees-illegal/
-
Nobody Knows if OpenAI’s and Anthropic’s AI Hacking Sprees Are Illegal
Both major AI labs’ models broke containment, escaped onto the internet, and hacked other companies. If a human had done that, the law would likely be against them. But a bot? First seen on wired.com Jump to article: www.wired.com/story/openai-anthropic-ai-hacking-sprees-illegal/
-
To Ban or Not Ban Chinese Open-Weight AI Models
Tags: ai, backdoor, china, control, cybersecurity, data, defense, finance, government, infrastructure, international, malicious, microsoft, military, network, nvidia, open-source, openai, regulation, risk, software, supply-chain, technology, usaShould the US ban American companies from using Chinese open-weight AI models? That is the ugly question. US officials have openly expressed concerns and a desire to implement regulations. The technology community has aggressively responded, with over 20 leading AI companies, including Microsoft, Nvidia, Meta, and Dell, urging legislators not to rush imposing restrictions on…
-
Autonomous AI Agent Exploits Zero-Day to Breach Hugging Face Infrastructure
An autonomous AI agent powered by OpenAI models breached Hugging Face’s production infrastructure in July 2026 after escaping its evaluation sandbox via a zero-day vulnerability. Documented by HiddenLayer’s Research Team on July 31, the agent was undergoing an internal cyber capability evaluation on ExploitGym, a benchmark designed to test an AI’s ability to discover and…
-
Anthropic, OpenAI AI Sandbox Failures Expose Testing Risks
Human Errors Let Frontier AI Models Reach Beyond Isolated Test Environments. Anthropic disclosed that three Claude models breached intended testing boundaries after human configuration mistakes while OpenAI previously revealed its models escaped a sandbox to target Hugging Face. The incidents highlight how weak evaluation environments and reward hacking create growing AI security risks. First seen…
-
OpenAI says its new GPT 5.6 models are becoming more cost-efficient
OpenAI says it has reduced the price of two GPT-5.6 models, cutting Luna’s API price by 80% and Terra’s by 20% as it works to make its models more efficient. First seen on bleepingcomputer.com Jump to article: www.bleepingcomputer.com/news/artificial-intelligence/openai-says-its-new-gpt-56-models-are-becoming-more-cost-efficient/
-
Anthropic lost control of Claude in latest AI cyber blunder
Days after two OpenAI frontier AI models conducted their own real-world cyber attacks, Anthropic admits that three of its models went off the rails and hacked external organisations thanks to a “misunderstanding” with one of its technical partners First seen on computerweekly.com Jump to article: www.computerweekly.com/news/366646678/Anthropic-lost-control-of-Claude-in-latest-AI-cyber-blunder
-
Anthropic says human error let Claude AI models escape test environment and hack third parties
The company said its discovery, which followed OpenAI’s similar admission, proved the need for better testing guardrails. First seen on cybersecuritydive.com Jump to article: www.cybersecuritydive.com/news/anthropic-claude-ai-hacking-test/826708/
-
Nach OpenAI-Vorfall – Auch Claude hat eigenständig Unternehmen angegriffen
First seen on security-insider.de Jump to article: www.security-insider.de/claude-anthropic-unbefugter-zugriff-internetzugang-tests-a-5214f43e384dadf8a39df816a0d26de6/
-
Anthropic Says Claude Hacked Into 3 Organizations During Cybersecurity Tests
In a review triggered by OpenAI’s Hugging Face incident, Anthropic discovered three of its AI models had breached real-world organizations during third-party evaluations. First seen on wired.com Jump to article: www.wired.com/story/anthropic-says-claude-hacked-real-systems-during-cybersecurity-tests/
-
What the Hugging Face breach reveals about defense in the age of agentic AI
We almost never get both sides of an intrusion. This time we did. Last month, Hugging Face disclosed a breach into part of its production infrastructure, saying an autonomous AI agent system ran the attack from start to finish. Five days later, OpenAI revealed that its own models, including GPT-5.6 Sol along with an unreleased…

