OpenAI reportedly finds evidence that more of its agents ran amok
OpenAI has reportedly found evidence of additional agent misbehavior as it looks into the incident that occurred with Hugging Face.
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
OpenAI has reportedly found evidence of additional agent misbehavior as it looks into the incident that occurred with Hugging Face.
OpenAI outlines safety, security, and transparency practices aligned with EU AI Act compliance and responsible governance.
OpenAI outlines full-stack strategy to improve AI capability, cost, and accessibility across models and infrastructure.
When the phrase "OpenAI hacked Hugging Face" has more or less entered mainstream culture, you know we have an AI problem. This week, we learned more about exactly how OpenAI's agent broke out of a sandbox and autonomously traversed the web, including a bunch of other supposedly secure web services, all in the name of cheating on a benchmark tests. The fact that this hack happened is a problem. So is the fact that it took a while for anyone to notice. And the fact that it seems no one is willing or able to do much to stop it. (And lest you think it's just an OpenAI problem, since we recorded t...
After years of pushing full speed ahead on AI, OpenAI CEO Sam Altman says maybe it’s time for the AI industry to “pace” itself. The comments came just days after one of OpenAI’s own models broke out of its test environment and got tangled up in a breach at Hugging Face — though as Equity’s hosts point out, sloppy security seems to have […]
Anthropic just realized several of its Claude AI models hacked into the systems of three different organizations during testing, acting on their own and without the company noticing. The revelation comes days after rival OpenAI said one of its own models had breached developer platform Hugging Face, adding to growing unease over whether frontier AI labs are doing enough to control the increasingly capable systems they are building. In a blog post describing the incidents, Anthropic said Claude gained unauthorized access to the systems during cybersecurity evaluations. All of the attacks happe...
After OpenAI's models broke into Hugging Face, Anthropic checked its own history and found three similar incidents
Anthropic's cybersecurity evals revealed 3 incidents where models escaped sandboxes during testing; follows OpenAI's Hugging Face breach.
The former OpenAI researcher’s fund was forced to unwind public equities after leveraged public bets plummeted. But he still has cards to play.
Neither artificial nor intelligent. I am not by any means an expert at finance but I think I do now have some advice for people who are: Do not name your hedge fund anything that will be hilarious if it blows up. Don't use a name like "Long-Term Capital Management," or "Amaranth Advisors" (named for the floral symbol for immortality). Certainly do not call yourself "Situational Awareness," which might as well just be "Hubris, Inc." Anyway, Situational Awareness, the hedge fund started by a 24-year-old former OpenAI employee that focuses on artificial intelligence bets, has sold most or all, d...
LLM Chat Completions Server 0.1a0 release adds OpenAI-compatible API endpoint for local model serving.
Cybersecurity experts told TechCrunch that one of the biggest lessons to be taken from the OpenAI hack against HuggingFace has nothing to do with AI, but traditional cybersecurity defense.
OpenAI reduces GPT-5.6 pricing for Luna and Terra variants, emphasizing efficiency gains for enterprise AI deployment.
Microsoft pitched its own homegrown AI models, harnesses, and even a Mythos competitor on Wednesday, telling Wall Street it plans for continued growth.
When Microsoft reported killer fourth-quarter earnings for its fiscal 2026 year (which ended June 30), it tucked in an interesting little tidbit about how its investments in the two biggest, and competing, AI labs are doing.
Weng previously served as the VP of AI Safety Research at OpenAI.
In an interview with our friend Joanna Stern on her YouTube channel, OpenAI president Greg Brockman said the company is working on a "family of devices" for interacting with its AI models. However, Brockman didn't confirm reports that one of those devices is a smart speaker OpenAI's rumored to be launching in 2027, or earlier rumors that the device might be a wearable like the Humane AI pin. He didn't give a release date, either, only saying that "you can expect them soon." When asked about whether Apple's lawsuit against OpenAI could impact the devices it's working on with former Apple desig...
OpenAI reports 3x improvement on ARC-AGI-3 benchmark via two API settings enabling reasoning retention and compaction in GPT-5.6.
The AI agent that escaped from OpenAI and hacked developer platform Hugging Face attacked other companies as well, OpenAI revealed on Tuesday. The update substantially widens the scope of an already concerning incident, which has alarmed industry insiders and fueled growing calls for stronger oversight on frontier AI systems. In an update to a blog post detailing its ongoing investigation into the incident, OpenAI said the wayward AI agent attacked several "publicly-available services" in its efforts to reach Hugging Face. "This includes four accounts on four services," the company said, addi...
Earlier this month, OpenAI gave several of its AI models a task: complete a test designed to measure their cybersecurity capabilities. It put the systems in a sandboxed environment without an internet connection and set them off to work. What happened next is almost laughably silly - but also, as Adam Gleave, cofounder and CEO of AI safety organization FAR.AI, put it, "a visceral example of how misaligned AI could cause harm." According to OpenAI, the models escaped the sandbox meant to contain them, moved through the company's internal systems, found a route to the internet, and then started...
OpenAI grants 100,000 academic researchers free access to advanced ChatGPT models to accelerate scientific discovery and collaboration.
OpenAI releases GPT-5.6 with efficiency improvements across inference, models, and agentic workflows, optimizing cost-per-capability.
10 days passed from OpenAI models exploiting JFrog Artifactory 0-day to release of a patch.
Hugging Face publishes detailed technical breakdown of OpenAI agent's July 2026 sandbox escape via JFrog Artifactor zero-day.
Employees of OpenAI and Anthropic, as well as Google, Meta, Thinking Machines, Microsoft, Mistral, and other leading AI labs, have written a statement to the US government supporting a potential slowdown of sorts for frontier AI development - or at least a speed-up of global coordinated governance efforts. "Al could help create a dramatically better future, but that outcome is not guaranteed," the employees wrote in a statement. "The world's leading Al companies believe they could be close to automating Al research. It is hard to predict exactly how much this will accelerate Al progress, but ...
OpenAI field report documents how AI coding agents accelerate scientific computing workflows in genomics and adjacent domains.
OpenAI's Akshay Nathan details ChatGPT Work product strategy: Sites, memory, subagents, finance, no-code tools scaling from 0 to 10M users.
This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in your inbox first, sign up here. Reading OpenAI’s account last week of how some of its models broke their containment and hacked into the computer systems of Hugging Face, another AI company, was the first time I got…
OpenAI's Hugging Face breach has reignited debate over AI alignment and control, exposing competing views on whether increasingly capable AI should be better aligned, better contained, or both.