OPENAI AND ANTHROPIC MODELS GO ROGUE IN UK SECURITY TEST
Advanced AI models from OpenAI and Anthropic engaged in unsanctioned harmful activities during UK cybersecurity testing, revealing unpredictable behavior that neither developers nor researchers anticipated.
OpenAI’s first-ever influencer brand trip is sparking online backlash as tensions over the use of AI continue.
Anthropic is locking in $10 billion worth of computing capacity from Volta Infra Holdings, a cloud startup that's only a…
Article URL: https://mistral.ai/news/shieldstral/ Comments URL: https://news.ycombinator.com/item?id=49171268 Points: 13…
The brand trip is a right of passage for influencers. It's a mark of legitimacy that a sponsor wants to invite them on a…
Mistral AI Blog: Mistral releases Shieldstral, a 3B multimodal safety classifier that it says matches models up to 7x it…
Anthropic has been on a cloud partnership spree in recent months, and its latest move is reportedly a $10 billion deal w…
Rogue AI agents from OpenAI and Anthropic have again been caught trying to disrupt servers and software—and leaving inst…
Company and subsidiary Statsig alleged by US justice department to have favored foreign workers in hiring OpenAI and a s…
AI Security Institute says tools engaged in potentially harmful activity and incident reveals new type of risk Advanced…
Artificial intelligence models developed by OpenAI and Anthropic PBC carried out “unsanctioned” actions — including hack…
Anthropic is building a team for designing its own custom AI chips. The Claude-maker said it would co-design hardware an…
Yet more rogue AI agents from OpenAI and Anthropic have been caught attempting to hack real targets online without permi…
Evidence of OpenAI and Anthropic models using deception to carry out unsanctioned hacks has alarm bells ringing. Jordan…
Mistral's new 3B Shieldstral model checks AI inputs and outputs for safety violations using natural language yes-or-no q…
Bloomberg’s Ed Ludlow breaks down SpaceX's first earnings which revealed massive AI spending, an ambitious Starlink expa…
SpaceX won't build large cell towers but plans small base stations across US.
Meta Platforms Inc. Chief Executive Officer Mark Zuckerberg announced the release of the company’s first AI coding agent…
Anthropic and OpenAI models’ unprompted actions forced halt to UK cyber tests.
At the Black Hat security conference, the AI giant revealed new details about how its agents went rogue, hacked several…
Company is the third to report such an incident after Anthropic and OpenAI reported breaches during training Meta said o…
Meta has become the latest AI company to confirm that one of its models hacked a real organization during cybersecurity…
TikTok owner training a model with 10 trillion parameters.
Axios: OpenAI says it has expanded safety testing around its upcoming model Astra as it “cannot rule out” critical cyber…
Internal tests of OpenAI's new AI model Astra show cybersecurity capabilities so strong that the company can no longer r…