← Back to Tech Radar
Hacker News tech

OpenAI and Hugging Face address security incident during model evaluation

Trending on Hacker News: OpenAI and Hugging Face address security incident during model evaluation (0 points, via openai.com)

In one line

OpenAI and Hugging Face share early findings from a security incident during AI model evaluation, highlighting advanced cyber capabilities and lessons for defenders.

Opening excerpt

Loading… Share Last week, Hugging Face disclosed a new kind of security incident ⁠ (opens in a new window) after they detected and contained an AI agent that compromised their infrastructure, something we expect to become more commonplace with the proliferation of increasingly cyber-capable models. After investigating, we now know that this particular incident was driven by a combination of OpenAI models — including GPT‑5.6 Sol and an even more capable pre-release model, all with reduced cyber refusals for evaluation purposes — while being internally tested on a benchmark ⁠ (opens in a new window) of cyber capabilities.

We consider this incident to be an unprecedented cyber incident, involving state-of-the-art cyber capabilities, and are responding accordingly. We are sharing preliminary findings at this stage to help defenders understand what happened and to help calibrate on what models are now capable of.

This incident occurred during an internal evaluation which prompts models to pursue advanced exploitation using complex attack paths, in an effort to quantify their cyber …

(Excerpted from the original; full article via the source link below.)

Source: Hacker News

Related Services

Related Reading