FRONTIERWIRE

ИИAI · 2026-09-03 · весь выпускfull edition

OpenAI впервые признала модель «критически опасной» для кибербезопасностиOpenAI just called one of its own models a critical cybersecurity risk

Модель Astra — первая у OpenAI, преодолевшая порог «критических» киберспособностей по их же框framework.Astra is the first OpenAI model to cross the company's own "critical" cybersecurity threshold.

OpenAI сообщила, что их модель под кодовым названием Astra стала первой, которая достигла порога «критических» киберспособностей по внутреннему Preparedness Framework — документу, где компания сама описывает, какие возможности ИИ считает достаточно опасными, чтобы требовать усиленных мер защиты. В посте компания говорит о «более сильных гарантиях безопасности» перед релизом, но конкретики — какие именно способности сработали как триггер, что это значит на практике — дигест не даёт.

Важно понимать формат такого заявления. Preparedness Framework — это собственная классификация OpenAI, не независимый внешний аудит. Компания сама определяет пороги, сама их измеряет и сама решает, что считать «критическим». С одной стороны, это шаг к прозрачности: раньше про такие пороги вообще не говорили публично так прямо. С другой — верить на слово тут особо нечему, пока нет независимой проверки методологии.

Тема не случайна: вопрос кибербезопасности ИИ обсуждается всё чаще именно потому, что модели становятся способны находить уязвимости в коде и писать эксплойты на уровне, доступном раньше только специалистам. Если Astra действительно перешагнула этот порог, то компания как минимум признаёт: инструмент, который она выпускает, может быть использован для атак, а не только для защиты.

Почему это важно

Это первый публичный случай, когда крупная лаборатория сама называет свою модель «критически» опасной по одному из параметров — а не просто говорит, что она «мощная» или «продвинутая». Стоит следить, появится ли за этим заявлением независимая проверка, или всё ограничится корпоративным пресс-релизом.

OpenAI announced that a model codenamed Astra is the first of theirs to meet the "Critical" cybersecurity capability threshold under the company's own Preparedness Framework — the document where OpenAI defines what AI capabilities it considers dangerous enough to warrant stronger safeguards. The post mentions "stronger safeguards for release," but the digest doesn't spell out which specific capability triggered the threshold or what that means in practice.

Worth being clear about the format here: the Preparedness Framework is OpenAI's own classification system, not an independent external audit. The company sets its own thresholds, measures against them itself, and decides what counts as "critical." On one hand, that's more transparency than labs used to offer — this kind of threshold crossing wasn't announced this directly before. On the other, there's not much to independently verify here yet.

The topic isn't random. AI cybersecurity keeps coming up because models are getting good enough to find code vulnerabilities and write exploits at a level that used to require specialist skill. If Astra genuinely crossed this line, the company is at minimum acknowledging that the tool it's shipping could be turned toward attacks, not just defense.

Why it matters

This is the first public instance of a major lab labeling its own model "Critical" on a specific safety axis, rather than just calling it "powerful" or "advanced." Worth watching whether independent scrutiny follows, or whether this stays a self-graded press release.

Источник:Source: https://openai.com/index/path-to-astra