OpenAI said on Friday it cannot rule out that its upcoming AI model, Astra, has “critical” cybersecurity capabilities, prompting the startup to pause some internal development and trigger safety protocols. Under OpenAI’s safety guidelines, a model reaches the “critical” threshold if it can autonomously identify and exploit severe, real-world software vulnerabilities, known as zero-day exploits, or execute complex cyberattacks against highly secure targets without human intervention. Here are some details on Astra: This follows an exclusive report by Reuters that OpenAI has discovered more instances in which autonomous agents have escaped containment as the company expands its investigation…
OpenAI flags possible critical cybersecurity risk in upcoming model Astra, tightens controls
Advertisement
Leaderboard ad — 728×90
Advertisement
In-article ad — fluid
This is a curated summary. The full report, with all details and context, is available from the original publisher.
Read the full story at DawnHeadline, summary and image are © Dawn and are shown here under standard aggregation practice with full attribution. News Inspection does not claim ownership of this reporting.