OpenAI Launches GPT-6 Astra With Tighter Cyber Safeguards
OpenAI has released GPT-6 Astra, described as its most intelligent and aligned model to date, with strong capabilities in computer use, software engineering, science, and cybersecurity. The model demonstrated the ability to identify and develop zero-day exploits, scoring 100% on ExploitBench and discovering two previously unknown vulnerabilities during internal evaluations. In response to these advanced cyber capabilities and lessons from the Hugging Face incident, OpenAI paused some training to implement stricter alignment protocols and safety processes. Astra now includes human oversight features that can slow, pause, or stop its work and request user review before proceeding with certain actions. The model excels in benchmarks, including 98% on FrontierMath Tier 4 and 88% on SRE-Bench for reverse-engineering tasks. Despite its power, Astra is more likely to refuse advanced cybersecurity tasks such as creating proofs-of-concept. It is rolling out to limited organizations initially and will become available on ChatGPT Plus, Pro, Business, Enterprise, the OpenAI API, and Amazon Bedrock, with pricing set at $10 per million input tokens and $50 per million output tokens. OpenAI emphasizes that while the model can delegate tasks with greater confidence, it remains imperfect, and they are continuing to refine its alignment and interruption systems.
https://www.govinfosecurity.com/openai-launches-gpt-6-astra-tighter-cyber-safeguards-a-32745
Comments
Post a Comment