OpenAI reveals upcoming Astra model may possess ‘critical’ hacking capabilities
OpenAI Group PBC today disclosed that one of its unreleased large language models may pose a significant cybersecurity risk.
The algorithm, which is known as Astra, was first detailed last week. OpenAI revealed in a Sunday blog post that the LLM had solved 10 long-running math problems. The company published the proofs and revealed that each one took about $2,000 worth of tokens to generate.
As part of its artificial intelligence safety efforts, OpenAI has published a 29-page document known as the Preparedness Framework. One of the document’s sections contains a rating system for AI risks. The system ranks the cybersecurity risks posed by an LLM as “High” or “Critical” depending on its capabilities.
OpenAI’s flagship GPT-5.6 Sol model and a few earlier algorithms were given a High rating. According to the company, Astra is the first of its LLMs that may qualify for a Critical designation. Its engineers drew that conclusion based on a series of recent cybersecurity tests.










