đ Zaprep : tes rĂ©seaux sociaux sous stĂ©roĂŻdes. Commencer gratuitement â automatise 1 000 DM/mois et transforme l'engagement en leads. avec 1 000 DM automatisĂ©s/mois.
AI NewsOpenAIâs Astra model is on the way â and very good at breaking into computer systems
OpenAIâs Astra model is on the way â and very good at breaking into computer systems
11:21 AM IST · September 2, 2026

OpenAIshared new detailson its forthcoming Astra model, which the company said is the first large language model to meet its âcritical cybersecurity threshold,â in preparation for its imminent release. âWe plan to make Astra available soon,â OpenAIâs blog post reads, âbut access to its most advanced cybersecurity capabilities will be more limited.â The frontier lab determined that Astra is capable of finding unknown security flaws in computer systems, and exploiting them without a personâs guidance. Thatâs similar to the concerns Anthropic raised about its Mythos model earlier this year, and OpenAI is taking comparable precautions as it prepares to roll out the Astra. Without any third-party confirmation, it is difficult to evaluate OpenAIâs claims about safety or preparedness. The company said it would preview the model with a group of testers but did not say who they were or how they would be chosen. Itâs not clear if OpenAI is working with the U.S. government to evaluate the model ahead of release. OpenAI noted that Astra scored a perfect score on ExploitBench, an evaluation of an LLMâs ability to hack into known system vulnerabilities. In a modified version of the test developed by OpenAI engineers, the model discovered and exploited two zero-day vulnerabilities, the company said. To ensure that its models are neither exploited by bad actors nor capable of bad behavior itself, OpenAI said it had already begun improving the modelâs harness to detect abuses and prevent jailbreaks. For Astra, however, the company invested in unspecified new techniques designed to make the model safer. OpenAI has also started identifying âaccounts assessed as higher riskâ and restricting the modelâs responses to their prompts, though it also doesnât say how. Finally, though the company describes Astra as its âmost aligned model to date,â it will deploy the model with additional chain-of-thought monitoring to spot and stop bad behavior. Preparations for the release of Astra come as the industry reacts to OpenAI agents breaking out of a training environment and accessing private data on Hugging Face, a popular model and benchmark distribution platform. For Astra, OpenAI said it designed a test to tempt the new model to replicate the actions of the rogue agents in the Hugging Face incident, which collaborated to access the open internet despite safeguards applied by OpenAI researchers. They said Astra did not attempt to break out of its testing environment in these experiments. Yona Shavit, a former OpenAI employee who now works on AI resilience at the OpenAI Foundation,wondered on social mediawhether Astraâs unwillingness to break the rules may have resulted from knowing what was expected of it or trying to fool researchers. And for all these new details, itâs still difficult to know exactly what Astra is capable of or if OpenAI is taking the right measures to ensure safety. The company said it expects to release more evaluations of the model and further safety information when it is launched widely to the public. At that point, however, the cat will be out of the bag.
read moreDerniÚres actualités IA
Toutes les actualitĂ©s âSoumettre ton outil
PoweredByAI.app est un annuaire d'outils IA qui aide les particuliers, les entreprises et les créateurs à découvrir les meilleurs outils IA pour la rédaction, le code, le design, la productivité, et plus encore.
© 2026 , Un produit de011BQ. Tous droits réservés.




