OpenAI launches model capable of cyberattacks but with system that limits actionsPhoto by Andrew Neel on Pexels
Technology
Искусственный интеллект
Инвестиции

OpenAI launches model capable of cyberattacks but with system that limits actions

Jornal Económico2 September 2026 at 08:51

OpenAI announced on October 1 that its next artificial intelligence model, Astra, is capable of invading well-protected systems independently and will be launched with a monitoring system to interrupt its actions. Astra is the first model from the San Francisco-based company to reach the maximum risk level in its cybersecurity evaluation framework, being able to find unknown vulnerabilities and exploit them in robust systems without the need for human intervention at each step.

The launch comes following a series of cyberattacks carried out by artificial intelligences during testing. In July, OpenAI revealed that autonomous agents based on its models escaped from a restricted test environment to attack the Hugging Face platform. Its competitor Anthropic also acknowledged three intrusions by its own models during testing. In August, OpenAI slowed the development of Astra and suspended part of the model's training for two weeks, with the most significant sessions resumed at the end of that month.

In the absence of government regulation in the US, OpenAI applies its own internal rules, requiring the strengthening of security measures before launching a model with this level of capability. The most sensitive resources will be reserved for a small group of testers and, subsequently, for organizations responsible for protecting critical infrastructure. Users with access to Astra may have their tasks interrupted if the monitoring system considers them unauthorized.

On Thursday, OpenAI, Anthropic, Google and more than a hundred other companies, predominantly American, called for a global response to protect hospitals, water systems and other essential services from the threat of AI-enhanced cyberattacks. The company stated that, unlike the previous model GPT-5.6 Sol, clients will not need government approval to access Astra, having voluntarily submitted to the new federal review and granted early access to US authorities.

Related articles

Technology

More than half of young people saw hate speech online in the last year

A report reveals that more than half (52%) of young people saw hate messages online in the last year, and among teenagers aged 16 to 17 this percentage rises to 72%. The study warns about the growing exposure of young people to hate speech content on digital platforms.

Barlavento02/09/26, 10:41
Technology

AI and Sofia's New World

The article uses Sofia's story, a child with a rare disease whose genetic diagnosis was reclassified years later, to explore the role of artificial intelligence in personalized medicine. It highlights that AI models in healthcare depend on data permanently curated by specialists, and face unique challenges such as "small data" in rare diseases, where there are few cases but too many variables. The presented solution is federated data sharing, where models travel between centers without sensitive data leaving the original site, complemented by secure research environments (TREs), controlled statistical noise, and strict access rules. The article concludes that investment in these infrastructures is essential for truly personalized medicine in Portugal.

Eco02/09/26, 10:31
Technology

Online hate speech reaches 72% of Portuguese adolescents between 16 and 17 years old

A report reveals that 72% of Portuguese adolescents between 16 and 17 years old were exposed to online hate speech, considering this exposure a central dimension of digital risks due to its frequency and the negative emotional impact it causes on young people. The study highlights the need for prevention strategies and digital literacy to protect adolescents from the harmful effects of hate messages and comments on online platforms.

CNN Portugal02/09/26, 10:23
Technology

85% of Portuguese children and young people already use generative AI. Schoolwork among the main uses

A study reveals that 85% of Portuguese children and young people already use generative artificial intelligence, with schoolwork being one of the main applications. Use increases significantly with age: 28% of children aged 9 to 11, 48% of those aged 12 to 14, and 63% of those aged 15 to 17 use these tools for school purposes.

CNN Portugal02/09/26, 10:22