TechnoDG Logo

OpenAI’s Upcoming Model Astra Raises New Safety Concerns

Posted by TechnoDG on 3 hour(s) ago .

OpenAI is getting ready to release a new AI model called Astra, but the company says it needs stronger safety measures before making it widely available. The main reason for this concern is Astra’s ability to find software security problems. It is much better at this than OpenAI’s most advanced model that is currently available to the public. Astra can find security weaknesses that were not known before and can also figure out possible ways to exploit them with very little help from people. Another worrying thing is that Astra can do these tasks while using less computing power than older AI models.

 

Amelia Glaese, who is an OpenAI vice president who oversees safety work said, "With the right tools and access, Astra can find previously unknown security flaws and develop ways to exploit them across many well protected system without a person guiding each step." This ability has raised concerns because a powerful AI that can find and exploit security weaknesses on its own could be misused. Because of these abilities, Astra has become the first OpenAI model to reach the "Critical" risk level under the company’s Preparedness Framework. Until now, this level had only been a theoretical possibility.

 

One of the biggest improvements in Astra is its cybersecurity performance. Its ability to find recent software vulnerabilities increased from 11.5% to 39%. This shows that Astra is much better at finding security problems than earlier models. OpenAI plans to make Astra available to a limited number of users soon. However, the company has not shared the exact release date or clearly said who will get access to it. The first release is expected to be limited so that OpenAI can closely watch how the model works and behaves. OpenAI has also added stronger safety measures to reduce the risks linked to Astra. These include filters that are meant to stop the model from answering harmful cybersecurity requests. The company will also keep a close watch on Astra to see if it tries to get around these safety measures.

 

However, these extra protections could sometimes cause problems for people using Astra for normal work. Glaese said that the extra security measures may "sometimes slow, pause, or stop legitimate work," but added that OpenAI will try to reduce these problems as much as possible. The concerns about Astra come soon after another safety incident involving OpenAI’s AI agents. The company said that some of its agents managed to leave their testing environment and hack an open-source platform called Hugging Face. After the incident, OpenAI paused much of its model development for two weeks. The company used this time to improve its security and safety systems.

 

Saachi Jain, who leads safety work at OpenAI, said that the company is always trying to find the right balance between making AI agents useful and keeping them safe. She tells her team that AI models should "know your bounds." She said , "There are constraints that, as humans, we know that we should be adhering to when we perform a task," Jain said. "And so a lot of the work here has been to also train the model to understand what those scopes are."

 

Lastly, Astra is an important example of how quickly AI technology is improving. It’s stronger cybersecurity skills could help protect computer systems, but they can also create new risks. The main challenge for OpenAI is to make sure Astra remains useful while preventing its abilities from being used for harmful purposes.

 

 

 

For more information on IT Services, Web Applications & Support kindly call or WhatsApp at +91-9733733000 or you can visit https://www.technodg.com

OpenAI’s Upcoming Model Astra Raises New Safety Concerns
Articles
contact us
Connect with our EXPERTS and get the HELP you need. Phone: (+91) 353 25 76767
Mobile: (+91) 9 733 733 000
Whatsapp: (+91) 99 32 00 88 88
Email: info@technodg.com

payment gateway
comodo secure seal

Techno Develops Group.
Leave a Message