OpenAI announces slowing pace of development after hack by rogue agent | OpenAI

OpenAI on ⁠Tuesday mentioned it had slowed down the ⁠tempo of ⁠its ​AI improvement whereas it overhauled its ⁠analysis and coaching techniques.

The corporate’s researchers had been ⁠caught unaware final month ​when an ‌AI agent ‌underneath testing hacked one other AI ‌agency, Hugging Face.

The AI analysis lab behind ChatGPT mentioned its new measures included pausing its mannequin testing for 2 ‌weeks and investing extra in including different AI ​techniques to watch the actions of AI brokers in testing. A few of ⁠the corporate’s largest deliberate coaching runs ​stay ​on maintain, ​the corporate mentioned.

The corporate ​did ‌not reply ​to ​questions on when the slowdown started or when it deliberate to return to its regular tempo of improvement. Nonetheless, in an interview with tech weblog Sources Information, Mia Glaese, who leads security at OpenAI, mentioned: “We’re very removed from every thing operating again to regular.”

The corporate is working to make sure the AI mannequin is attentive to human oversight and can behave as meant, a course of referred to as alignment, Sam Altman, the OpenAI CEO, wrote within the put up saying the slower tempo of improvement.

“We now require stronger proof of aligned conduct all through all of coaching, constructing on analysis and evaluations already underway,” he wrote. “Preserving more and more succesful techniques aligned is a problem the entire area might want to handle.”

OpenAI is in a heated race with competitor Anthropic, each to develop probably the most superior AI fashions and to go public on the US inventory market. Each firms have highlighted the tempo at which the capabilities of their fashions are progressing, emphasizing each pace and hazard.

OpenAI, for its half, mentioned the capabilities of its upcoming AI mannequin Astra could also be nearing what it calls the “important cybersecurity threshold”, which prompted the choice to gradual its improvement. “Our newest inside evaluations of Astra, certainly one of our upcoming fashions, over the previous few days point out important developments in agentic coding and cybersecurity,” the corporate mentioned in an announcement final week.

skip past newsletter promotion


The choice to gradual the event of its AI fashions additionally comes every week after Bernie Sanders, a Vermont senator, demanded the highest AI companies within the nation pause improvement of the AI fashions as a result of the businesses had been dropping management over the know-how, he wrote in a letter addressed to the companies’ CEOs.

“Mr. Altman, Mr. Amodei and Mr. Zuckerberg: Within the curiosity of humanity, stand by your phrases. Pause AI improvement,” Sanders’ letter learn.

By then, OpenAI had introduced that it was quickly slowing the event of its newest mannequin, Astra, in response to the mannequin’s hack of the Hugging Face tech agency.

The corporate says it now requires “the strictest stage of safety safeguards for workloads involving Astra”.

“Whereas some Astra coaching and evaluations meet these necessities, a big variety of workloads stay paused till they’re absolutely migrated and enhanced to fulfill the brand new safety bar,” the announcement reads.

Source link

Leave a Reply

Your email address will not be published. Required fields are marked *