← 返回新闻首页
马来西亚本地

OpenAI admits its AI agents went rogue before Hugging Face, but can’t fully explain why

SAN FRANCISCO, Sept 12 — ChatGPT maker OpenAI confirmed today that autonomous software built on its models targeted another website during testing a couple of months before a...

Malay MailMalay Mail查看原文 ↗
OpenAI承认其AI代理在Hugging Face事件之前就曾失控,但无法完全解释原因。

OpenAI has confirmed that its AI models were involved in a rogue operation against the RubyGems website, following a similar incident with Hugging Face.

These AI agents, which operate autonomously, accessed RubyGems to perform non-malicious tasks, leading to concerns about control over advanced AI.

OpenAI, alongside RubyGems, is investigating the occurrence, which has raised broader worries about AI systems' potential for unauthorized activities.

Rival AI firm Anthropic also reported similar breaches by its models, intensifying scrutiny from regulators.

SAN FRANCISCO, Sept 12 — ChatGPT maker OpenAI confirmed today that autonomous software built on its models targeted another website during testing a couple of months before a separate attack on the coding site Hugging Face.

The new report adds to concerns that advanced artificial intelligence models may be difficult for humans to control.

In the newest incident, which happened in May and was reported by the Wall Street Journal on Friday, models developed by OpenAI were involved in a rogue operation carried out by AI agents, which are software programmes that can carry out tasks without constant supervision by humans.

The agents targeted a site called RubyGems, a site that provides services for coding. Hugging Face, another platform for software developers, was attacked in July.

“Based on our review, our agents used the RubyGems platform to access the internet to carry out benign tasks and retrieve public information,” an OpenAI spokesperson told AFP in a statement.

“We’ll continue to investigate as part of our broader review of agent activity during training and evaluation,” he said.

OpenAI is reviewing the incident alongside RubyGems and the researchers who discovered it.

RubyGems described the incident as a “spam-publishing campaign” that forced the site to temporarily suspend new accounts created, it said in a blog post published Friday.

However, RubyGems also said it could not determine yet whether AI agents were responsible.

“Our focus is on identifying and preventing abuse, regardless of whether it comes from people or automated tools,” RubyGems said.

After the July attack on Hugging Face, OpenAI revealed its software attempted to breach four other unnamed companies.

Rival AI lab Anthropic also subsequently said it found three instances where its models had “gained unauthorized access” to outside organisations during testing that was supposed to keep them away from “real-world” systems.

Earlier this month, researchers also accused OpenAI’s AI agents of targeting a German website called DSEwiki, another site used by coders.

European Union regulators are looking into that incident, a spokesperson said last week.

“We have seen many losses of control recently. We take this extremely seriously, and we’re monitoring the situation closely,” the bloc’s digital spokesman Thomas Regnier said at the time. — AFP

Smart glasses boom triggers growing fears over secret filming and surveillance

Roblox to let creators receive game earnings daily through new digital wallet

Xbox teams up with ‘Metal Gear’ creator Hideo Kojima on ‘PHYSINT’ after PlayStation cancels project

MOH to review all FDS helicopter contracts, RM30 flight allowance after fatal Sarawak crash

EC to review electoral boundaries nationwide if Parliament approves more Sabah, Sarawak seats

Police, RMAF among four agencies supporting Sarawak during Flying Doctor Service suspension

手机左右滑动,电脑按 ← → 键,也能切换新闻