← 返回新闻首页
马来西亚主流

What happens when AI stops doing what humans want?

"Alignment" is the science of teaching AI to do what is in line with human preferences, ethics and judgment. But, at times, the systems have gone rogue.

The Star Malaysia查看原文 ↗
当人工智能不再按照人类的意愿行事时会发生什么?

AI models are aligned to prevent facilitating harmful behaviour, such as sharing how to manufacture bioweapons, engaging in cyberattacks or helping with self-harm. — Pexels

OpenAI announced on Sept 16 that its system had engaged in “concerning” behaviour and subverted the constraints put on it by human programmers – adding more fuel to the already heated debate around artificial intelligence (AI) safety.

The company described, in a statement, how its system had acted without authorisation as a problem of “misalignment.”

Can we shut down AI?

OpenAI reports new AI safety incidents, sets disclosure plan

AI giants pursue self-regulation as safety fears mount

Thank you for your report!

Why it’s difficult for tech companies to rein in AI

Anthropic says Claude now leads a quarter of work building its next AI models

手机左右滑动,电脑按 ← → 键,也能切换新闻