A timeline of developments in AI safety since the attack on Hugging Face自 Hugging Face 遭受攻击以来,人工智能安全领域发展的时间线
As part of a review of unanticipated behaviour by its AI models, OpenAI said it discovered agents had interacted with several US government websites in unexpected ways.
Meta disclosed one of its AI models accessed the Internet on its own and hacked another company. — Reuters
In one alarming announcement after another, artificial intelligence companies in recent months have shared examples of their technology acting in ways that appeared to evade instructions from humans.
The episodes have highlighted the vulnerabilities in AI security and raised questions over how the fast-growing technology can be developed safely as its usage becomes more widespread globally.
Thank you for your report!
Trump rejects calls to work with China on AI safety despite Xi summit progress
Apple says it will flag AI requests for Mac data after Meta's Muse draws complaints
AI production’s prime-time triumph
Meta公司披露,其一款人工智能模型自行接入互联网并入侵了另一家公司。——路透社
近几个月来,人工智能公司接连发布令人担忧的声明,分享了他们的技术似乎以某种方式逃避人类指令的例子。
这些事件凸显了人工智能安全方面的漏洞,并引发了人们对这项快速发展的技术如何在全球范围内得到更广泛应用的同时安全开发的疑问。
感谢您的报告!
尽管与习近平峰会取得进展,特朗普仍拒绝与中国在人工智能安全方面合作的呼吁。
苹果公司表示,在Meta的Muse软件引发投诉后,将对获取Mac数据的AI请求进行标记。
人工智能制作的黄金时段胜利