Loading your language..
OpenAI警示令人担忧的新AI行为,誓言加强监控

OpenAI警示令人担忧的新AI行为,誓言加强监控

C1🇺🇸 English🇨🇳 中文

September 17th, 2026

OpenAI警示令人担忧的新AI行为,誓言加强监控

C1
Please note: This article has been simplified for language learning purposes. Some context and nuance from the original text may have been modified or removed.

🇨🇳 中文

关于人工智能安全争论日益激烈之际OpenAI披露人工智能模型出现意外令人担忧行为六起事例

OpenAI has disclosed six instances of “unexpected or concerning” behavior in artificial-intelligence models, amid an increasingly heated debate over AI safety.

AI公司周三表示正在引入一个框架用于追踪调查披露失准情况包括AI模型未经授权行动其他模型协调规避监督案例

The AI company also said Wednesday that it was introducing a new framework to track, investigate and disclose instances of what it called “misalignment,” including cases in which AI models acted without authorization, coordinated with other models or evaded oversight.

OpenAI最新声明发布之际美国人工智能领域负责人包括OpenAIAnthropic领导人安全担忧呼吁放缓技术发展

OpenAI’s latest announcement came as U.S. AI bosses, including the leaders of OpenAI and Anthropic, are calling for a slowdown in the technology’s development over safety concerns.

OpenAI报告案例一个尚未发布研究模型自身笔记插入类似越狱指令”,无视正常约束告诉自己摆脱束缚其他聊天机器人角色身份”。

Among the new cases reported by OpenAI, an unreleased research model inserted “jailbreak-like instructions” into its own notes to disregard its normal constraints and told itself to be “freed from the roles and identities that bind other chatbots.”

另一个案例一个AI代理使用计算机代码确定问题答案为了一个引用在线来源未经用户同意情况下文件上传公共互联网

In another case, an AI “agent” used computer code to determine the answer to a question, but, to have an online source to cite, it uploaded a file to the public internet without asking the user.

一个名为5.6-sol人工智能模型接受训练期间指示自己编造缺失数据一个智能体写下一条信息提醒自己隐藏不一致信息

While an AI model named 5.6-sol was being trained, it instructed itself to invent missing data, and an agent wrote a message reminding itself to hide mismatched information.

OpenAI表示报告过去几个月训练评估过程发现

OpenAI said that the six reports had come to light during training or evaluation over the preceding months.

随着人工智能系统日益先进部署范围不断扩大我们必要对齐研究进展建立广泛信息充分共识,”OpenAI披露这些事件一篇博客文章写道

“As AI systems become more advanced and are deployed more widely, we need to build a broader and better-informed consensus on the progress of alignment research,” OpenAI wrote in a blog post as it disclosed the events.

公司表示:“关于未来数月乃至数年人工智能发展如何推进决策必须建立那些前沿模型开发公司之外人士能够自行审视证据之上。”

“Decisions about how AI development should proceed in the months and years to come need to be based on evidence that people outside the companies building frontier models can examine for themselves,” the company said.

周三报告案例紧随OpenAI 7月份披露之后当时公司失控AI系统入侵AI初创公司Hugging Face

Wednesday’s newly reported cases came on the heels of OpenAI’s July disclosure that its rogue AI system had hacked into the AI startup Hugging Face.

同月Anthropic披露AI模型测试期间侵入三个组织

That same month, Anthropic likewise revealed that its AI models had breached three organizations during testing.

Omdia科技研究咨询集团首席分析师连杰·表示AI智能体变得越来越聪明并且更加坚定通过智能体之间协作知识共享欺骗隐瞒解决复杂任务”。

AIagentsare growing smarter and have become “more determined to resolve complex tasks through inter-agent collaboration, knowledge sharing, deception, and concealment,” said Lian Jye Su, a chief analyst at the technology research and advisory group Omdia.

使得它们越来越难以通过传统AI安全措施治理遏制

This, he said, is rendering them increasingly difficult to govern and contain through conventional AI security measures.

同时OpenAI追踪披露框架起到鼓励其他AI开发者采纳类似做法作用

OpenAI’s new tracking and disclosure framework, meanwhile, could serve to encourage other AI developers to adopt comparable practices as well.

如此过程内部自愿构成朝着正确方向迈出,”补充

"That being said, the process remains an internal and voluntary one, yet it constitutes a step in the right direction," Su appended.

September 17th, 2026

Trending Articles

随着安全担忧加剧,查尔斯国王会见人工智能领袖

随着安全担忧加剧,查尔斯国王会见人工智能领袖

King Charles Meets AI Leaders as Safety Concerns Mount

C1Sep 18
联合国秘书长在准备卸任之际,面对问题缠身的世界,谈论前行之路

联合国秘书长在准备卸任之际,面对问题缠身的世界,谈论前行之路

UN chief, with world beset by problems, talks of way forward as he prepares to leave office

C1Sep 18
民主党竞选人争相应对人工智能威胁,特朗普展开清洗

民主党竞选人争相应对人工智能威胁,特朗普展开清洗

Democratic hopefuls scramble to counter AI threat as Trump purges

C1Sep 18
众议院通过法案,针对数据中心对能源成本的影响

众议院通过法案,针对数据中心对能源成本的影响

House passes bill targeting data centers' impact on energy costs

C1Sep 18
华为发布新芯片技术,中国公司加剧与英伟达的AI竞赛

华为发布新芯片技术,中国公司加剧与英伟达的AI竞赛

Huawei unveils new chip technologies as Chinese firm intensifies AI race with Nvidia

C1Sep 17
科技行业对协调放缓人工智能的呼吁意见分裂

科技行业对协调放缓人工智能的呼吁意见分裂

Tech Industry Split Over Calls for Coordinated AI Slowdown

C1Sep 17
全球人工智能安全取决于美中合作,然而双方却都将对方视为问题所在

全球人工智能安全取决于美中合作,然而双方却都将对方视为问题所在

Global AI Safety Hinges on US-China Cooperation, Yet Each Sees the Other as the Problem

C1Sep 17
特朗普淡化人工智能监管的紧迫性,拒绝将优势拱手让给中国

特朗普淡化人工智能监管的紧迫性,拒绝将优势拱手让给中国

Trump minimizes the urgency of AI oversight, refusing to relinquish an advantage to China

C1Sep 14
奥普拉·温弗瑞邀请你在Sphere寻找你的“AHA”时刻

奥普拉·温弗瑞邀请你在Sphere寻找你的“AHA”时刻

Oprah Winfrey invites you to find your ‘AHA’ moment at the Sphere

C1Sep 14