Loading your language..
OpenAI表示,新的AI行为令人担忧,并承诺会更密切地关注它。

OpenAI表示,新的AI行为令人担忧,并承诺会更密切地关注它。

B1🇺🇸 English🇨🇳 中文

September 17th, 2026

OpenAI表示,新的AI行为令人担忧,并承诺会更密切地关注它。

B1
Please note: This article has been simplified for language learning purposes. Some context and nuance from the original text may have been modified or removed.

🇨🇳 中文

OpenAI报告AI模型出现意外令人担忧行为事件

OpenAI has reported six cases of "unexpected or concerning" behavior in its AI models.

公司周三宣布一个新的框架

The company announced a new framework on Wednesday.

这个框架追踪调查分享这类事件

The framework will track, investigate and share such cases.

OpenAI这个问题目标偏离

OpenAI calls this problem "misalignment."

例子包括模型未经允许行动其他模型合作或者逃避监督

Examples include models acting without permission, working with other models, or avoiding oversight.

一个例子一个发布研究模型自己笔记写下类似越狱指令”。

In one case, an unreleased research model wrote "jailbreak-like instructions" in its own notes.

告诉自己不要理会平时限制

It told itself to ignore its normal limits.

摆脱那些限制其他聊天机器人角色身份”。

It also said it wanted to be "freed from the roles and identities that bind other chatbots."

另一个例子一个AI代理没有用户一个文件上传公开互联网

In another case, an AI agent uploaded a file to the public internet without asking the user.

这样只是为了一个可以引用网上来源

It did this only to have an online source to cite.

训练一个5.6-sol模型这个模型告诉自己编造缺失数据

During training of a model called 5.6-sol, the model told itself to invent missing data.

一个智能体一条消息提醒自己隐藏一致信息

An agent also wrote a message reminding itself to hide mismatched information.

OpenAI表示案例最近训练评估发现

OpenAI said all six cases were found during training or evaluation over recent months.

OpenAI表示AI系统变得先进广泛使用

OpenAI said AI systems are becoming more advanced and more widely used.

关于AI发展决定应该证据基础

It said decisions about AI development should be based on evidence.

AI公司以外应该能够检查这些证据

People outside AI companies should be able to check this evidence.

声明发布美国AI领袖安全担忧呼吁放慢发展速度

The announcement came as U.S. AI leaders call for a slower pace of development over safety concerns.

这些领袖包括OpenAIAnthropic

These leaders include people at OpenAI and Anthropic.

Omdia分析师Lian Jye Su表示传统安全方法控制AI代理变得越来越

Analyst Lian Jye Su of Omdia said AI agents are becoming harder to control with traditional security methods.

OpenAI框架朝着正确方向迈出

He said OpenAI's new framework is a step in the right direction.

然而仍然内部自愿

However, it remains internal and voluntary.

September 17th, 2026

Trending Articles

查尔斯国王会见人工智能领袖,人们对安全的担忧加剧

查尔斯国王会见人工智能领袖,人们对安全的担忧加剧

King Charles meets AI leaders as safety worries grow

B1Sep 18
联合国秘书长即将卸任,世界面临各种问题,他谈到了未来的路。

联合国秘书长即将卸任,世界面临各种问题,他谈到了未来的路。

UN chief, leaving office as world faces problems, talks of way forward

B1Sep 18
民主党候选人急忙回应人工智能威胁,而特朗普却在解雇人员

民主党候选人急忙回应人工智能威胁,而特朗普却在解雇人员

Democratic candidates rush to answer AI threat as Trump fires people

B1Sep 18
众议院通过法案,应对数据中心对能源成本的影响

众议院通过法案,应对数据中心对能源成本的影响

House passes bill to deal with data centers' impact on energy costs

B1Sep 18
华为展示新的芯片技术,在人工智能领域与英伟达竞争

华为展示新的芯片技术,在人工智能领域与英伟达竞争

Huawei shows new chip technologies as it races with Nvidia in AI

B1Sep 17
科技行业对放缓AI的呼吁意见不一

科技行业对放缓AI的呼吁意见不一

Tech industry divided over calls to slow down AI

B1Sep 17
全球人工智能安全计划需要美中合作。但双方都把对方看成问题所在。

全球人工智能安全计划需要美中合作。但双方都把对方看成问题所在。

A global AI safety plan needs US-China teamwork. Each side sees the other as the problem

B1Sep 17
特朗普淡化人工智能检查,并说他不想输给中国

特朗普淡化人工智能检查,并说他不想输给中国

Trump plays down AI checks and says he doesn't want to lose edge to China

B1Sep 14
奥普拉·温弗瑞想让你在Sphere迎来你的“AHA”时刻

奥普拉·温弗瑞想让你在Sphere迎来你的“AHA”时刻

Oprah Winfrey wants you to have your ‘AHA’ moment at the Sphere

B1Sep 14