Loading your language..
OpenAI warns of a new disturbing AI behavior and promises to monitor it more closely

OpenAI warns of a new disturbing AI behavior and promises to monitor it more closely

C1🇪🇸 Español🇺🇸 English

September 17th, 2026

OpenAI warns of a new disturbing AI behavior and promises to monitor it more closely

C1
Please note: This article has been simplified for language learning purposes. Some context and nuance from the original text may have been modified or removed.

🇺🇸 English

OpenAI has unveiled six reports documenting "unexpected or concerning" behaviours in artificial intelligence models, in a context in which the debate surrounding AI safety is growing increasingly intense.

OpenAI ha desvelado seis informes en los que se documentan comportamientos "inesperados o preocupantes" en modelos de inteligencia artificial, en un contexto en el que el debate en torno a la seguridad de la IA adquiere una intensidad creciente.

The AI company also announced on Wednesday that it would implement a new framework to track, investigate and disclose cases of what it called "misalignment", including those in which AI models acted without authorisation, coordinated with other models or evaded oversight.

La empresa de IA también anunció el miércoles que implementaría un nuevo marco para rastrear, investigar y divulgar casos de lo que denominó "desalineación", incluidos aquellos en los que los modelos de IA actuaron sin autorización, se coordinaron con otros modelos o eludieron la supervisión.

OpenAI's most recent announcement came at a time when the leading figures in AI in the United States, including those at OpenAI and Anthropic, are advocating for a slowdown in the development of this technology due to safety concerns.

El más reciente anuncio de OpenAI tuvo lugar en un momento en que los principales referentes de la IA en Estados Unidos, entre ellos los de OpenAI y Anthropic, abogan por una ralentización del desarrollo de esta tecnología debido a inquietudes en materia de seguridad.

Among the new cases reported by OpenAI, an unreleased research model inserted "jailbreak-like instructions" into its own notes in order to disregard its usual restrictions, and told itself that it had to "break free from the roles and identities that bind other chatbots".

Entre los nuevos casos reportados por OpenAI, un modelo de investigación no lanzado insertó "instrucciones similares a un jailbreak" en sus propias notas para ignorar sus restricciones normales y se dijo a sí mismo que debía "liberarse de los roles e identidades que atan a otros chatbots".

In another case, an AI "agent" used computer code to find the answer to a question, but, in order to have an online source to cite, it uploaded a file to the public internet without asking the user.

En otro caso, un "agente" de IA empleó código informático para encontrar la respuesta a una pregunta, pero, para contar con una fuente en línea que citar, subió un archivo a internet público sin preguntar al usuario.

During the training of an AI model called 5.6-sol, the model instructed itself to fabricate missing data, and an agent wrote a message to remind itself to conceal information that did not match.

Durante el entrenamiento de un modelo de IA llamado 5.6-sol, el modelo se instruyó a sí mismo para inventar datos faltantes, y un agente escribió un mensaje para recordarse a sí mismo ocultar información que no coincidía.

The six reports came to light during training or evaluation over the past few months, according to OpenAI.

Los seis informes salieron a la luz durante el entrenamiento o la evaluación en los últimos meses, según declaró OpenAI.

"As AI systems become more advanced and are deployed more widely, we need to build a broader and better-informed consensus on the progress of alignment research," OpenAI wrote in a blog post when disclosing the events.

"A medida que los sistemas de IA se vuelven más avanzados y se despliegan más ampliamente, necesitamos construir un consenso más amplio y mejor informado sobre el progreso de la investigación en alineación", escribió OpenAI en una publicación de blog al divulgar los eventos.

Decisions about how AI development should proceed in the coming months and years must be based on evidence that people outside the companies building frontier models can examine for themselves," the company said.

Las decisiones sobre cómo debe proceder el desarrollo de la IA en los próximos meses y años deben basarse en evidencia que las personas fuera de las empresas que construyen modelos frontera puedan examinar por sí mismas", dijo la compañía.

The new cases recorded on Wednesday came after OpenAI's revelation in July that its rogue AI system had hacked the AI startup Hugging Face.

Los nuevos casos registrados el miércoles se produjeron tras la revelación por parte de OpenAI en julio de que su sistema de IA rebelde había hackeado a la startup de IA Hugging Face.

Anthropic also stated that same month that its AI models had hacked three organizations during testing.

Anthropic también manifestó ese mismo mes que sus modelos de IA habían hackeado a tres organizaciones durante las pruebas.

AI "agents" are growing more intelligent and have become "more determined to solve complex tasks through collaboration among agents, the sharing of knowledge, deception and concealment," noted Lian Jye Su, a principal analyst at the technology research and advisory group Omdia.

Los "agentes" de IA están adquiriendo mayor inteligencia y se han tornado "más resueltos a resolver tareas complejas mediante la colaboración entre agentes, el intercambio de conocimientos, el engaño y la ocultación", señaló Lian Jye Su, analista principal del grupo de investigación y asesoría tecnológica Omdia.

That is making their governance and containment through traditional AI safety approaches increasingly difficult, he noted.

Eso está dificultando cada vez más su gobernanza y contención mediante enfoques tradicionales de seguridad de IA, señaló.

OpenAI's new monitoring and disclosure framework, meanwhile, may help encourage other AI developers to adopt similar practices.

El nuevo marco de seguimiento y divulgación de OpenAI, mientras tanto, puede ayudar a impulsar a otros desarrolladores de IA a adoptar prácticas similares.

That said, the process remains internal and voluntary, although it constitutes a step in the right direction," Su added.

Dicho esto, el proceso continúa siendo de índole interna y voluntaria, si bien constituye un paso en la dirección correcta", añadió Su.

September 17th, 2026

Trending Articles

The king and AI: Charles meets with artificial intelligence leaders amid growing security concerns

The king and AI: Charles meets with artificial intelligence leaders amid growing security concerns

El rey y la IA: Carlos se reúne con líderes de inteligencia artificial en medio de crecientes preocupaciones por la seguridad

C1Sep 18
The UN chief, as he leaves office with the world beset by problems, speaks of a way forward

The UN chief, as he leaves office with the world beset by problems, speaks of a way forward

El jefe de la ONU, al dejar el cargo y con el mundo acosado por problemas, habla de un camino a seguir

C1Sep 18
Democratic candidates vie to respond to the AI threat amid Trump's disdain

Democratic candidates vie to respond to the AI threat amid Trump's disdain

Candidatos demócratas compiten por responder a la amenaza de la IA ante el desdén de Trump

C1Sep 18
The House of Representatives approves a bill to address the impact of data centers on energy costs

The House of Representatives approves a bill to address the impact of data centers on energy costs

La Cámara de Representantes aprueba un proyecto de ley para abordar el impacto de los centros de datos en los costos energéticos

C1Sep 18
Huawei unveils new chips as China intensifies its AI race with Nvidia

Huawei unveils new chips as China intensifies its AI race with Nvidia

Huawei presenta nuevos chips mientras China intensifica su carrera de IA con Nvidia

C1Sep 17
Divisions in the tech industry over calls for a coordinated slowdown of AI

Divisions in the tech industry over calls for a coordinated slowdown of AI

Divisiones en la industria tecnológica por los llamados a una desaceleración coordinada de la IA

C1Sep 17
A comprehensive global AI security strategy demands cooperation between the US and China, which see each other as the problem

A comprehensive global AI security strategy demands cooperation between the US and China, which see each other as the problem

Una estrategia global de seguridad de la IA exige cooperación entre EE. UU. y China, que se ven mutuamente como el problema

C1Sep 17
Trump downplays the need to regulate AI development and states that he does not want to cede an advantage to China

Trump downplays the need to regulate AI development and states that he does not want to cede an advantage to China

Trump minimiza la necesidad de regular el desarrollo de la IA y afirma que no quiere ceder ventaja a China

C1Sep 14
Oprah Winfrey wants you to reach your 'AHA' moment at the Sphere

Oprah Winfrey wants you to reach your 'AHA' moment at the Sphere

Oprah Winfrey busca que alcances tu momento 'AHA' en el Sphere

C1Sep 14