Get the latest tech news

How we monitor internal coding agents for misalignment


How OpenAI uses chain-of-thought monitoring to study misalignment in internal coding agents—analyzing real-world deployments to detect risks and strengthen AI safety safeguards.

None

Get the Android app

Or read this on Hacker News

Read more on:

Photo of OpenAI

OpenAI

Photo of misalignment

misalignment

Related news:

News photo

Research acceleration: The view inside OpenAI

News photo

OpenAI, Microsoft face copyright lawsuit from local newspapers

News photo

OpenAI admits to 'wiki incident' after its agents were discovered using a programming hub to communicate — says more transparency is needed regarding misalignments