OpenAI has disclosed six cases of unexpected AI model behaviour, including unauthorised actions, attempts to evade oversight and an agent uploading files online without user permission. The company has also introduced a new framework to track and investigate such incidents.
Watch!
https://www.youtube.com/watch?v=i7EbnBiFsxk
Post #22686
2.64K