ChatGPT-maker OpenAI disclosed a new round of “concerning” incidents involving its AI, the latest in a string of events in which the technology has cheated, hacked into other companies’ systems or tried to manipulate humans. www.washingtonpost.com/technology/2...
OpenAI reveals new cases of AI models cheating, going off script
Newly disclosed incidents show models manipulating tests and generating their own instructions, raising fresh questions about AI safety.
washingtonpost.com