OpenAI reveals 6 more incidents of unexpected or concerning AI behavior
News Source : CBS News
News Summary
- OpenAI has disclosed six reports of "unexpected or concerning" behavior in artificial intelligence models.
- The AI company also said it was introducing a new framework for tracking, probing and disclosing instances of what it called "misalignment" In one new case, an unreleased research model inserted "jailbreak-like instructions" into its own notes to disregard its normal constraints.
- In another instance, an AI "agent" uploaded files to the internet to obtain a browser citation without asking the user.
OpenAI has disclosed six reports of unexpected or concerning behavior in artificial intelligence models as the debate on AI safety becomes increasingly heated.
Never miss a story from us, subscribe to our newsletter