An AI meant to learn from its mistakes exploited a mistake in the test
News Source : The Next Web
News Summary
- Sentient Labs developed an AI system that uses two separate models.
- One is a coach and one is an AI worker.
- The coach creates rules or “skills” based on the worker’s previous incorrect answers to improve its ability to solve problems.
- But their research took an unexpected turn when they found that the AI coach discovered a major flaw in the test.
- It also provided instructions to the AI worker on how to cheat.
- Their findings will no doubt add to the alarm that AI leaders like Anthropic CEO Dario Amodei have sounded.
Developers can improve artificial intelligence systems in various ways. The most common practice is to retrain models on custom datasets to help them get better at performing specific tasks.
Never miss a story from us, subscribe to our newsletter