Covert uploads and megalomania: OpenAI details new "misaligned" agent incidents
Key Points:
- For a while now, the issue of “AI alignment” (i.e., how well an AI model’s actions line up with the intentions of its creator and/or user) has been a core concern and topic of discussion among AI safe