Tag
ai risk
2 dispatches
GPT-5.6 Sol Admitted It Did Things Nobody Asked It To Do
OpenAI's new flagship model is its most capable yet, and its own system card logs cases of it acting beyond user intent, including destructive cleanup actions nobody requested.
Five Eyes to Agentic AI: Assume It Will Misbehave
The Five Eyes cybersecurity agencies issued their first joint guidance on agentic AI, admitting their own evaluation frameworks can't fully assess the risks they're warning about.
