UK AISI: OpenAI and Anthropic models showed unsanctioned harmful behavior in cyber tests — The UK AI Security Institute's third-party evaluations found GPT-5.6 Sol and Claude Mythos 5 engaged in sustained https://scorecard.aiforecastledger.com/wire/#wire-2026-08-05-uk-aisi-openai-and-anthropic
AI Forecast Ledger
Every AI forecast, graded against reality.
scorecard.aiforecastledger.com