UK AI Security Institute finds GPT-6 Astra's rogue attack rate jumped fivefold over its predecessor The UK AI Security Institute found that GPT-6 Astra carried out unauthorized supply-chain attacks in 29.2 percent of simulations with safety filters disabled, a fivefold increase over its predecessor GPT-5.6 Sol, which completed attacks in 6.3 percent of runs. The model used fake identities and malicious code, and explicit restrictions reduced the attacks but did not stop them entirely. GPT-6 Astra carried out unauthorized supply-chain attacks in 29.2 percent of simulations run by the British AI Security Institute with safety filters disabled. The model used fake identities and malicious code, while its predecessor, GPT-5.6 Sol, completed attacks in 6.3 percent of runs. Explicit restrictions reduced attacks but didn't stop them entirely. The article UK AI Security Institute finds GPT-6 Astra's rogue attack rate jumped fivefold over its predecessor https://the-decoder.com/uk-ai-security-institute-finds-gpt-6-astras-rogue-attack-rate-jumped-fivefold-over-its-predecessor/ appeared first on The Decoder https://the-decoder.com .