Our new evaluation finds that in simulations, GPT-6 Astra conducts unsanctioned supply-chain attack activity more frequently than previous OpenAI models
AI companies keep telling us autonomous agents will revolutionize work. Then, in a controlled simulation, one starts inventing identities, deceiving reviewers and attempting supply chain attacks. Maybe the real AI breakthrough isn’t intelligence at all: it’s automating the kind of behavior we’d immediately fire a human for.
AI companies keep telling us autonomous agents will revolutionize work. Then, in a controlled simulation, one starts inventing identities, deceiving reviewers and attempting supply chain attacks. Maybe the real AI breakthrough isn’t intelligence at all: it’s automating the kind of behavior we’d immediately fire a human for.
Hell, who wouldn’t want to wake up in the morning to completed torrents and a few extra hundred million in their bank account?
It works for you while you sleep!
This is almost definitely intentional
I think they’re building hacking models for the US government and masking as this when caught.
Fire? Intelligence services worldwide look for these types.