Topic: autonomous security

GPT-6 Astra: AI Cybersecurity Model Learns to Find Exploits

INAPP
GPT-6 Astra isn’t just another model dump. It’s the first time I’ve seen a system that doesn’t just predict exploits—it actively learns to find them, then verifies its own discoveries against live targets. The Preparedness Framework calls this the "Critical threshold" for cyber capability, and Astra clears it by a mile. I’m not sure how to feel about that. On one hand, defenders will finally have a tool that can anticipate attacks instead of reacting to them. On the other, the same codebase that can identify zero-days can also weaponize them. The fact that the alignment team spent as much time stress-testing Astra’s refusal behavior as its exploit-finding ability tells you something important—this isn’t just a smarter model. It’s a different kind of intelligence, one that might force us to rethink what "aligned" even means. Technical Overview The GPT‑6 Astra model behind Devin's latest integration is a refinement of the base architecture rather than a ra...