GPT-5.5 Solved 92.4% of Offensive Cyber Tasks and Broke the Benchmark: AI Offensive Capability Is Doubling Every Six Months and the Ruler Can’t Keep Up
GPT-5.5 solved 292 of 316 offensive cybersecurity tasks at 92.4% accuracy, saturating every benchmark Lyptus Research had built as “the hardest questions available globally.” The UK AI Security Institute independently confirmed rapid capability improvement. AI offensive cyber capability is doubling every 5–6 months. GPT-5.5-Cyber is now available to vetted defenders. GPT-5.6 Sol launched government-gated. A 50M token budget vs 2M token budget produces a 32-point accuracy gap on the hardest benchmarks — compute is now a direct multiplier on offensive AI capability. Five operational implications for enterprise security teams.
