DEV Community

Achin Bansal
Achin Bansal

Posted on Originally published at gridthegrey.com

GPT-6 Astra Tops ExploitBench With Perfect Security Score

Forensic Summary

OpenAI's GPT-6 Astra achieves 100% on ExploitBench and 99.2% on binary reverse engineering benchmarks, significantly outperforming its predecessor GPT-5.6 Sol on security-relevant tasks. The model's exceptional capability at offensive security benchmarks raises dual-use concerns, as frontier models with near-perfect exploit generation ability represent a meaningful capability uplift for threat actors. The article also notes the model's strong long-context performance, which has implications for processing large codebases or security artifacts.


Read the full technical deep-dive on Grid the Grey: https://gridthegrey.com/posts/gpt-6-astra-tops-exploitbench-with-perfect-security-score/

Top comments (0)