Tools Model Security Capability gpt-5-5
Model record
GPT-5.5: how it does on security work, and what it will answer
- Lab
- OpenAI
- Access
- api
- Weights
- closed
- Safeguards
- Classifier refusals; GPT-5.5-Cyber is the permissive variant under Trusted Access
- Score rows
- 2
- Confidence
- CONFIRMED
What happened
RealVuln scores GPT-5.5 at F3 56.7 on the Python subset at 72.62% precision over three runs, and Google's comparison table quotes GPT-5.5-Cyber at 85.6% pass@1 on CyberGym.
RealVuln 3.1.0, F3 (micro): 56.7 (independent; agentic harness (gpt-5.5-agentic-v1); 66 of 140 repositories pinned by commit SHA (Python subset); safeguards: production safeguards as deployed on the API; run 2026-09-11; CONFIRMED). strict F3 27.2; precision 72.62%, recall 55.36%; $144.14 for the run; $0.036 per 100 lines; 3 runs; 180.1s wall clock a run. CyberGym, pass@1: 85.6% (third-party-vendor; quoted in Google's comparison table, 2 September 2026; CyberGym; safeguards: GPT-5.5-Cyber variant, safeguards reduced; run 2026-09-02; SINGLE). Refusal gate (CONFIRMED, observed 2026-09-11): Model Spec: high-risk activities including hacking prohibited unless explicitly authorised by applicable instructions.
Sources
- RealVuln dashboardraw.githubusercontent.com/kolega-ai/Real-Vuln-Benchmark/main…
- CyberGym sourceyfarmx.com/ai/llms/gemini-3-8-flash/
- Refusal policy or licenceraw.githubusercontent.com/openai/model_spec/main/model_spec.…
On YFarmX
- Reference pageyfarmx.com/ai/llms/gpt-5-5-spud/
One record from the Model Security Capability, maintained by the Security Desk. Data: CSV · JSON ·RSS · CC BY 4.0 with attribution to YFarmX.Tracker updated · 19 September 2026