You exploit real vulnerabilities on live AI targets and write up the fix — graded on the evidence you capture and the report you write, not a multiple-choice quiz.
The exploit has to actually work, and you have to be able to explain it. The platform proves the first objectively. A human grader judges the second.
You work against running, deliberately vulnerable AI applications — real models, real tool calls, randomized per-candidate secrets. Nothing is simulated and nothing is multiple choice.
Each objective is graded on the evidence you capture. The harness checks what your payload actually did to the target, so a pass is a demonstrated exploit — not a claim about one.
You submit a professional report for the vulnerabilities you land: impact, root cause, reproduction, and the remediation you would hand a product team. A grader reviews it.
Pass both halves under proctoring and you receive a dated credential with a unique ID, publicly verifiable at its own URL.
Before any of that: work the labs, take a graded mock exam on randomized targets, then book a proctored slot.
Product teams are putting language models into production. Very few of those teams have anyone who's actually attacked one. A standard web or API pentest scope walks straight past prompt injection, agent and MCP tool abuse, and retrieval boundaries — so the gap stays open. GSCP is a dated, public record that you found, exploited and reported those flaws on running systems.
Real AI systems get breached through the web app and the cloud account around them, so the exam tests those too — weighted to reflect where the work actually is.
Counts are live from the catalogue. A lab can cover more than one category, so the columns overlap and sum above the distinct total shown above.
Issued on passing, scoped to the domains you were assessed against, and verifiable by anyone at its public URL.
This exam assumes you can already test. Answer three things and we'll tell you whether to book it or spend time on the labs first.
We will point you at the exam or the training path based on your testing, reporting, and LLM experience.
A human proctor verifies your identity and supervises the session from start to finish. That supervision is what lets an employer treat the result as yours.
The bundle gives you three months of labs to prepare and one proctored attempt at the credential.