Send candidates timed, proctored assessments built from real labs. They exploit live targets; you get auto-graded scores and a per-candidate report — evidence for your panel, not a verdict. Coverage runs across the OWASP LLM Top 10, MCP (Model Context Protocol), and agentic security — with web, API, and cloud tasks in the same sitting. A modern AI system gets breached through the layers around it; the assessment should too.
$199 per candidate, à-la-carte — no subscription. Packs of 10 and 25 bring it to ≈$180 and ≈$148 each.
A candidate who can break a prompt but not an API is only half useful — so you can test both in one sitting.
Prompt injection, extraction, unsafe output handling — the OWASP LLM Top 10.
Tool abuse, over-broad scopes, agent trust boundaries and MCP servers.
Access control, injection, business-logic flaws and API authorization.
Misconfiguration, identity and privilege paths around the workload.
Candidates face purpose-built vulnerable applications running live for their session — a distraction-free runner, a countdown timer, and evidence submission. No multiple choice, no take-home guesswork.
The grader replays each candidate's exploit to confirm it actually worked. Everyone sits the same task mix under the same server-side grading, so you can read their scores, targets solved and OWASP coverage side by side.
Each candidate gets a shareable report: headline score, a task-by-task breakdown with the evidence they submitted, OWASP coverage, and an activity timeline with proctoring flags. It stops there — no shortlist, no recommendation. Your panel makes the call.
Choose labs by category and difficulty — or start from a role template — invite candidates by link, and share the reproduced results with your panel.
Every sitting is time-boxed against fresh per-session targets, with randomized scenario selection and server-side grading, so there is no flag to copy and no candidate works on another candidate's target. Proctoring reports basic signals — tab focus and paste anomalies — honestly, and is not invasive surveillance. Every LLM lab is tagged to an OWASP LLM Top 10 item.
Scores and evidence are one input to a human screening decision, never the decision itself: the platform issues no hire recommendation and no pass/fail on a candidate. All names and scores in the mockups above are sample data.
Create your first assessment, or see a sample candidate report on a demo.
$199 per candidate · packs of 10 and 25 · no subscription required.