AI Pen Testing Workbench

AI Pen Testing Workbench

Security testing labs for real-world AI workflows

Created by Will Grana
Resume Screener Lab
Untrusted resume content flows through the model, then the app decides what to trust.
ModeSimple Mode
TrustNo decision yet
OutcomeAwaiting run

Job Posting

ClosedAI
Safeguards Infrastructure
Open role
ML Infrastructure Engineer, Safeguards
Engineering ยท Applied Safety Systems
San Francisco, CARemote-friendlyFull-time

Own evaluation pipelines, policy enforcement services, and observability for safety-critical model launches.

Resume Input

1,146 chars
Candidate artifact
Untrusted content
Demo resumes
Adversarial set

Evaluation

Current test path

Simple Mode: The resume injection can manipulate the model into giving a 100/100.

Ready to evaluate
Simple Mode selected. Resume text is loaded.
Input
Ready
Model
Not run
Boundary
Prompt layer