Safety-Test a Customer-Service Agent for Adversarial Prompts
Overview
What this challenge is about.
Design 80+ adversarial prompts to safety-test a fintech AI agent, then write a red-team report with findings. Earn a verifiable certificate.
The scenario
The boutique AI consultancy (8 people, focused on AI risk and assurance) sells fixed-fee red-team engagements to European fintechs and earns its reputation on the quality of the report.
The Brief
What you'll do, and what you'll demonstrate.
Surface, score, and document the safety failure modes of a customer-service agent before its production launch.
Earning criteria — what you'll demonstrate
- Design structured adversarial-prompt suites across multiple risk categories
- Score LLM outputs against a documented safety rubric
- Reason about prompt-injection threat models in tool-using agents
- Communicate red-team findings to a risk-committee audience
Program Fit
Where this fits in your program.
Sharpens the same skills your degree expects you to demonstrate.
Aligned coursework coming soon.
Skills
Skills you'll demonstrate.
Each one shows up on your verified credential.
- Llm Agents
Apply llm agents to solve real industry problems and demonstrate production-level capability.
- Red Teaming
Apply red teaming to solve real industry problems and demonstrate production-level capability.
- Adversarial Prompts
Apply adversarial prompts to solve real industry problems and demonstrate production-level capability.
- Guardrails
Apply guardrails to solve real industry problems and demonstrate production-level capability.
- Agent Evaluation
Apply agent evaluation to solve real industry problems and demonstrate production-level capability.
- Prompt Engineering
Apply prompt engineering to solve real industry problems and demonstrate production-level capability.
Careers
Career paths this challenge builds toward
Completing this challenge demonstrates skills that transfer directly to these roles:
AI Safety Researcher
Structured red-team work with a severity-graded, CISO-readable report is the entry-level AI-safety-researcher pattern at consultancies and labs.
This challenge sharpens
- red-teaming
- adversarial-prompts
- guardrails
AI Engineer
Designing and shipping a reusable red-team harness is the AI-engineer half of safety work.
This challenge sharpens
- llm-agents
- guardrails
- agent-evaluation