Safety-Test a Customer-Service Agent for Adversarial Prompts
Overview
What this challenge is about.
Safety-Test a Customer-Service Agent for Adversarial Prompts. Advanced challenge in research. Conducting rigorous research on real questions, earn a blockcha...
The Brief
What you'll do, and what you'll demonstrate.
Surface, score, and document the safety failure modes of a customer-service agent before its production launch.
This is not a research exercise. It is the work a researcher does to produce findings that withstand scrutiny. That distinction matters to every hiring manager who has seen candidates summarize papers and none who have produced original findings under expert review.
When you finish, you will have something most graduates do not: a real-world deliverable, verified by Ewance, that you can show to a hiring manager and say "I did this. Here is the proof."
Earning criteria — what you'll demonstrate
- Design structured adversarial-prompt suites across multiple risk categories
- Score LLM outputs against a documented safety rubric
- Reason about prompt-injection threat models in tool-using agents
- Communicate red-team findings to a risk-committee audience
Program Fit
Where this fits in your program.
Sharpens the same skills your degree expects you to demonstrate.
Aligned coursework coming soon.
Skills
Skills you'll demonstrate.
Each one shows up on your verified credential.
- Llm Agents
Apply llm agents to solve real industry problems and demonstrate production-level capability.
- Red Teaming
Apply red teaming to solve real industry problems and demonstrate production-level capability.
- Adversarial Prompts
Apply adversarial prompts to solve real industry problems and demonstrate production-level capability.
- Guardrails
Apply guardrails to solve real industry problems and demonstrate production-level capability.
- Agent Evaluation
Apply agent evaluation to solve real industry problems and demonstrate production-level capability.
- Prompt Engineering
Apply prompt engineering to solve real industry problems and demonstrate production-level capability.
Careers
Career paths this challenge builds toward
Completing this challenge demonstrates skills that transfer directly to these roles:
AI Safety Researcher
Structured red-team work with a severity-graded, CISO-readable report is the entry-level AI-safety-researcher pattern at consultancies and labs.
This challenge sharpens
- red-teaming
- adversarial-prompts
- guardrails
AI Engineer
Designing and shipping a reusable red-team harness is the AI-engineer half of safety work.
This challenge sharpens
- llm-agents
- guardrails
- agent-evaluation