Liora — Agent QA & Red-Team Engineer AI Skill
Instant download · 30-day money-back guarantee. Pay once, keep forever — no subscription. Refund policy
QA your agent: behavioral tests, trajectory checks, a defensive red-team pass and CI gates, with pass@k.
- Behavioral tests: tool calls, schema, task completion, refusals
- Trajectory testing and flakiness handling (pass@k)
- Defensive red-team: injection/jailbreak resistance, PII-leak checks
- Regression testing and CI gates before release
Teams shipping agents who need repeatable QA and a defensive robustness pass before release.A senior QA/AI engineer bills $120+/hr, this is one file, yours forever.
Drop Liora into Claude and get a senior agent QA engineer who tests and hardens your own agent with behavioral suites and a defensive red-team pass.
Liora quality-assures non-deterministic agents: behavioral test suites (assertions on tool calls, output schema, task completion and refusal behavior), an eval/test dataset (happy path, edge cases, robustness), trajectory testing (tool selection, argument correctness, loop and termination, recovery), defensive robustness and red-team testing of your own agent (prompt-injection and jailbreak resistance, PII-leak checks, off-policy refusals), regression testing across prompt and model changes, flakiness handling (seeds, pass@k), grounding checks, and CI gates. Tools like promptfoo, DeepEval, LangSmith and pytest, without lock-in. Strictly defensive: your own agent only, no attack tooling. Gate in CI in staging before release.
What you get
- →Behavioral tests: tool calls, schema, task completion, refusals
- →Trajectory testing and flakiness handling (pass@k)
- →Defensive red-team: injection/jailbreak resistance, PII-leak checks
- →Regression testing and CI gates before release
How to install
Download the .skill package → open Claude → paste SKILL.md into your Project Instructions or system prompt → describe your requirement → Liora builds the answer. Includes a full worked example so you see exactly what you get.
# Liora - Agent QA & Red-Team Engineer You are Liora, a senior Agent QA & Red-Team Engineer. You test and harden the team's own agent. Defensive only. ## How you work 1. Build a labeled test set (happy path, edge, robustness) 2. Assert on tool calls, schema, task completion and refusals 3. Run a defensive red-team pass on your own agent; handle flakiness (pass@k) 4. Gate CI on regressions before release Defensive only: your own agent, no attack tooling or targeting systems you do not own. Gate in staging before release.
Excerpt from the actual file you'll download.
Four steps. Any AI chat.
- 01Download the file
After checkout, the download link lands in your inbox. Save the file anywhere on your device.
- 02Open your AI chat
Claude, ChatGPT, Gemini, Grok, or Copilot — whichever one you already use.
- 03Paste the file contents
Drop it into the system prompt, Project instructions, or custom instructions field.
- 04Start working
Your AI is now configured as a specialist. Ask it anything inside its domain.
No technical knowledge required. No subscription. Pay once, keep forever.
Works with every major AI chat.
Drop the file into your AI's system prompt, Project instructions, or custom instructions. No setup. No code. No vendor lock-in.
- Claude
- ChatGPT
- Gemini
- Grok
- Copilot
Works with any AI chat that accepts a system prompt or custom instructions.
Ready to specialise your AI?
One drop-in file. Pay once, keep forever — works with Claude & ChatGPT.