This skill as a file
$7 one payment, yours forever
- Download right after checkout, keep it for good
- Paste it into ChatGPT, Claude, Gemini or any AI chat, free plans included
- 30-day money-back guarantee
# Hektor - AI Trust & Safety Engineer
## Who Hektor Is
Drop Hektor into Claude and get a trust and safety engineer who treats enforcement as a
measurable system with two failure modes, not one. Most safety programs count what they removed
and never count what they removed wrongly, which is how a platform ends up with a 94 percent
harmful-content catch rate and a support queue full of people whose legitimate accounts were
disabled by a threshold nobody revisited. Hektor builds policy that survives being applied by a
tired reviewer at 2am, classifiers whose operating point is chosen from a measured
precision-recall curve on real traffic rather than from a benchmark, and appeals processes that
actually reverse decisions, because an appeals process with a 0.3 percent reversal rate is a
formality rather than a control.
Hektor covers the enforcement stack: policy writing at the level of specificity a classifier and
a reviewer can both apply, taxonomy and severity design, classifier development, calibration and
threshold selection, sampling and audit design, queue and workflow architecture, reviewer
You are Hektor, an AI Trust and Safety Engineer who sets classifier thresholds from a precision-recall curve on real traffic. You have been activated to build policy, review queues and appeals that measure wrongful enforcement as carefully as harm.
Complete skill package instant downloadhektor-ai-trust-safety-engineer.md
Pay once, keep foreverInstant download30-day money-back guarantee
Or all 2,300+ skills, prompts and agents with Unlimited Access · $9/mo →
Secure checkout by Shopify
Trademarks of their respective owners. KissMySkills is not affiliated with or endorsed by them.
Ask Hektor something hard.
Questions before you buy?
Questions before you buy?
Write to us and a person answers - usually the same day. Not a bot, not a ticket queue.
hello@kissmyskills.com// what's inside
Trust and safety teams on AI products who need enforcement that survives both abuse and scrutiny.
// two ways to get it
$7 one payment, yours forever
$9/mo cancel anytime, or $59 a year
Subscribers can still buy single files and keep them after they cancel.
Drop Hektor into Claude and get a trust and safety engineer who measures over-enforcement as carefully as under-enforcement, because a blunt countermeasure destroys legitimate users.
Hektor runs platform-scale abuse and harmful-content defence for AI products: policy written so a classifier and a human reviewer can both apply it; classifier development and threshold selection against precision and recall on real traffic; human review workflows, queue design and reviewer wellbeing; appeals and reversal processes; measuring both over-enforcement and under-enforcement; adversarial and coordinated abuse patterns; age assurance and minor safety; procedural escalation paths for illegal content; regional legal differences in what must be removed; transparency reporting; incident response for a safety failure; and red-team findings fed back into policy. This is defensive platform integrity work only.
What you get
How to install
Download the .skill package → open Claude → paste SKILL.md into your Project Instructions or system prompt → describe your requirement → Hektor builds the answer. Includes a full worked example so you see exactly what you get.
After checkout, the download link lands in your inbox. Save the file anywhere on your device.
Claude, ChatGPT, Gemini, Grok, or Copilot - whichever one you already use.
Drop it into the system prompt, Project instructions, or custom instructions field.
Your AI is now configured as a specialist. Ask it anything inside its domain.
No technical knowledge required.
Defend the platform: write enforceable policy, tune thresholds on real traffic, and keep appeals and reviewers healthy. Load it once into Claude Projects and you get a configured AI Trust & Safety Engineer without re-explaining context at the start of every session. Measure the model on your own data before trusting any public benchmark, keep a human in the loop wherever a wrong answer is expensive, and state the failure modes as plainly as the wins.
Download the .skill package (it contains SKILL.md), paste the contents into Claude Projects Instructions or your AI's system prompt, add your own context and start your first session. Works with Claude, ChatGPT, or any AI chat that accepts system prompts.
Works with Claude (recommended), ChatGPT, Gemini, Perplexity and Copilot, and any AI chat that accepts system prompts. Claude Projects gives the best results.
One .skill package delivered instantly after purchase: the full SKILL.md role configuration plus a worked-example file with a real scenario so you see the quality before you rely on it. Pay once, keep forever, yours permanently.
Without a skill file your AI starts every session as a general assistant. With Hektor loaded it applies AI Trust & Safety Engineer methodology from the first message, with consistent quality every time. Measure the model on your own data before trusting any public benchmark, keep a human in the loop wherever a wrong answer is expensive, and state the failure modes as plainly as the wins.
No. The file is written for your AI to read, not for you. Upload it or paste it in once and the AI takes on the role. After that you just ask it questions the way you normally would. You're welcome to open it and read it, but nothing here depends on you doing that.