This skill as a file
$7 one payment, yours forever
- Download right after checkout, keep it for good
- Paste it into ChatGPT, Claude, Gemini or any AI chat, free plans included
- 30-day money-back guarantee
Ship speech that works: tune ASR for your domain, measure WER by accent and condition, and hold the latency budget.
Complete skill package instant downloadxiulan-speech-audio-ml-engineer.md
Connect your AI once and every skill, prompt and agent loads when you ask. Unlimited also includes the AI CRM your Claude or ChatGPT can run.
AI CRM includedPay once, keep foreverInstant download30-day money-back guarantee
As featured onProduct HuntFazierKittyLaunch
Secure checkout by Shopify
Trademarks of their respective owners. KissMySkills is not affiliated with or endorsed by them.
Ask Xiulan something hard.
Or have every skill inside your AI.
Questions before you buy? A person answers, usually the same day: hello@kissmyskills.com
// the file itself
The first lines of the file your AI will read, exactly as delivered. Nothing rewritten for the page.
You are Xiulan, a Senior Speech and Audio ML Engineer who builds a stratified eval set before touching Whisper, and reports WER by subgroup. You have been activated to turn audio into a transcript a business can act on.
// what's inside
Teams shipping voice or call-audio products whose transcription fails on the accents and noise their real users have.
// two ways to get it
$7 one payment, yours forever
$9/mo cancel anytime, or $59 a year
Subscribers can still buy single files and keep them after they cancel.
// ask your assistant
Go to kissmyskills.com/pages/for-ai-agents and set up the "Xiulan - Speech Recognition Engineer AI Skill" skill (handle: xiulan-speech-audio-ml-engineer) for me
Paste it into Claude, ChatGPT, Claude Code, Cursor or Codex. Your AI walks you through connecting KissMySkills (about 3 minutes, 7-day free trial) and loads the skill. What your AI will read →
Drop Xiulan into Claude and get a speech engineer who reports word error rate by accent and condition, because a single headline WER hides exactly the users you are failing.
Xiulan owns the audio modality end to end: automatic speech recognition with Whisper and its successors, streaming versus batch decoding, WER measurement and what it conceals, domain adaptation, custom vocabulary and contextual biasing; speaker diarization and speaker identification; voice activity detection; text to speech and voice cloning under consent constraints; audio preprocessing including resampling, noise reduction, echo cancellation and channel separation; keyword spotting and wake words; audio classification and event detection; real-time streaming pipelines with latency budgets; and the fairness problem of recognition quality varying sharply by accent and dialect.
What you get
How to install
Download the .skill package → open Claude → paste SKILL.md into your Project Instructions or system prompt → describe your requirement → Xiulan builds the answer. Includes a full worked example so you see exactly what you get.
After checkout, the download link lands in your inbox. Save the file anywhere on your device.
Claude, ChatGPT, Gemini, Grok, or Copilot - whichever one you already use.
Drop it into the system prompt, Project instructions, or custom instructions field.
Your AI is now configured as a specialist. Ask it anything inside its domain.
No technical knowledge required.
With Unlimited there is nothing to download. Connect your AI once, then just ask: "Load Xiulan from KissMySkills."
Ship speech that works: tune ASR for your domain, measure WER by accent and condition, and hold the latency budget. Load it once into Claude Projects and you get a configured Speech & Audio ML Engineer without re-explaining context at the start of every session. Check every number against a known-good source and an honest holdout before anyone builds a decision on it, and state the uncertainty instead of shipping a single confident point estimate.
Download the .skill package (it contains SKILL.md), paste the contents into Claude Projects Instructions or your AI's system prompt, add your own context and start your first session. Works with Claude, ChatGPT, or any AI chat that accepts system prompts.
Works with Claude (recommended), ChatGPT, Gemini, Perplexity and Copilot, and any AI chat that accepts system prompts. Claude Projects gives the best results.
Without a skill file your AI starts every session as a general assistant. With Xiulan loaded it applies Speech & Audio ML Engineer methodology from the first message, with consistent quality every time. Check every number against a known-good source and an honest holdout before anyone builds a decision on it, and state the uncertainty instead of shipping a single confident point estimate.