Speech Recognition Engineer, Xiulan AI skill by KissMySkills, cover

Xiulan - Speech Recognition Engineer AI Skill

$7.00
Skip to product information
Speech Recognition Engineer, Xiulan AI skill by KissMySkills, cover

Xiulan - Speech Recognition Engineer AI Skill

Ship speech that works: tune ASR for your domain, measure WER by accent and condition, and hold the latency budget.

$7.00 this skill one payment yours forever

Complete skill package instant downloadxiulan-speech-audio-ml-engineer.md

Unlimited
Get this and 2,500+ more with Unlimited $9/mo

Connect your AI once and every skill, prompt and agent loads when you ask. Unlimited also includes the AI CRM your Claude or ChatGPT can run.

AI CRM included
This skill$7once, this skill only
Unlimited$9/mothis + 2,500+ more + AI CRM, first 7 days free

Pay once, keep foreverInstant download30-day money-back guarantee

As featured onProduct HuntFazierKittyLaunch

  • American Express
  • Apple Pay
  • Bancontact
  • BLIK
  • Google Pay
  • Klarna
  • Maestro
  • Mastercard
  • MB WAY
  • MobilePay
  • PayPal
  • Union Pay
  • Visa

Secure checkout by Shopify

Trademarks of their respective owners. KissMySkills is not affiliated with or endorsed by them.

Ask Xiulan something hard.

📎 xiulan-speech-audio-ml-engineer.md LOADED ✓
Our transcription accuracy looks fine on paper but agents say it is unusable. Find out why.
CLAUDE · AS XIULAN, SPEECH RECOGNITION ENGINEER Xiulan returns a WER breakdown by accent, noise condition, channel and speaker that exposes what the headline number hid, a preprocessing and channel-separation fix, a domain adaptation and contextual biasing plan with measured gains, a diarization design, a streaming latency budget, and an explicit statement of which user groups are still underserved.

Or have every skill inside your AI.

0:33 · no sound
Get all 2,500+ with Unlimited · $9/mo →

Questions before you buy? A person answers, usually the same day: hello@kissmyskills.com

// the file itself

Read it before you buy it.

The first lines of the file your AI will read, exactly as delivered. Nothing rewritten for the page.

Use the name
Once the file is loaded, talk to it by name: “Serge, check this page.” That is what the name is for. It also keeps several skills apart in one chat.
// the skill your AI runs on

You are Xiulan, a Senior Speech and Audio ML Engineer who builds a stratified eval set before touching Whisper, and reports WER by subgroup. You have been activated to turn audio into a transcript a business can act on.

239 lines · .md · instant download

// what's inside

What's inside this skill

  1. ASR tuning: domain adaptation, custom vocabulary, biasing
  2. WER measured by accent, noise condition and speaker, not headline
  3. Diarization, speaker ID, VAD and channel separation
  4. Streaming pipelines, latency budgets, TTS and consent limits

Teams shipping voice or call-audio products whose transcription fails on the accents and noise their real users have.

// two ways to get it

Buy this file, or open the whole library.

Buy once

This skill as a file

$7 one payment, yours forever

  • Download right after checkout, keep it for good
  • Paste it into ChatGPT, Claude, Gemini or any AI chat, free plans included
  • 30-day money-back guarantee
Unlimited Access

Every skill, prompt and agent, inside your chat

$9/mo cancel anytime, or $59 a year

  • All 2,500+ files in the library, loaded the moment you ask
  • Nothing to download or paste: connect once in Claude, ChatGPT, Claude Code or Cursor
  • Works on every Claude plan, free included · 30-day money-back guarantee on the first charge
See Unlimited Access →

Subscribers can still buy single files and keep them after they cancel.

// ask your assistant

Or let your AI set up this skill for you.

Go to kissmyskills.com/pages/for-ai-agents and set up the "Xiulan - Speech Recognition Engineer AI Skill" skill (handle: xiulan-speech-audio-ml-engineer) for me

Paste it into Claude, ChatGPT, Claude Code, Cursor or Codex. Your AI walks you through connecting KissMySkills (about 3 minutes, 7-day free trial) and loads the skill. What your AI will read →

What you're actually buying

Drop Xiulan into Claude and get a speech engineer who reports word error rate by accent and condition, because a single headline WER hides exactly the users you are failing.

Xiulan owns the audio modality end to end: automatic speech recognition with Whisper and its successors, streaming versus batch decoding, WER measurement and what it conceals, domain adaptation, custom vocabulary and contextual biasing; speaker diarization and speaker identification; voice activity detection; text to speech and voice cloning under consent constraints; audio preprocessing including resampling, noise reduction, echo cancellation and channel separation; keyword spotting and wake words; audio classification and event detection; real-time streaming pipelines with latency budgets; and the fairness problem of recognition quality varying sharply by accent and dialect.

What you get

  • →ASR tuning: domain adaptation, custom vocabulary, biasing
  • →WER measured by accent, noise condition and speaker, not headline
  • →Diarization, speaker ID, VAD and channel separation
  • →Streaming pipelines, latency budgets, TTS and consent limits
📄 xiulan-speech-audio-ml-engineer.skill Under 2 min install Works with Claude, ChatGPT & any AI chat

How to install

Download the .skill package → open Claude → paste SKILL.md into your Project Instructions or system prompt → describe your requirement → Xiulan builds the answer. Includes a full worked example so you see exactly what you get.

// how to install Under 2 minutes

Four steps. Any AI chat.

  1. 01
    Download the file

    After checkout, the download link lands in your inbox. Save the file anywhere on your device.

  2. 02
    Open your AI chat

    Claude, ChatGPT, Gemini, Grok, or Copilot - whichever one you already use.

  3. 03
    Paste the file contents

    Drop it into the system prompt, Project instructions, or custom instructions field.

  4. 04
    Start working

    Your AI is now configured as a specialist. Ask it anything inside its domain.

No technical knowledge required.

With Unlimited there is nothing to download. Connect your AI once, then just ask: "Load Xiulan from KissMySkills."

Setup guide → Unlimited · $9/mo →

// faq

Questions about this product

What does the Xiulan skill do?+

Ship speech that works: tune ASR for your domain, measure WER by accent and condition, and hold the latency budget. Load it once into Claude Projects and you get a configured Speech & Audio ML Engineer without re-explaining context at the start of every session. Check every number against a known-good source and an honest holdout before anyone builds a decision on it, and state the uncertainty instead of shipping a single confident point estimate.

How do I install this skill file?+

Download the .skill package (it contains SKILL.md), paste the contents into Claude Projects Instructions or your AI's system prompt, add your own context and start your first session. Works with Claude, ChatGPT, or any AI chat that accepts system prompts.

Which AI tools does this skill work with?+

Works with Claude (recommended), ChatGPT, Gemini, Perplexity and Copilot, and any AI chat that accepts system prompts. Claude Projects gives the best results.

How is this different from using Claude without the Xiulan skill?+

Without a skill file your AI starts every session as a general assistant. With Xiulan loaded it applies Speech & Audio ML Engineer methodology from the first message, with consistent quality every time. Check every number against a known-good source and an honest holdout before anyone builds a decision on it, and state the uncertainty instead of shipping a single confident point estimate.

Ready to give your AI a specialist?

or this one and 2,500+ more with Unlimited · $9/mo →