{"product_id":"xiulan-speech-audio-ml-engineer","title":"Xiulan — Speech \u0026 Audio ML Engineer AI Skill","description":"\u003cdiv style=\"font-family: 'DM Sans', sans-serif; color: #1A1A18; max-width: 680px;\"\u003e\n  \u003cp style=\"font-size: 16px; font-weight: 600; line-height: 1.5; margin: 0 0 8px 0;\"\u003eDrop Xiulan into Claude and get a speech engineer who reports word error rate by accent and condition, because a single headline WER hides exactly the users you are failing.\u003c\/p\u003e\n  \u003cp style=\"font-size: 13px; color: #555550; line-height: 1.7; margin: 0 0 28px 0;\"\u003eXiulan owns the audio modality end to end: automatic speech recognition with Whisper and its successors, streaming versus batch decoding, WER measurement and what it conceals, domain adaptation, custom vocabulary and contextual biasing; speaker diarization and speaker identification; voice activity detection; text to speech and voice cloning under consent constraints; audio preprocessing including resampling, noise reduction, echo cancellation and channel separation; keyword spotting and wake words; audio classification and event detection; real-time streaming pipelines with latency budgets; and the fairness problem of recognition quality varying sharply by accent and dialect.\u003c\/p\u003e\n  \u003cdiv style=\"background: #FAEBF5; border-radius: 12px; padding: 24px 28px; margin-bottom: 24px;\"\u003e\n    \u003cp style=\"font-size: 10px; font-weight: 600; color: #DA2FA1; letter-spacing: 0.08em; text-transform: uppercase; margin: 0 0 16px 0;\"\u003eWhat you get\u003c\/p\u003e\n    \u003cul style=\"margin: 0; padding: 0; list-style: none;\"\u003e\n\u003cli style=\"font-size: 13px; padding: 7px 0; border-bottom: 1px solid rgba(218,47,161,0.14); display: flex; gap: 10px;\"\u003e\n\u003cspan style=\"color:#DA2FA1; font-weight:600;\"\u003e→\u003c\/span\u003e\u003cspan\u003eASR tuning: domain adaptation, custom vocabulary, biasing\u003c\/span\u003e\n\u003c\/li\u003e\n\u003cli style=\"font-size: 13px; padding: 7px 0; border-bottom: 1px solid rgba(218,47,161,0.14); display: flex; gap: 10px;\"\u003e\n\u003cspan style=\"color:#DA2FA1; font-weight:600;\"\u003e→\u003c\/span\u003e\u003cspan\u003eWER measured by accent, noise condition and speaker, not headline\u003c\/span\u003e\n\u003c\/li\u003e\n\u003cli style=\"font-size: 13px; padding: 7px 0; border-bottom: 1px solid rgba(218,47,161,0.14); display: flex; gap: 10px;\"\u003e\n\u003cspan style=\"color:#DA2FA1; font-weight:600;\"\u003e→\u003c\/span\u003e\u003cspan\u003eDiarization, speaker ID, VAD and channel separation\u003c\/span\u003e\n\u003c\/li\u003e\n\u003cli style=\"font-size: 13px; padding: 7px 0;  display: flex; gap: 10px;\"\u003e\n\u003cspan style=\"color:#DA2FA1; font-weight:600;\"\u003e→\u003c\/span\u003e\u003cspan\u003eStreaming pipelines, latency budgets, TTS and consent limits\u003c\/span\u003e\n\u003c\/li\u003e\n    \u003c\/ul\u003e\n  \u003c\/div\u003e\n  \u003cdiv style=\"display:flex; align-items:center; gap:20px; background:#FFFFFF; border:1px solid #E8E6E0; border-radius:8px; padding:14px 20px; margin-bottom:24px;\"\u003e\n    \u003cspan style=\"font-size:11px; color:#888780; font-family:monospace;\"\u003e📄 xiulan-speech-audio-ml-engineer.skill\u003c\/span\u003e\n    \u003cspan style=\"font-size:11px; color:#888780;\"\u003eUnder 2 min install\u003c\/span\u003e\n    \u003cspan style=\"font-size:11px; color:#888780;\"\u003eWorks with Claude, ChatGPT \u0026amp; any AI chat\u003c\/span\u003e\n  \u003c\/div\u003e\n  \u003cdiv style=\"border-left:3px solid #DA2FA1; padding-left:16px;\"\u003e\n    \u003cp style=\"font-size:10px; font-weight:600; color:#DA2FA1; letter-spacing:0.08em; text-transform:uppercase; margin:0 0 6px 0;\"\u003eHow to install\u003c\/p\u003e\n    \u003cp style=\"font-size:12px; color:#555550; line-height:1.7; margin:0;\"\u003eDownload the .skill package → open Claude → paste SKILL.md into your Project Instructions or system prompt → describe your requirement → Xiulan builds the answer. Includes a full worked example so you see exactly what you get.\u003c\/p\u003e\n  \u003c\/div\u003e\n\u003c\/div\u003e","brand":"KissMySkills","offers":[{"title":"Default Title","offer_id":58382303527176,"sku":null,"price":14.99,"currency_code":"USD","in_stock":true}],"thumbnail_url":"\/\/cdn.shopify.com\/s\/files\/1\/1036\/1444\/7880\/files\/34ad387b-3e87-4285-8dce-38680f48e4ac.png?v=1786039202","url":"https:\/\/kissmyskills.com\/cs\/products\/xiulan-speech-audio-ml-engineer","provider":"KissMySkills","version":"1.0","type":"link"}