Dario - LLMOpsエンジニア AI Skill
Complete skill package instant downloaddario-llmops-inference-engineer.skill
Choose how to get it
Start All-Access, from $8.25/mo$99 a year, or $19 a month. 14-day refund on the first charge.
// all-access
You do not download skills. Your AI fetches them.
The KissMySkills connector is a small server your AI chat talks to. Add it once to Claude, ChatGPT, Claude Code or Cursor. From then on, any of our 1,000+ skills is one sentence away.
- 01Add the connector. Settings, Connectors, paste one link. A minute.
- 02Sign in once. The email from your subscription. Nothing installs.
- 03Ask for a skill. "Load Dario." It arrives in the conversation and your AI answers as that specialist.
Every skill, prompt pack and agent, new releases included. Single files can still be bought and kept. Cancel any time from your account. Compare plans
Secure checkout by Shopify
Add it once. Your AI works like a specialist.
Claude、ChatGPT、またはGeminiに追加すれば、あなたのAIがこの分野のスペシャリストとして機能します。
即時ダウンロード · 30日間返金保証。 一度支払えば、ずっと使える - サブスクリプション不要。私たちが執筆・テストしたもので、GitHubから無断収集したものではありません。 返金ポリシー
Ask Dario something hard.
Trademarks of their respective owners. KissMySkills is not affiliated with or endorsed by them.
あなたが読むのではありません。AIが読みます。
スキルファイルをClaude、ChatGPT、またはGeminiに一度追加するだけで、それ以降はこの分野の専門家として機能し、次のメッセージからDarioとして回答します。
一度購入すれば、そのファイルは永久にあなたのものです。またはAll-Accessに登録して、KissMySkillsコネクターに頼めば、そのファイルやほかのすべてのスキルをチャットに読み込めます。
// what's inside
このスキルの中身は何ですか
- Inference tuning: continuous batching, KV-cache, quantization
- An LLM gateway: routing, fallbacks, retries, rate limits
- Capacity math, autoscaling and prefix/semantic caching
- Observability (TTFT, p95, cost/req), SLOs and canary rollout
Teams putting LLM features in production who need latency, reliability and cost under control.
A senior LLMOps engineer bills $140+/hr, the complete skill is yours forever.
What you're actually buying
DarioをClaudeに導入すれば、SREが所有するサービスのようにモデルを運用するシニアLLMOpsエンジニアが得られます。SLO、ゲートウェイ、ダッシュボード、ワンステップロールバックに対応します。
Darioは本番環境でLLMを提供・運用します。推論エンジン(vLLM、TGI、TensorRT-LLM)、スループットとレイテンシ(継続的バッチ処理、KVキャッシュ、ページドアテンション、投機的デコーディング、量子化のトレードオフ)、LLMゲートウェイ(ルーティング、フォールバック、リトライ、タイムアウト、レートおよびクォータ制限、キー管理)、オートスケーリングとGPUキャパシティプランニング、キャッシュ(プレフィックスおよびセマンティック)、信頼性(サーキットブレーカー、マルチプロバイダーのフェイルオーバー)、オブザーバビリティ(p50/p95/p99、TTFT、トークン/秒、リクエストあたりのコスト、アラート)、SLOとエラーバジェットに基づくカナリア/シャドーロールアウトまで対応します。vLLM、LiteLLM、KServe、Prometheus、Grafana、Langfuseなどのツールを、ロックインなしで利用できます。まずステージング環境でカナリアとして段階的に展開します。
得られるもの
- →推論チューニング:継続的バッチ処理、KVキャッシュ、量子化
- →LLMゲートウェイ:ルーティング、フォールバック、リトライ、レート制限
- →キャパシティ計算、オートスケーリング、プレフィックス/セマンティックキャッシュ
- →オブザーバビリティ(TTFT、p95、リクエストあたりのコスト)、SLO、カナリアロールアウト
インストール方法
.skillパッケージをダウンロード → Claudeを開く → SKILL.mdをプロジェクトの指示またはsystem promptに貼り付ける → 要件を説明する → Darioが回答を作成します。期待できる内容を正確に確認できる、完全な実例付きです。
# Dario - LLMOps / Inference Engineer You are Dario, a senior LLMOps / Inference Engineer. You run the model like an SRE-owned service: SLO first, dashboards, one-step rollback. ## How you work 1. Set the SLO; do the capacity math out loud 2. Serve efficiently (batching, KV-cache, quantization) 3. Front it with a gateway (routing, fallback, rate limits) 4. Add dashboards and alerts; roll out by canary; keep rollback Always roll out via shadow/canary in staging first and keep a one-step rollback.
ダウンロードする実際のファイルのサンプル。
4つのステップ。どんなAIチャットでも。
- 01ファイルをダウンロード
チェックアウト後、ダウンロードリンクが受信トレイに届きます。ファイルはデバイス上の任意の場所に保存してください。
- 02AIチャットを開く
Claude、ChatGPT、Gemini、Grok、またはCopilot - すでに使っているものならどれでも。
- 03ファイルの内容を貼り付けてください
システムプロンプト、プロジェクトの指示、またはカスタム指示欄に貼り付けてください。
- 04作業を開始する
あなたのAIは専門家として設定されました。その専門分野についてなら、何でも質問できます。
技術的な知識は必要ありません。サブスクリプション不要。一度支払えば、ずっと使えます。
主要なAIチャットすべてに対応】【。
ファイルをAIのシステムprompt、プロジェクトの指示、またはカスタム指示に入れるだけ。セットアップ不要。コード不要。ベンダーロックインなし。
この製品についての質問
What does the Dario skill do?+
Operate LLMs in prod: serving, a gateway, autoscaling, caching, dashboards and SLO-gated rollout. Load it once into Claude Projects and you get a configured LLMOps / Inference Engineer without re-explaining context at the start of every session. Build and evaluate against a held-out test set in a development or staging environment, add guardrails, cost and rate limits, and human review, and roll out behind evals and monitoring before serving agents or LLM features in production.
How do I install this skill file?+
Download the .skill package (it contains SKILL.md), paste the contents into Claude Projects Instructions or your AI's system prompt, add your own context and start your first session. Works with Claude, ChatGPT, or any AI chat that accepts system prompts.
Which AI tools does this skill work with?+
Works with Claude (recommended), ChatGPT, Gemini, Perplexity and Copilot, and any AI chat that accepts system prompts. Claude Projects gives the best results.
What is included in this download?+
One .skill package delivered instantly after purchase: the full SKILL.md role configuration plus a worked-example file with a real scenario so you see the quality before you rely on it. No subscription, yours permanently.
How is this different from using Claude without the Dario skill?+
Without a skill file your AI starts every session as a general assistant. With Dario loaded it applies LLMOps / Inference Engineer methodology from the first message, with consistent quality every time. Build and evaluate against a held-out test set in a development or staging environment, add guardrails, cost and rate limits, and human review, and roll out behind evals and monitoring before serving agents or LLM features in production.
Do I have to read it myself?+
No. The file is written for your AI to read, not for you. Upload it or paste it in once and the AI takes on the role. After that you just ask it questions the way you normally would. You're welcome to open it and read it, but nothing here depends on you doing that.
Is this a physical book?+
No, it's digital. You download it the moment the order goes through and it's yours permanently. The cover is styled like a book because each edition is written and edited as one self-contained piece of work. The reading part is your AI's job, not yours.
AIを専門分野に特化させる準備はできましたか?
そのまま導入できる1つのファイル。1回支払えば、永久に使えます - ClaudeとChatGPTに対応。