bbuf

bbuf/sglang-auto-driven-skills

13 resources in this repository

GitHub
🎯13

🎯Skills13

🎯sglang-prod-incident-triage🎯Skill

A skill for triaging SGLang production serving incidents using a replay-first approach, helping diagnose queue growth, timeouts, wrong outputs, crashes, and distributed stalls by preserving evidence and reproducing the request path before patching.

sglang-prod-incident-triage
🎯llm-serving-auto-benchmark🎯Skill

Part of Lifeskills, a curated collection of non-coding skills for AI agents focused on business-critical communication, strategy, negotiation, and influence with decision-ready outputs.

llm-serving-auto-benchmark
🎯llm-torch-profiler-analysis🎯Skill

An agent skill for AI infrastructure engineers that provides operational playbooks for torch profiler analysis, LLM serving benchmarks, and SGLang optimization. Includes skills for splitting prefill/decode profiler evidence and turning traces into kernel fusion opportunities.

llm-torch-profiler-analysis
🎯model-architecture-diagram🎯Skill

Agent-ready operational playbooks for AI infrastructure engineers, covering LLM serving benchmarks across SGLang, vLLM, and TensorRT-LLM, torch-profiler trace triage, kernel optimization opportunities, SGLang patch review, and production incident replay.

model-architecture-diagram
🎯model-pr-history-knowledge🎯Skill

Agent-ready playbooks for AI infrastructure engineers, covering LLM serving benchmarks, capacity planning, torch-profiler analysis, compute simulation, and SGLang/vLLM optimization. Includes 58 model PR histories and production incident triage skills.

model-pr-history-knowledge
🎯sglang-sota-humanize-loop🎯Skill

An agent-ready playbook for LLM serving benchmarks, capacity planning, torch-profiler triage, SGLang/vLLM optimization, and production incident analysis. Provides structured workflows for AI infrastructure performance tuning and human code review.

sglang-sota-humanize-loop
🎯sglang-humanize-review🎯Skill

Agent-ready playbooks for AI infrastructure engineers, covering LLM serving benchmarks, capacity planning, torch profiler analysis, pipeline inspection, compute simulation, and SGLang/vLLM optimization.

sglang-humanize-review
🎯llm-serving-capacity-planner🎯Skill

Agent-ready playbooks for AI infrastructure engineers covering LLM serving benchmarks, capacity planning, torch-profiler analysis, pipeline inspection, compute simulation, and SGLang/vLLM optimization with production incident triage.

llm-serving-capacity-planner
🎯vllm-sota-humanize-loop🎯Skill

An agent-ready playbook for LLM serving optimization, providing operational memory for benchmarking SGLang, vLLM, and TensorRT-LLM, analyzing serving capacity from logs, profiling at kernel level, and handling production incidents.

vllm-sota-humanize-loop
🎯model-compute-simulation🎯Skill

A collection of agent-ready playbooks for AI infrastructure engineers, covering LLM serving benchmarks, capacity planning, profiler triage, compute simulation, and SGLang/vLLM optimization with real maintainer discussion patterns.

model-compute-simulation
🎯llm-pipeline-analysis🎯Skill

An agent-ready playbook for AI infrastructure engineers that provides forward-pass, layer-level, and kernel-level timing analysis from torch profiler traces, part of a broader skill set covering LLM serving benchmarks, capacity planning, and SGLang/vLLM optimization.

llm-pipeline-analysis
🎯h100🎯Skill

Agent-ready playbooks for AI infrastructure engineers, providing operational skills for LLM serving benchmarks (SGLang, vLLM, TensorRT-LLM), torch-profiler triage, kernel optimization, SGLang code review, production incident replay, and model-family PR history tracking.

h100
🎯h100-sglang-diffusion🎯Skill

A collection of agent-ready playbooks for AI infrastructure engineers, covering LLM serving benchmarks, torch-profiler triage, SGLang optimization, code review, production incident handling, and model PR intelligence.

h100-sglang-diffusion