
vLLM Agent Skills
An inference and serving engine for large language models.
Repository snapshot · main · da3c07b
Scanned 2026-09-17
1–4 of 4 skills
ci-fails-buildkite
Fetch and diagnose vLLM Buildkite CI failure logs. Use when investigating failing CI jobs on a PR or build, when the user pastes a buildkite.com URL, or asks to fetch/diagnose CI logs.
.agents/skills/ci-fails-buildkite/SKILL.md
debug-ima
Debug CUDA illegal memory access (IMA) errors in vLLM with CUDA core dumps and cuda-gdb.
.agents/skills/debug-ima/SKILL.md
kernel-microbenchmark
Build, debug, and interpret vLLM GPU kernel microbenchmarks for CUDA, Triton, and CuteDSL, including CUPTI timing, correctness checks, generated-code inspection, multi-GPU measurements, and SOL sanity checks.
.agents/skills/kernel-microbenchmark/SKILL.md
triton-kernel-writing
Write or review Triton kernels for vLLM, with practical guidance for generated-code inspection, launch grids, indexing, specialization, tuning, and representative performance validation.
.agents/skills/triton-kernel-writing/SKILL.md
Discovery details
Tracked SKILL.md files, excluding tests, fixtures, dependencies, and vendored directories. The Instructions analysis keeps its own pinned revision.
Give your agents the whole story.
Skills teach agents how your team works. Modem shows them what customers said, who is affected, and what changed.