kernel-microbenchmark

Build, debug, and interpret vLLM GPU kernel microbenchmarks for CUDA, Triton, and CuteDSL, including CUPTI timing, correctness checks, generated-code inspection, multi-GPU measurements, and SOL sanity checks.

Install
npx skills add 'https://github.com/vllm-project/vllm/tree/main/.agents/skills/kernel-microbenchmark'
Download bundle ↓
main · da3c07bScanned 2026-09-17

Contributors

GitHub-linked commit authors for this SKILL.md at the saved revision. Co-authors and history before file renames are not included.

File history ↗

agents/openai.yaml

agents/openai.yamlBrowse 4 files
View on GitHub
← Back to SKILL.md
interface:  display_name: "Kernel Microbenchmark"  short_description: "Build and debug reliable GPU kernel microbenchmarks."  default_prompt: "Use the kernel microbenchmark workflow to build or debug a vLLM GPU kernel benchmark."