home/categories/containers/project-hami-hami-skill-k8s-debug-gpu-pod-skill-md
containersdevops

k8s-gpu-pod-troubleshooter

A comprehensive diagnostic skill for troubleshooting GPU pod scheduling and allocation issues in Kubernetes clusters using HAMi (Heterogeneous AI Computing Virtualization Middleware). It identifies GPU resource constraints, webhook configuration problems, device plugin issues, and scheduler policy misconfigurations to provide actionable remediation guidance.

Project-HAMi
maintainer
Project-HAMi
Atualizado 3/2/2026
Estrelas
3256
Forks
506
quick start

Installation and usage

A comprehensive diagnostic skill for troubleshooting GPU pod scheduling and allocation issues in Kubernetes clusters using HAMi (Heterogeneous AI Computing Virtualization Middleware). It identifies GPU resource constraints, webhook configuration problems, device plugin issues, and scheduler policy misconfigurations to provide actionable remediation guidance.

Instalação
$ install --globalskills.sh
Uso

Depois de instalar, você pode usar esta skill executando o seguinte comando no terminal:

skills use k8s-gpu-pod-troubleshooter