home/categories/containers/bagelhole-devops-security-agent-skills-infrastructure-local-ai-llm-inference-scaling-skill-md

containersdevops

llm-inference-scaling

Auto-scale LLM inference clusters on Kubernetes using KEDA, custom GPU metrics, and horizontal pod autoscaling. Handle traffic spikes, implement queue-based scaling, and optimize cost with spot instances for AI workloads.

स्रोत देखें containers

maintainer

BagelHole

अपडेट किया गया 3/2/2026

स्टार

फोर्क

quick start

Installation and usage

इंस्टॉलेशन

$ install --globalskills.sh

उपयोग

इंस्टॉल करने के बाद, आप टर्मिनल में यह कमांड चलाकर इस स्किल का उपयोग कर सकते हैं:

skills use llm-inference-scaling