Karpenter’s defaults aren’t production-ready. This guide covers 10 specific practices to prevent real cluster failures: SQS interruption handling, NodePool isolation, disruption budgets, AMI pinning, and Prometheus alerting. All examples use the v1 stable API.