- Use proper scaleTargetRef structure for http.keda.sh/v1alpha1
- Uncomment keda-vllm.yaml in kustomization
- This should allow Flux to reconcile the scaler
- Scale to 0 when idle (saves money)
- 15 min cooldown before scale-down
- Triggers on 3+ pending HTTP requests
- Targets openclaw-brain-vllm deployment