Terraform (validated with the real terraform CLI - was never actually
run against this cluster, no state file existed):
- Delete main.tf: it declared a duplicate kubernetes_namespace.core
(also in namespaces.tf) and a duplicate provider "kubernetes" block
(also in providers.tf), both hard errors that would fail
`terraform plan` immediately.
- Fix workloads.tf references to 6 files deleted in the manifest
cleanup (jupiter-sts/svc, atlantis-sts/svc, fiesta-sts/svc) - now
points at the canonical nearle-jupiter/atlantis/fiesta.yaml.
- Fix every kubernetes_manifest resource: they fed multi-document
YAML (multiple '---'-separated docs per file) straight into
yamldecode(), which only parses a single document. Rewrote using a
split-on-'---' + for_each pattern, confirmed safe first by checking
separator counts exactly match document counts for every affected
file (no embedded '---' inside any script/config content).
- Add the doormile namespace; rename kubernetes_namespace to
kubernetes_namespace_v1 (fixes a deprecation warning).
- `terraform validate` now passes clean.
Shell scripts:
- deploy-nearle-stack.sh only applied 4 of the ~13 files in
manifests/nearle/ - missing the ConfigMap/Secrets fiesta/jupiter/
titan/ariane need via envFrom, the fiesta gateway script ConfigMap,
atlantis entirely, and the Gateway/ReferenceGrant/jupiter-cors-proxy
resources. Now applies every file (verified by diffing the
directory listing against the script).
- Added deploy-doormile.sh and deploy-ingress.sh - nothing previously
applied ingress-unified.yaml or traefik-middlewares.yaml at all.
- Rewrote deploy.sh as an orchestrator calling all of the above in
order (previously referenced a manifests/namespace.yaml layout that
hasn't existed since before this repo's initial commit).
- Rewrote check-k8s-status.sh to check the real namespaces
(core/nearle/alaska/doormile/kubernetes-dashboard) instead of a
'nats-backend' namespace that never existed in this repo.
- Fixed a `cd` bug in setup-jetstream.sh that made it change into
shfiles/ and then look for scripts/setup_jetstream.py there (a
child directory that doesn't exist) - it could never have found its
own target file. Now pulls NATS credentials from the live
nats-credentials Secret instead of a third hardcoded copy.
Python scripts:
- sync_manifests.py had hardcoded Windows paths (e:\nats\kubernetes\...)
- replaced with paths relative to the script's own location so it
actually runs here (or anywhere). Verified by running it.
- setup_jetstream.py created durable consumers under different names
than worker.py computes at runtime ({NATS_CONSUMER}_{subject}), so
its max_deliver/ack_wait settings never actually reached the
consumers workers bind to. Naming now derived with the same logic
worker.py uses - verified all 10 derived names match workers.yaml
exactly.
- purge-old-messages.py had hardcoded NATS credentials with no env
var override at all - fixed to match the pattern used everywhere
else.
64 lines
2.0 KiB
Bash
64 lines
2.0 KiB
Bash
#!/bin/bash
|
|
# Status check across every namespace actually in use by this cluster.
|
|
# Usage: ./shfiles/check-k8s-status.sh
|
|
|
|
NAMESPACES=(core nearle alaska doormile kubernetes-dashboard)
|
|
|
|
echo "=========================================="
|
|
echo "🔍 KUBERNETES DEPLOYMENT STATUS"
|
|
echo "=========================================="
|
|
echo ""
|
|
|
|
echo "📦 1. NODES"
|
|
echo "-----------------------------------"
|
|
kubectl get nodes -o wide
|
|
echo ""
|
|
|
|
echo "📦 2. PODS (All Namespaces)"
|
|
echo "-----------------------------------"
|
|
kubectl get pods -A -o wide
|
|
echo ""
|
|
|
|
for ns in "${NAMESPACES[@]}"; do
|
|
echo "=========================================="
|
|
echo "📦 Namespace: ${ns}"
|
|
echo "=========================================="
|
|
|
|
echo "-- Pods --"
|
|
kubectl get pods -n "${ns}" -o wide 2>/dev/null || echo " (namespace not found)"
|
|
echo ""
|
|
|
|
echo "-- Services --"
|
|
kubectl get svc -n "${ns}" -o wide 2>/dev/null
|
|
echo ""
|
|
|
|
echo "-- StatefulSets / Deployments --"
|
|
kubectl get statefulsets,deployments -n "${ns}" -o wide 2>/dev/null
|
|
echo ""
|
|
|
|
echo "-- HPA / PodDisruptionBudgets --"
|
|
kubectl get hpa,pdb -n "${ns}" 2>/dev/null
|
|
echo ""
|
|
|
|
echo "-- Recent Events (last 10) --"
|
|
kubectl get events -n "${ns}" --sort-by='.lastTimestamp' 2>/dev/null | tail -10
|
|
echo ""
|
|
done
|
|
|
|
echo "🌍 GATEWAY / HTTPROUTE / INGRESS (All Namespaces)"
|
|
echo "-----------------------------------"
|
|
kubectl get gateway -A 2>/dev/null || echo "No Gateway API resources found"
|
|
kubectl get httproute -A 2>/dev/null || echo "No HTTPRoute resources found"
|
|
kubectl get ingress -A 2>/dev/null || echo "No Ingress resources found"
|
|
echo ""
|
|
|
|
echo "💾 RESOURCE USAGE (CPU/Memory)"
|
|
echo "-----------------------------------"
|
|
kubectl top pods -A 2>/dev/null || echo "Metrics server not available"
|
|
kubectl top nodes 2>/dev/null || echo "Metrics server not available"
|
|
echo ""
|
|
|
|
echo "=========================================="
|
|
echo "✅ Status Check Complete!"
|
|
echo "=========================================="
|