I can help troubleshoot this Kubernetes issue.
Service overview
I have hands-on experience with EKS/GKE, Helm-based add-ons, ingress controllers, monitoring, autoscaling, and production readiness checks.
My approach would be:
1. Review the current cluster/workload state
2. Identify failing resources, events, logs, and configuration issues
3. Validate networking, ingress, DNS, RBAC, storage, and autoscaling if relevant
4. Apply the fix or provide exact remediation steps
5. Share a short report so the issue does not repeat
I can start by reviewing the symptoms, current manifests/Helm values, and recent changes.
