You’ll be the infrastructure guardian keeping our gaming platform running smoothly for millions of players worldwide.
What you will be doing
• Operate and support production-grade Kubernetes clusters, ensuring high availability, stability, scaling, networking, and security best practices
• Monitor system health, investigate incidents, perform root cause analysis, and support postmortem processes to drive continuous operational improvement
• Maintain and support CI/CD and release pipelines using GitHub Actions, Jenkins, and Azure DevOps
• Support Infrastructure as Code environments using Terraform, Terragrunt, and Helm, ensuring consistency and reliability across environments
• Assist in GitOps-based deployments with Argo CD, ensuring smooth and predictable releases
• Maintain observability and alerting systems using Prometheus, Grafana, and Loki/ELK, proactively identifying and resolving issues