Requirements for the role:
8+ years in cloud infrastructure, with 3+ years hands-on AWS - owning complex systems end-to-end and driving technical decisions across engineering teams
Deep expertise in AWS services, such as: EC2, EKS, ECS, RDS, Elasticache, S3, IAM, VPC, Route 53, CloudFront.
Expert-level Terraform (Terragrunt a strong plus), able to define custom internal modules to bootstrap complete solutions.
Confident managing multi-account AWS environments.
Strong Linux systems knowledge — kernel, networking, performance tuning, package management.
Container expertise: Docker, Kubernetes, and Helm chart authoring.
Kubernetes Topology: leverage Taints and tolerations, Affinities, Topology Spreads to build reliable systems.
Strong scripting skills in Bash and at least one general-purpose language (Python preferred).
Experience with CI/CD platforms: Jenkins, GitHub Actions, or CircleCI.
Source control mastery: Git (GitHub), branching strategies, code review workflows, GitOps.
Experience authoring reliable automation pipelines: from regular scripting to orchestration workflows using tools such as GitHub Actions or Step Function.
Observability & Security:
Hands-on experience with configuring monitoring/logging stacks: Prometheus/Grafana, ELK, and CloudWatch.
Nginx configuration, SSL/TLS management, and certificate lifecycle automation.
VPN implementation: OpenVPN.
Zero Trust Implementation: Cloudflare Tunnels.
Familiarity with cloud security best practices: IAM least-privilege, secrets management (Vault / AWS Secrets Manager), network segmentation.
Experience with Containers and Kubernetes security: Network Policies, RBAC, Logs and events auditing, containers hardening, Seccomp.
Experience supporting relational databases (PostgreSQL), caches (Redis, Memcache), and message buses (Kafka/RedPanda, RabbitMQ), NoSQL Databases (MongoDB, DynamoDB).
Working knowledge of networking fundamentals: DNS, TCP/IP, load balancing, CDN.