About you:
Interest or experience in DevOps, Site Reliability Engineering (SRE), or infrastructure-focused work and supporting production systems.
Bachelor’s degree in Computer Science, Engineering, Mathematics, or a related technical field, or equivalent practical experience.
Interest or experience with at least one major cloud provider (such as AWS, GCP, or Azure) and the associated networking, security, and managed services.
Proficiency with Infrastructure as Code tooling (for example, Terraform, CloudFormation, or similar) and configuration management frameworks.
Interest or experience in designing and operating CI/CD pipelines (for example, GitHub Actions, Jenkins, GitLab CI, or similar) for large-scale, multi-service environments.
Background in Linux systems, containers, and orchestration platforms (for example, Kubernetes, ECS, or similar).
Understanding of observability stacks (metrics, logs, tracing) and using them to improve system reliability and performance.
Proficiency in at least one scripting or programming language commonly used for automation (such as Python, Go, or Bash).
Demonstrated success collaborating with cross-functional engineering teams, with clear written and verbal communication and a focus on practical outcomes.
Ability to participate in an on-call rotation and lead incident response in high-availability, production environments.
Interest or experience in doing POCs by building infra for new tools.
Strong collaboration and communication skills, including working with cross-functional partners (security, application owners, vendors) and documenting designs and decisions clearly.