· Operate and harden SUSE Harvester environments: lifecycle management, upgrades, node/cluster health, HA, capacity planning, and incident response.
· Manage Kubernetes clusters (multi‑cluster, multi‑tenant) including cluster creation, upgrades, admission control, API server health, and etc.
· Own CNI operations with Antrea: policy design, network performance, and east‑west traffic observability.
· Run KubeVirt for VM workloads on Kubernetes: plan migrations, right‑size resources, and build reliable pipelines for VM lifecycle.
· Use Rancher to standardize cluster fleet management: provisioning (CAPI), templates, RBAC, and centralized policy/upgrade orchestration.
· Implement GitOps with FluxCD: define release pipelines, drift detection, progressive delivery, and automated rollbacks.
· Provision cloud/on‑prem resources with Crossplane: compose abstractions, manage providers, and enforce guardrails for day‑2 operations.
· Build and maintain SLOs/SLIs: availability, latency, error budgets; automate alerts and runbooks tied to service health.
· Reduce toil through automation: scripting, operators, controllers, and self‑service tooling for developers.
· Participate in on‑call rotations, post‑incident reviews, and reliability roadmaps; drive corrective actions and platform improvements.