•Pod delivery — Own delivery for the platform scale and reliability pod, including planning, execution, quality, and operational readiness across the NBA platform.
•Platform reliability — Own uptime, performance, resilience, and operational excellence across platform services; proactively identify bottlenecks, failure points, and scaling constraints before they become production incidents.
•Quality engineering — Build and maintain automated testing frameworks spanning unit, integration, end-to-end, regression, and performance testing; establish quality standards that all delivery pods must meet before release.
•Hands-on development — Contribute production code alongside the team, building reliability tooling, test automation, observability capabilities, and platform infrastructure improvements.
•Observability and monitoring — Own the platform’s logging, metrics, distributed tracing, alerting, and monitoring strategy; ensure engineering teams have deep visibility into system behavior across environments.
•Error and failure management — Define and enforce patterns for exception handling, fault tolerance, recovery, and operational diagnostics across distributed services.
•Scalability engineering — Evaluate platform behavior under load and lead initiatives that improve performance, throughput, reliability, and cost efficiency as membership volume and engagement activity grow.
•Production readiness — Define and enforce release readiness criteria, operational quality gates, and handoff standards that services must satisfy before entering production ownership.
•Team leadership — Manage and mentor engineers and contractors; conduct code reviews, uphold engineering standards, and foster a culture of quality and operational excellence.
•Cross-team coordination — Partner closely with decisioning, orchestration, activation, platform, and data teams to identify reliability risks early and ensure production concerns are addressed throughout delivery.