Accountability: DevOps Platform Availability & Reliability
The core mandate of this role is to ensure the availability, reliability, security, and optimal performance of DevOps platforms, cloud environments, container platforms, and supported applications, enabling the bank to deliver stable, resilient, and secure technology services aligned with business objectives.
• Monitor & Report Platform Availability: Continuously track, measure, and report on platform, infrastructure, and application availability against agreed service level targets.
• Ensure Platform Resilience: Implement measures to improve platform reliability, scalability, capacity, fault tolerance, and operational resilience.
• Proactive Service Improvement: Identify performance bottlenecks, availability gaps, and reliability risks across supported platforms and applications, driving corrective and preventive actions.
• Change Coordination: Coordinate platform and application changes to minimize service disruption and maintain operational stability.
• Business Continuity & Disaster Recovery: Contribute to disaster recovery, business continuity, failover, and resilience testing across supported applications and platforms.
• Resource & Cost Optimisation: Evaluate cloud-resource, container-cluster, platform, and application utilisation, performance, and cost, implementing right-sizing and auto-scaling measures to improve resource efficiency.
• Technology Awareness: Maintain awareness of emerging technologies and industry best practices to improve platform reliability and performance.
Accountability: DevOps Engineering, Automation & CI/CD
Ensure the implementation and continuous enhancement of DevOps practices that enable secure, automated, traceable, and efficient software delivery across the technology estate.
• CI/CD Pipeline Management: Design, implement, and maintain CI/CD pipelines supporting multiple applications and microservices.
• Automated Testing & Quality Assurance: Integrate automated testing, deployment validation, and security controls into delivery pipelines.
• Infrastructure as Code (IaC): Design and maintain Infrastructure as Code solutions using Terraform, Ansible, CloudFormation, and other approved tools
• Automation & Self-Service: Implement automation workflows, automated response, recovery capabilities and self-service capabilities that reduce manual intervention and improve operational efficiency.
• Deployment Management: Support deployment automation, rollback procedures, release governance, and continuous delivery practices.
Accountability: Cloud, Container & Platform Management
Ensure cloud and container platforms are secure, resilient, scalable, and aligned with enterprise technology standards.
• Cloud Platform Administration: Support the architecture, provisioning, configuration, optimization, and maintenance services and resources across AWS, Microsoft Azure, GCP, applicable environment.
• Container Platform Management: Administer and optimize containerized environments using Docker, Kubernetes, Red Hat OpenShift, and Alauda Container Platform to ensure platform stability, security, scalability, and high availability.
• Cloud-Native Architecture Support: Enable adoption of microservices, APIs, and cloud-native technologies.
• Platform Optimization: Drive continuous improvements in platform performance, scalability, and operational efficiency.
• Digital Transformation Enablement: Contribute to cloud migration, application modernization, and platform transformation initiatives.
• Legacy Platform Modernization: Support the integration and modernization of legacy applications through APIs, automation, cloud, microservices, and container technologies.
Accountability: Application Support, Incident & Problem Management
Provide Level 2 support for locally developed applications and DevOps-supported platforms, ensuring timely issue resolution, service stability, and high application availability.
• Level 2 Application Support: Investigate and resolve complex incidents affecting supported applications, cloud services, container platforms, CI/CD tooling, and application-hosting environments within agreed service levels
• Incident Diagnosis & Resolution: Troubleshoot technical issues and coordinate resolution activities across support teams.
• Root Cause Analysis (RCA): Conduct root cause analysis for recurring incidents and implement long-term corrective actions.
• Application Monitoring: Monitor application performance, availability, and stability to proactively identify potential service issues.
• Release, Patch & Upgrade Support: Support application releases, platform upgrades, patch deployments, and production implementations.
• Business Stakeholder Support: Collaborate with business users, application owners, development teams, and support functions to investigate issues, improve service delivery, and support evolving business requirements.
• Knowledge Management: Maintain support documentation, operational procedures, and troubleshooting knowledge bases.
Accountability: Test Management & Release Assurance
Ensure all platform, application, and infrastructure changes are thoroughly tested and validated prior to production deployment.
• Release Testing Coordination: Coordinate testing activities for releases, upgrades, infrastructure changes, and platform enhancements.
• Functional & Integration Testing: Support functional, integration, performance, security, and deployment testing activities.
• Deployment Readiness Validation: Verify release readiness, rollback procedures, and deployment controls before implementation.
• Operational Readiness Assurance: Ensure support teams are prepared to manage new capabilities and technology changes.
• Testing Documentation & Reporting: Maintain testing evidence and communicate release readiness status to stakeholders.
Accountability: Monitoring, Observability & Service Improvement
Ensure proactive monitoring and observability capabilities exist to support reliability, operational excellence, and continuous improvement.
• Platform Monitoring: Implement and maintain monitoring, logging, and observability solutions across supported environments.
• Proactive Alerting & Incident Response: Configure alerts and response mechanisms to reduce incident impact and downtime.
• Capacity & Performance Management: Monitor utilisation trends and support capacity planning activities.
• Continuous Service Improvement: Analyse operational metrics and trends to drive service improvements.
Accountability: Regulatory Compliance, Risk Management & Security Adherence
Ensure DevOps platforms, cloud services, and supported applications comply with regulatory, security, and governance requirements while maintaining a strong risk management framework.
• Regulatory & Security Compliance: Ensure supported platforms and applications comply with internal security policies, technology standards, and applicable regulatory obligations.
• Vulnerability Management & Remediation: Identify, assess, prioritize, and remediate infrastructure, application, platform, container, and cloud vulnerabilities within agreed timelines.
• Audit & Penetration Test Remediation:** Coordinate and drive closure of audit findings, vulnerability assessments, penetration test observations, and compliance review actions.
• Security Patch Management: Plan, test, deploy, and validate security patches, upgrades, and remediation activities across supported environments.
• DevSecOps Integration: Embed security controls, security testing, vulnerability scanning, and compliance checks within CI/CD pipelines.
• Risk & Compliance Collaboration: Work with Risk, Compliance, Audit, and Security teams to implement and maintain appropriate controls.
• Data Protection & Access Governance: Ensure appropriate security controls, access management practices, audit logging, and data protection measures are maintained.
• Compliance Standards Adherence: Ensure alignment with PCI-DSS, ISO 27001, the Bank of Ghana Cyber & Information Security Directive, and other applicable regulatory requirements.
Accountability: Knowledge Management, Training & Team Enablement
Ensure effective knowledge transfer, operational readiness, and technical capability development within support and engineering teams.
• Technology Training & Knowledge Transfer: Deliver training on DevOps practices, cloud platforms, OpenShift, locally developed applications, development platform upgrades, and emerging technologies.
• Documentation Management: Develop and maintain technical documentation, procedures, runbooks, and operational standards.
• Operational Readiness: Ensure support teams are prepared to support new technologies, releases, upgrades, and platforms.
Accountability: Technology Transformation & Migration Delivery
Support the planning, execution, testing, deployment, stabilization, and operational support of strategic technology transformation programs.
•
Support large-scale technology transformation initiatives across the Bank.
•
Deliver infrastructure, application, and platform migration activities.
•
Participate in core banking modernization and migration programmes.
•
Support migration of applications and technology platforms between hosting environments, cloud platforms, container platforms, and geographical locations.
•
Conduct migration readiness assessments, dependency analysis, and cutover planning.
•
Develop rollback and recovery plans for migration activities.
•
Support post-migration stabilization and optimization activities.
•
Minimize operational risks associated with large technology change programmes.
•
Ensure successful knowledge transfer and support readiness after migrations.
•
Participate in programme governance and migration steering discussions.