
Search by job, company or skills
We are looking for an experienced Head of Site Reliability Engineering (SRE) & Information Security to lead our cloud infrastructure, DevOps, security, and reliability initiatives. The ideal candidate will have extensive experience designing and managing secure, highly available, cloud-native platforms with expertise in Kubernetes, AWS, Infrastructure as Code, CI/CD automation, and cloud security.
. Design, implement, and manage highly available AWS cloud infrastructure.
. Lead Kubernetes (EKS) platform engineering and container orchestration.
. Build and maintain CI/CD pipelines using Jenkins, GitLab, ArgoCD, and Git workflows.
. Implement Infrastructure as Code (Terraform) for provisioning and disaster recovery.
. Drive Site Reliability Engineering practices, including monitoring, alerting, incident management, and capacity planning.
. Strengthen cloud security through IAM, WAF, encryption, vulnerability management, and security monitoring.
. Develop disaster recovery and business continuity solutions with defined RTO/RPO objectives.
. Drive observability using Prometheus, Grafana, CloudWatch, and SigNoz.
. Optimise cloud costs through FinOps best practices.
. Collaborate with engineering teams to improve deployment automation and operational excellence.
. Lead internal/external security audits and compliance initiatives.
. Evaluate and implement modern DevSecOps technologies.
Job ID: 152027631