
Search by job, company or skills
Key Responsibilities
1. Work with L1 vendor engineers and the internal team to oversee day-to-day operational support and maintenance of critical infrastructure platforms across on-premises and hybrid environments.
2. Provide technical oversight across compute, storage, virtualisation, operating systems, backup, disaster recovery (DR), high availability (HA), and related infrastructure platforms.
3. Lead incident management for complex issues, including technical deep dives, root cause analysis, recovery planning, and timely resolution follow-up.
4. Manage and coordinate vendor activities to ensure operational tasks, incidents, service requests, and technical deliverables are completed effectively and within expected timelines.
5. Review and assess technical changes, service requests(SRs), implementation plans, and recovery approaches to ensure they are practical, supportable, and aligned with operational, security, and architectural requirements.
6. Support technical and security governance activities, including change review, risk awareness, compliance alignment, and operational readiness checks.
7. Support infrastructure modernisation and the multi-year technology refresh programme while maintaining stable day-to-day operations with minimal downtime.
8. Plan, coordinate, and track technical infrastructure activities across internal teams and vendors for refresh, upgrade, migration, and other technical project initiatives.
9. Track project risks, issues, dependencies, action items, and status updates to ensure milestones, operational commitments, and implementation readiness are met.
10. Work with application, network, security, and vendor teams to resolve platform-related issues and maintain runbooks, SOPs, governance artefacts, and supporting documentation.
Requirements
1. At least 5 years of experience in infrastructure or platform operations
2. Experience supporting production or mission-critical environments
3. Strong technical understanding of infrastructure operations across virtualisation, operating systems, storage, backup, networking, and related platforms
4. Experience working with vendor-supported operating models and managing external support teams
5. Able to participate in technical troubleshooting, issue deep dives, and problem resolution for complex incidents
6. Experience supporting technical governance, security governance, change review, or service request approval processes is preferred
7. Experience supporting technical project delivery, including coordination, issue and risk tracking, dependency management, and stakeholder follow-up
8. Good to have data center operations experience
9. Strong communication and coordination skills with both technical and non-technical stakeholders
Technical Requirements
1. Experience with infrastructure platforms such as VMware, Hyper-V, or Nutanix
2. Experience with Red Hat Linux and/or Windows Server environments
3. Familiarity with enterprise storage, backup, HA, and DR solutions
4. Strong understanding of monitoring, observability, logging, and alerting practices
5. Solid understanding of networking fundamentals such as TCP/IP, DNS, routing, firewalls, load balancing, and segmentation
6. Familiarity with hybrid infrastructure environments, including AWS or Azure
Desired Technical Skills
1. Experience with automation tools such as Ansible, Puppet, Chef, or Terraform
2. Familiarity with scripting languages such as Python, PowerShell, or Bash
3. Familiarity with containerisation and orchestration platforms such as Docker or Kubernetes
4. Experience with Infrastructure as Code (IaC) for automated provisioning and configuration
Job ID: 151981271
Skills:
Switches, Wlan, Dns, Group Policy, Firewalls, DHCP, Vulnerability Assessments, Windows Servers, Lan, Routers, Network Design, network security policies, Storage Solutions, endpoint protection, Active Directory, intrusion detection, virtualization platforms, Access Control, VPN infrastructure, cybersecurity best practices
Skills:
VMware, Backup, Windows Server, Logging, Dns, Firewalls, routing, Docker, Terraform, Nutanix, Python, AWS, PowerShell, Bash, Red Hat Linux, High Availability, Ansible, Load Balancing, Puppet, Azure, Kubernetes, Hyper-V, Infrastructure as Code, Chef, Disaster Recovery, alerting, Segmentation, observability, Monitoring, enterprise storage
Skills:
VMware, Windows Server, Backup, Dns, Docker, Terraform, Nutanix, Python, AWS, Logging, Load Balancing, PowerShell, Routing, Bash, Red Hat Linux, Ansible, Firewalls, Puppet, Azure, Kubernetes, Hyper-V, Alerting, Segmentation, Chef, enterprise storage, Networking fundamentals, Observability, Monitoring
Skills:
VMware, Cortex, Switches, Windows Server Administration, Network Printer, Palo Alto, Vpn, Group Policies, Fortigate Firewall, Veeam Backup, Qualys, Active Directory, Network Storage
Skills:
Ansible, PowerShell, Splunk, Cyberark, Python, Nessus