SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Volta is a vertically integrated AI infrastructure platform building compute utilities at hyperscale. Launched with a $10B strategic partnership with a leading frontier AI lab, Series A funding from Andreessen Horowitz, and a $5B AI Infrastructure Fund, the company operates AI factories across multiple regions and is scaling from 100+ to hundreds of employees.
Reporting to the Regional Data Centre Operations Director, EMEA, you will lead the deployment, operation, and continuous improvement of IT infrastructure across Volta's AI factories in Europe, Middle East, and Africa. You will oversee highly available, secure, and scalable compute, network, and facilities-adjacent technology services across multiple data centre campuses, combining strong technical expertise with operational leadership, commercial judgment, vendor management, and large-scale programme delivery.
Key responsibilities include:
**Infrastructure Leadership**: Own deployment, operation, maintenance, and lifecycle management of large-scale data centre IT infrastructure across EMEA. Lead infrastructure operations across compute, storage, networking, and platform services. Define standards for resilient, secure, and repeatable deployments. Ensure IT infrastructure and site teams—both insourced and outsourced—are ready to support high-availability services. Partner with the CTO architecture team, operations engineering, security, facilities, and product teams. Oversee providers and integrators, holding them to strict KPIs and ensuring customer SLAs are met.
**Data Centre Operations**: Oversee operational readiness and performance of data centre IT infrastructure. Establish and maintain processes for installation, maintenance, upgrades, decommissioning, and asset lifecycle management. Ensure effective capacity planning for network connectivity, infrastructure, and hardware availability. Manage incident, problem, change, and event management processes. Lead or support large-scale event response, service restoration, root cause analysis, and corrective action programmes. Maintain operational documentation, playbooks, escalation procedures, and disaster recovery plans. Ensure infrastructure environments are audit-ready and meet regulatory, contractual, and internal control requirements.
**Reliability, Security, and Resilience**: Set service-level objectives, KPIs, and operational health metrics for IT infrastructure services. Drive improvements in uptime, mean time to detect, mean time to recover, change success rate, capacity utilisation, and operational risk. Ensure IT infrastructure supports business continuity, disaster recovery, backup, replication, and recovery testing. Promote a culture of resilience engineering, preventative maintenance, and continuous improvement. Contribute to and review the EMEA technical risk register.
**Automation and Operational Excellence**: Reduce manual intervention through automation and standardised tooling. Define IT engineering and operational standards for repeatable deployments across multiple sites and regions. Use data and operational metrics to identify bottlenecks, recurring failures, capacity risks, and optimisation opportunities. Establish effective governance for IT infrastructure changes, technology standards, and platform roadmaps.
**People and Stakeholder Management**: Promote a generative safety culture, inclusive leadership, operational discipline, and knowledge sharing. Hire, lead, develop, and retain a high-performing team. Define team structures, responsibilities, career paths, performance expectations, and succession plans. Build strong partnerships with the CTO organisation and service owners. Communicate technical risks, investment needs, service performance, and operational priorities to senior leadership. Coordinate with vendors, contractors, managed service providers, and hardware manufacturers.
**Financial and Vendor Management**: Contribute to infrastructure operating and capital expenditure planning. Manage supplier relationships, SLAs, contracts, procurement activity, and vendor performance. Identify opportunities to reduce total cost of ownership without compromising availability, security, or service quality. Track budgets, forecasts, resource utilisation, and programme benefits.
Success will be measured by infrastructure availability, reliability, and performance against agreed service objectives; fewer major incidents and unplanned outages; better incident response and recovery times; capacity and infrastructure expansion programmes delivered on schedule and within budget; improved IT infrastructure utilisation and total cost of ownership; compliance with security and operational control requirements; team engagement, retention, and leadership effectiveness; and strong stakeholder partnerships.
**Requirements**
Essential:
- Bachelor's degree in computer science, information technology, or related field, or equivalent practical experience
- Typically 10 or more years' experience in IT infrastructure, data centre operations, cloud infrastructure, systems engineering, network engineering, or related discipline
- Several years managing technical teams and leading large-scale IT infrastructure operations or transformation programmes
- Track record of operating business-critical or mission-critical infrastructure at enterprise or hyperscale
- Strong knowledge of data centre technologies, including servers, storage, networking, and monitoring
- Experience with disaster recovery, capacity planning, incident management, and change management
- Experience leading cross-functional teams and influencing senior stakeholders
- Excellent written, verbal, analytical, and organisational skills
- Able to take part in or oversee an on-call and major incident response rota
Desirable:
- Experience in hyperscale cloud, colocation, internet service provider, telecoms, or large enterprise data centre environments
- Managing IT infrastructure across multiple regions, countries, or geographically distributed sites
- Public cloud platforms and hybrid or multi-cloud operating models
- Knowledge of data centre facilities interfaces, including power, cooling, environmental monitoring, physical security, and structured cabling
- IT service management frameworks, reliability engineering, and ISO 27001, SOC 2, NIST, or similar control frameworks
- Relevant certifications such as ITIL, CISCP, CCNP, PMP, AWS, Azure, Google Cloud, or equivalent