Remote job
Remote confirmedSenior Manager, IT Infrastructure
Volta
Our assessment
- Our reading of the full posting text confirms it: fully remote.
- 18 more open roles from this employer in our index. 4 of them fully remote.
Automated assessment by nomado24, not an employer statement.
Job description
ABOUT VOLTA
Volta is the category-defining, fully vertically integrated AI infrastructure platform – from capital to clusters to software, under a founder-led enterprise. Our mission is The Utility of Compute™: AI infrastructure as dependable and available as electricity, for every organization that needs it. Launched with a $10B strategic partnership with one of the leading frontier AI labs, a Series A led by Andreessen Horowitz, and a $5B AI Infrastructure Fund, Volta is building the infrastructure layer of the AI era from the ground up. We are 100+ people across London, Palo Alto, and New York, with rapid growth expectations to hundreds.
About The Role
Reporting to the Regional Data Centre Operations Director, EMEA, the Senior Manager, IT Infrastructure will lead the deployment, operation and continuous improvement of IT infrastructure across Volta's AI factories in EMEA. You will oversee highly available, secure and scalable compute, network and facilities-adjacent technology services across multiple data centre campuses. The role combines strong technical expertise with operational leadership, commercial judgement, vendor management and the ability to run infrastructure programmes at scale. You will lead the IT infrastructure site teams, set operational standards, drive automation and make sure IT infrastructure services meet demanding requirements for availability, performance, security, capacity and cost efficiency.
What You Will Be Doing
Infrastructure leadership
- Own the deployment, operation, maintenance and lifecycle management of large-scale data centre IT infrastructure across EMEA
- Lead infrastructure operations across compute, storage, networking and platform services, as designed by Volta's CTO organisation
- Define standards for resilient, secure and repeatable infrastructure deployments
- Ensure Volta's IT infrastructure and site teams, both insourced and outsourced, are ready to support high-availability services
- Partner with the CTO architecture team, operations engineering, security, facilities and product teams to align infrastructure capabilities with business requirements
- Oversee Volta's providers and integrators, holding them to strict KPIs and making sure customer SLAs are fully met
Data centre operations
- Oversee the operational readiness and performance of data centre IT infrastructure environments
- Establish and maintain processes for installation, maintenance, upgrades, decommissioning and asset lifecycle management of Volta's AI platforms
- Ensure effective capacity planning for network connectivity, infrastructure and hardware availability
- Manage incident, problem, change and event management processes for IT infrastructure services
- Lead or support large-scale event response, service restoration, root cause analysis and corrective action programmes
- Maintain operational documentation, playbooks, escalation procedures and disaster recovery plans for IT infrastructure
- Ensure infrastructure environments are audit-ready and meet regulatory, contractual and internal control requirements
Reliability, security and resilience
- Set service-level objectives, KPIs and operational health metrics for IT infrastructure services, including for partners and providers SENIOR MANAGER, IT INFRASTRUCTURE VOLTA / JOB DESCRIPTION PAGE 3
- Drive improvements in uptime, mean time to detect, mean time to recover, change success rate, capacity utilisation and operational risk
- Ensure IT infrastructure supports business continuity, disaster recovery, backup, replication and recovery testing requirements
- Promote a culture of resilience engineering, preventative maintenance and continuous improvement
- Contribute to and review the EMEA technical risk register, making sure mitigation plans are documented, prioritised and delivered
Automation and operational excellence
- Reduce manual intervention through automation and standardised tooling where relevant
- Define IT engineering and operational standards for repeatable deployments across multiple sites and regions
- Use data and operational metrics to identify bottlenecks, recurring failures, capacity risks and opportunities to optimise
- Establish effective governance for IT infrastructure changes, technology standards and platform roadmaps
People and stakeholder management
- Promote a generative safety culture, inclusive leadership, operational discipline and knowledge sharing
- Hire, lead, develop and retain a high-performing team
- Define team structures, responsibilities, career paths, performance expectations and succession plans
- Build strong partnerships with your main internal customers: the CTO organisation and the service owners
- Communicate technical risks, investment needs, service performance and operational priorities clearly to senior leadership
- Coordinate with vendors, contractors, …
This role is provided by an external source. Applications are handled on the source website.
