Infrastructure Operations Manager
NAVNscale
📍 Sauda, Rogaland
📂 Kontor og økonomi
📅 Publisert 1 uke siden
Infrastructure Operations Manager
NAVNscale
📍 Sauda, Rogaland
📂 Kontor og økonomi
📅 Publisert 1 uke siden
Stillingsbeskrivelse
About the Role
We are hiring an Infrastructure Operations Manager to lead the day-to-day operation of one or more of Nscale's AI data centre facilities.
Reporting to the Head of Infrastructure Operations, you will be responsible for ensuring operational excellence across critical infrastructure, leading a team of Infrastructure Operations Operators, and delivering exceptional service against operational KPIs and customer SLAs.
You will oversee site operations, people leadership, infrastructure reliability, vendor management, and operational readiness while partnering closely with Engineering, Deployment, Supply Chain, and Customer Operations teams to support the continued growth of Nscale's global AI cloud platform.
This role includes participation in an on-call rotation and occasional travel to other Nscale data centre locations.
What you'll be doing
Site Operations & Infrastructure Management
• Lead the day-to-day operation of Nscale's AI data centre infrastructure.
• Ensure high availability, operational reliability, and performance across critical infrastructure supporting AI workloads.
• Oversee hardware installation, maintenance, break-fix activities, and infrastructure lifecycle management.
• Monitor power, cooling, environmental systems, and operational health, proactively addressing risks before they impact service.
• Ensure all operational activities are performed safely and in accordance with established procedures and SLAs.
Team Leadership & Development
• Lead, coach, and develop a team of Infrastructure Operations Operators.
• Foster a culture of ownership, accountability, safety, and continuous improvement.
• Plan shift coverage, resource allocation, and on-call schedules to maintain 24/7 operational support.
• Support recruitment, onboarding, training, and career development for team members.
Customer & Vendor Management
• Act as the primary operational point of contact for customers and key stakeholders.
• Provide regular reporting against operational KPIs, SLAs, and service performance.
• Manage relationships with vendors and contractors supporting infrastructure maintenance and upgrades.
• Coordinate procurement activities and ensure timely delivery of hardware and operational services.
Inventory & Asset Management
• Maintain accurate inventory of spare parts, operational tooling, and infrastructure assets.
• Monitor stock levels and coordinate replenishment to minimize operational risk.
• Ensure asset records remain accurate throughout the infrastructure lifecycle.
Technical Leadership
• Maintain strong technical understanding of HPC, GPU infrastructure, networking, and data centre operations.
• Support troubleshooting complex infrastructure issues and coordinate with specialist engineering teams where required.
• Stay current with emerging technologies and operational best practices within AI infrastructure.
Operational Excellence & Reporting
• Monitor operational performance using defined KPIs and reporting frameworks.
• Produce regular reports on infrastructure availability, operational performance, incidents, and service improvements.
• Identify opportunities to improve reliability, scalability, efficiency, and cost optimization.
• Drive continuous improvement initiatives across operational processes and standards.
About You
Required Experience
• 5+ years of experience managing infrastructure operations or data centre environments.
• Proven experience leading technical operations teams within mission-critical environments.
• Strong understanding of HPC, GPU infrastructure, server hardware, and data centre operations.
• Experience managing operational performance against customer SLAs and KPIs.
• Excellent leadership, coaching, and stakeholder management skills.
• Strong customer-facing communication and reporting experience.
• Experience managing inventory, operational assets, and vendor relationships.
Technical Knowledge
• Strong understanding of power, cooling, networking, environmental monitoring, and data centre infrastructure.
• Experience supporting GPU and HPC deployments.
• Ability to troubleshoot infrastructure issues and coordinate cross-functional technical teams.
• Familiarity with operational monitoring, reporting, and infrastructure management tools.
Preferred Experience
• Certifications such as CDCP, CDCS, or equivalent data centre qualifications.
• Experience with NVIDIA GPU platforms, CUDA, and AI infrastructure.
• Familiarity with hybrid cloud or HPC environments.
• Experience supporting hyperscale or enterprise data centre operations.
Leadership & Attributes
• Calm and decisive under pressure.
• Strong operational mindset with excellent organizational skills.
• Passion for developing people and building high-performing teams.
• Highly collaborative with the ability to influence multiple functions.
• Strong sense of ownership and commitment to operational excellence.
What we can offer you
At Nscale, you'll find a collaborative, supportive, and innovative environment where your contributions spark real impact. We're building something extraordinary, and we want you at the core.
Highly competitive package (base + equity) with reviews every 12 months. 🚀
Join one of the fastest-growing AI infrastructure companies --- your opportunity to lead mission-critical operations powering the next generation of AI. ✨
Expect a dynamic progression plan tailored to your ambitions. Grow by leading exceptional teams, driving operational excellence, and helping scale one of the world's most advanced AI infrastructure platforms.
Human-First Flexibility: We treat you as humans first. 🫶🏽 Our flexible workplace trusts Nscalers to deliver, giving you the autonomy to shape your day around life's moments.
Remote-first collaboration with the option for office-based work. Geography is no barrier to impact or connection.
We are hiring an Infrastructure Operations Manager to lead the day-to-day operation of one or more of Nscale's AI data centre facilities.
Reporting to the Head of Infrastructure Operations, you will be responsible for ensuring operational excellence across critical infrastructure, leading a team of Infrastructure Operations Operators, and delivering exceptional service against operational KPIs and customer SLAs.
You will oversee site operations, people leadership, infrastructure reliability, vendor management, and operational readiness while partnering closely with Engineering, Deployment, Supply Chain, and Customer Operations teams to support the continued growth of Nscale's global AI cloud platform.
This role includes participation in an on-call rotation and occasional travel to other Nscale data centre locations.
What you'll be doing
Site Operations & Infrastructure Management
• Lead the day-to-day operation of Nscale's AI data centre infrastructure.
• Ensure high availability, operational reliability, and performance across critical infrastructure supporting AI workloads.
• Oversee hardware installation, maintenance, break-fix activities, and infrastructure lifecycle management.
• Monitor power, cooling, environmental systems, and operational health, proactively addressing risks before they impact service.
• Ensure all operational activities are performed safely and in accordance with established procedures and SLAs.
Team Leadership & Development
• Lead, coach, and develop a team of Infrastructure Operations Operators.
• Foster a culture of ownership, accountability, safety, and continuous improvement.
• Plan shift coverage, resource allocation, and on-call schedules to maintain 24/7 operational support.
• Support recruitment, onboarding, training, and career development for team members.
Customer & Vendor Management
• Act as the primary operational point of contact for customers and key stakeholders.
• Provide regular reporting against operational KPIs, SLAs, and service performance.
• Manage relationships with vendors and contractors supporting infrastructure maintenance and upgrades.
• Coordinate procurement activities and ensure timely delivery of hardware and operational services.
Inventory & Asset Management
• Maintain accurate inventory of spare parts, operational tooling, and infrastructure assets.
• Monitor stock levels and coordinate replenishment to minimize operational risk.
• Ensure asset records remain accurate throughout the infrastructure lifecycle.
Technical Leadership
• Maintain strong technical understanding of HPC, GPU infrastructure, networking, and data centre operations.
• Support troubleshooting complex infrastructure issues and coordinate with specialist engineering teams where required.
• Stay current with emerging technologies and operational best practices within AI infrastructure.
Operational Excellence & Reporting
• Monitor operational performance using defined KPIs and reporting frameworks.
• Produce regular reports on infrastructure availability, operational performance, incidents, and service improvements.
• Identify opportunities to improve reliability, scalability, efficiency, and cost optimization.
• Drive continuous improvement initiatives across operational processes and standards.
About You
Required Experience
• 5+ years of experience managing infrastructure operations or data centre environments.
• Proven experience leading technical operations teams within mission-critical environments.
• Strong understanding of HPC, GPU infrastructure, server hardware, and data centre operations.
• Experience managing operational performance against customer SLAs and KPIs.
• Excellent leadership, coaching, and stakeholder management skills.
• Strong customer-facing communication and reporting experience.
• Experience managing inventory, operational assets, and vendor relationships.
Technical Knowledge
• Strong understanding of power, cooling, networking, environmental monitoring, and data centre infrastructure.
• Experience supporting GPU and HPC deployments.
• Ability to troubleshoot infrastructure issues and coordinate cross-functional technical teams.
• Familiarity with operational monitoring, reporting, and infrastructure management tools.
Preferred Experience
• Certifications such as CDCP, CDCS, or equivalent data centre qualifications.
• Experience with NVIDIA GPU platforms, CUDA, and AI infrastructure.
• Familiarity with hybrid cloud or HPC environments.
• Experience supporting hyperscale or enterprise data centre operations.
Leadership & Attributes
• Calm and decisive under pressure.
• Strong operational mindset with excellent organizational skills.
• Passion for developing people and building high-performing teams.
• Highly collaborative with the ability to influence multiple functions.
• Strong sense of ownership and commitment to operational excellence.
What we can offer you
At Nscale, you'll find a collaborative, supportive, and innovative environment where your contributions spark real impact. We're building something extraordinary, and we want you at the core.
Highly competitive package (base + equity) with reviews every 12 months. 🚀
Join one of the fastest-growing AI infrastructure companies --- your opportunity to lead mission-critical operations powering the next generation of AI. ✨
Expect a dynamic progression plan tailored to your ambitions. Grow by leading exceptional teams, driving operational excellence, and helping scale one of the world's most advanced AI infrastructure platforms.
Human-First Flexibility: We treat you as humans first. 🫶🏽 Our flexible workplace trusts Nscalers to deliver, giving you the autonomy to shape your day around life's moments.
Remote-first collaboration with the option for office-based work. Geography is no barrier to impact or connection.
Stillingsdetaljer
- Kategori
- Kontor og økonomi
- Sted
- Sauda, Rogaland
- Arbeidstid
- Heltid
- Arbeidssted
- På arbeidsplass
- Ansettelsestype
- Fast stilling
- Publisert
- 1 uke siden
Om bedriften
N
Nscale
Andre stillinger innen samme område
Sauda kommune Sauda DMS
Sauda, Rogaland
Heltid
Nscale
Sauda, Rogaland
Heltid