Back to jobs
New

Global Controls GPU Network and Facilities Operations Lead

Singapore

About Nscale

Nscale is taking on the hyperscalers by building a vertically integrated cloud built for AI. We own the data centres, software, and applications that power today’s AI stack using sustainable technology solutions. We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency.

As a Nscaler, you’ll build trust through openness and transparency, where everyone is inspired to do their best work. Collaboration is key, and we work together swiftly and respectfully, embracing adaptability and resilience in all we do.

Global Controls GPU Network and Facilities Operations Lead

The Global Controls, GPU Network, and Facilities Operations Lead is a senior technical contributor who supports the integration of controls systems, GPU network operations, and facilities operations across Nscale’s AI data centre portfolio. The role serves as a technical point of contact within a global team of subject-matter experts and works closely with Engineering, Construction, Commissioning, Operations, Controls, Networking, Security, OEMs, and external partners.

The lead provides technical guidance, reviews designs and operational requirements, coordinates specialist input across disciplines, and identifies integration risks. The role does not act as the single technical authority or sole point of contact. Accountable functional owners and Nscale governance processes retain final decision-making authority for their respective systems and programs.

The role supports the development and consistent application of engineering standards, operational readiness practices, telemetry and observability approaches, and high-density AI data centre deployment methods. It contributes practical expertise across controls, facility infrastructure, GPU networking interfaces, commissioning, and operations.

What You’ll be Doing

Global Technical Coordination and Standards

  • Act as a technical point of contact for assigned projects, workstreams, and stakeholders, coordinating with other designated technical points of contact and relevant global SMEs
  • Contribute specialist input to global strategies, reference architectures, and standards for controls systems, GPU network operations interfaces, NOC/FOC integration, and facilities readiness
  • Support the development and application of repeatable AI data centre deployment approaches, including modular and distributed-compute environments
  • Identify cross-functional risks, dependencies, and opportunities for improvement, and escalate material matters to the appropriate accountable functional owner

Controls Engineering and Commissioning Support

  • Provide SME review and technical support for the design, deployment, and lifecycle governance of electrical controls, BMS, EPMS, SCADA, and PLC/RTU architectures
  • Support the development and implementation of controls commissioning strategies for high-density AI facilities in partnership with project commissioning leads and system owners
  • Provide technical input to electrical design reviews, failure mode analysis, system integration, functional performance criteria, and corrective-action planning
  • Coordinate with global controls, network, and cybersecurity SMEs to ensure solutions align with approved standards and operating practices

GPU Network and NOC Integration

  • Serve as a technical interface between facility controls, telemetry, operations, and GPU network teams to support integrated monitoring, observability, and response practices
  • Support reviews of GPU fabric, fiber, and high-bandwidth interconnect requirements where they interface with data centre infrastructure, telemetry, and operational readiness
  • Contribute to NOC observability, telemetry, and responsiveness requirements in collaboration with accountable GPU network and NOC owners
  • Help identify facility-to-compute dependencies that affect GPU training and inference workloads, including power, cooling, alarm, and operating-envelope information

Facilities Operations and FOC Support

  • Support FOC and site operations stakeholders with controls, telemetry, power, cooling, and critical-environment integration requirement
  • Contribute technical input to operational-readiness reviews, MOP/SOP/EOP governance, incident response practices, and resiliency standards
  • Assist in harmonizing controls and facilities operational practices across global data center sites, including modular mechanical and electrical systems, immersion cooling platforms, and high-density electrical distribution

Data Centre Deployment and AI Factory Programs

  • Support the integration of electrical, mechanical, controls, facilities operations, and GPU networking requirements into standardized deployment templates
  • Provide technical SME support to owner’s engineering, design, construction, and commissioning teams for high-density AI data centre programs
  • Review technical compliance, constructability, systems integration, and commissioning readiness within the role’s assigned workstreams

Stakeholder Management and Knowledge Sharing

  • Coordinate with internal stakeholders, customers, utilities, OEMs, vendors, consultants, and regulatory bodies as appropriate to assigned workstreams
  • Provide clear technical communication on risks, design decisions, dependencies, and recommended actions, while escalating decisions outside the role’s authority to the accountable owner
  • Capture lessons learned and contribute to knowledge sharing, reference designs, standards, and the development of technical capability across the global team

About You

  • Bachelor Degree in Electrical Engineering, Computer Engineering, Controls/Automation Engineering or equivalent
  • 10+ years of experience in mission-critical infrastructure, electrical engineering, controls systems, data centre operations, GPU infrastructure, or comparable complex technical environments
  • Demonstrated experience supporting global or multi-site engineering, operations, commissioning, or infrastructure programs
  • Strong technical understanding of high-density power systems, medium- and low-voltage distribution, UPS, backup generation, and advanced cooling systems
  • Experience with controls systems such as BMS, EPMS, SCADA, and PLCs; commissioning practices; and reliability engineering
  • Working understanding of GPU fabrics, high-bandwidth fiber networks, or comparable high-performance-compute environments, particularly at their interface with facility infrastructure
  • Strong communication and collaboration skills across technical disciplines, delivery teams, vendors, and operational stakeholders.

Good to have

  • Professional Engineering designation or equivalent professional registration
  • Project Management Professional certification or equivalent project-delivery experience
  • Experience with NVIDIA GPU architectures, large-scale cluster operations, or next-generation compute platforms
  • Experience supporting owner’s engineering roles for large-scale or rapidly deployed data centre development
  • Familiarity with OT networking, telemetry, observability, and cybersecurity practices for critical infrastructure

Key Competencies

  • Cross-domain controls, infrastructure, and operations integration
  • Technical coordination within a global SME team
  • Controls systems and commissioning support
  • GPU network and facility infrastructure interface awareness
  • Reliability, risk identification, and structured problem solving
  • Clear stakeholder communication and disciplined escalation
  • Standards development and continuous improvement

Why this role matters

This role is central to Nscale’s ability to scale AI infrastructure reliably in the US. As high-density compute demand grows, strong facilities leadership is critical to keeping sites resilient, efficient, and ready for expansion. Your work will directly influence uptime, operational excellence, and Nscale’s ability to deliver trusted AI infrastructure at scale.

What We Can Offer You

You’ll have the opportunity to help shape the operating standards behind a next-generation AI cloud platform, working on complex infrastructure challenges with real ownership and impact. This is a chance to play a meaningful role in scaling high-performance, sustainable data centre operations in a fast-moving environment.

Equal Opportunities Statement

We strongly encourage applications from people of colour, the LGBTQ+ community, people with disabilities, neurodivergent people, parents, carers, and people from lower socio-economic backgrounds.

If there’s anything we can do to accommodate your specific situation, please let us know.

The responsibilities outlined in this job description are not exhaustive and are intended to provide a general overview of the position. The employee may be required to perform additional duties, tasks, and responsibilities as assigned by management, consistent with the skills and qualifications required for the role.

For information on how Nscale handles candidate personal data, please see our Employee & Candidate Privacy Notice: Here.

Apply for this job

*

indicates a required field

Phone
Resume/CV*

Accepted file types: pdf, doc, docx, txt, rtf

Cover Letter

Accepted file types: pdf, doc, docx, txt, rtf


Select...

Nscale uses AI-powered tools to assist in reviewing and prioritising applications against the requirements of this role. All final hiring decisions are made by humans. To learn more about how AI is used and your rights, click "Learn more" below.

Learn more