Nebius

Nebius

Technical Program Manager, Data Centers

EMEARemotePosted 2 days ago
Full TimeSeniorRemote

See how this job matches your profile

Sign in for an AI-powered fit score, breakdown, and a tailored resume.

Sign in

Job Description

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model

Key Highlights

  • Own the operational health of Nebius COLO and BTS sites, ensuring each facility runs to expectation across power, cooling, space, connectivity, security, and environmental controls.
  • Track, monitor, and enforce SLA compliance across landlords and colocation providers; identify breaches, drive remediation, and hold providers accountable to contractual commitments.
  • Manage and coordinate site maintenance schedules — preventive and corrective — including planning and approving maintenance windows, reviewing Methods of Procedure (MOPs), and minimizing risk to live workloads.
  • Plan and drive site audits covering compliance, capacity, power/cooling performance, physical security, and safety; track findings to closure.
  • Serve as the primary day-to-day interface with data center landlords and operators, managing the operational relationship, escalations, and coordination of on-site activity.

Qualifications

Required Qualifications

  • 10+ years of experience in technical program management, data center operations, or critical facilities/infrastructure management.
  • Experience managing data center infrastructure and operations (power, cooling, space, connectivity) in colocation, build-to-suit, or owned environments.
  • Experience managing third-party vendors, landlords, or service providers against SLAs and contractual obligations.
  • Demonstrated ability to manage multiple programs, sites, or workstreams simultaneously and drive them to measurable outcomes.
  • Bachelor’s degree in a relevant field, or equivalent practical experience.
  • Direct experience with colocation (COLO) and build-to-suit (BTS) data center models, including operating across multiple landlords and operators.
  • Working knowledge of data center SLAs, MOPs/SOPs, maintenance regimes, and audit and compliance frameworks (e.g., Uptime Institute Tier standards, SOC 2, ISO 27001).
  • Experience supporting AI/HPC, GPU cluster, or other high-density compute infrastructure.
  • Strong familiarity with incident management and root cause analysis in a critical facilities context.
  • Experience building reporting mechanisms, dashboards, or operational scorecards for infrastructure health and risk.
  • PMP, Uptime ATD, or equivalent program/operations certification.
  • Proficiency with program and ticketing tools (e.g., Jira, ServiceNow) and comfort working with operational data.
  • Willingness to travel to sites as needed.

Skills & Technologies

Jira

Interested in this role?

Sign in or create a free account to see how this job matches your skills, apply with one click, and let our AI tailor your resume.

Sign in to apply
AI-powered resume optimization
Save and track your applications

Job Details

Employment Type

Full Time

Experience Level

Senior

Location

EMEA

Work Mode

Remote

Posted

2 days ago