--- title: 'Datacentre Operations Engineer at Radiant' canonical: 'https://feeny.ai/job/datacentre-operations-engineer-radiant-london-w5xycwe74qz7' type: 'job' last_seen: '2026-09-06' --- # Datacentre Operations Engineer at Radiant - **Company:** Radiant - **Location:** London, United Kingdom - **Employment:** full-time - **Work type:** hybrid - **Posted:** 2026-08-25 - **Last confirmed live:** 2026-09-06 - **Apply:** https://jobs.ashbyhq.com/radiant/bdb45a71-d3c6-4d79-9e1d-cea8103ba2e0 ## Job description ## ABOUT US Radiant is powered intelligence, delivering scalable, high-performance compute infrastructure purpose-built for AI and HPC workloads. Operating across global data centres, we run mission-critical environments where uptime, throughput, and ultra-low latency are non-negotiable. ## ROLE OVERVIEW We are seeking a deeply technical, hardware-passionate Datacentre Operations Engineer to execute on-the-ground operations for our new East London deployment—Radiant's newest AI infrastructure site. This role focuses on delivering precise, repeatable physical practices—including advanced smart-hands support, complex cabling, and hands-on operation of high-density, air-cooled compute systems—to guarantee world-class SLAs on next-generation hardware architecture. Working closely with Infrastructure (HPC) SRE, Network Engineering, and Datacentre Strategy teams, you will uphold uncompromising standards across East London's data centre floor, spanning multiple data halls. You will live and breathe the hardware, maintaining elite facility reliability through hands-on deployment, proactive maintenance, rapid incident response, and structured break/fix execution across high-density air cooling systems, busbar-based power distribution, and next-generation GPU compute platforms. The role centres on technical execution and optimised output. You will turn global engineering standards into flawless, repeatable daily routines, continually honing on-the-ground practices to keep our most advanced hardware running at peak performance. East London's hardware platform is NVIDIA B300-class, fully air-cooled compute, delivered through a new OEM partner being onboarded to Radiant's fleet; while the interfaces and overall architecture closely align with other OEM platforms you may already know, you'll be one of the first teams operating it on the ground, so a fast learning curve and strong first-principles troubleshooting matter more than prior exposure to this specific vendor. You will be part of the founding team standing up 24x7x365 on-site coverage at East London to meet tight customer SLAs, and as Radiant expands its EMEA footprint, your relentless drive for hardware perfection and proven field expertise with high-density environments will serve as the operational blueprint to scale execution models efficiently across the region. WHAT’S IN IT FOR YOU? Join a team operating some of the world’s most advanced high-performance computing infrastructure. As a Datacentre Operations Engineer, you’ll work hands-on with cutting-edge GPU and CPU platforms — including the latest NVIDIA architectures — powering dense, large-scale compute environments used for AI, machine learning, and next-generation workloads. This is an opportunity to build expertise at the forefront of modern infrastructure, where reliability, scale, and performance matter every day. You’ll collaborate with experienced engineers across a globally distributed organisation that values openness, inclusion, technical excellence, and continuous learning. We move quickly, solve meaningful challenges, and give people the space to make an impact. If you thrive in fast-paced environments, enjoy working with advanced technology, and want to help shape the future of high-performance compute, you’ll find both challenge and opportunity here. ## YOU CAN ALSO EXPECT: - Exposure to industry-leading GPU and AI infrastructure - Opportunities to grow alongside a rapidly scaling global business - A collaborative, inclusive, and supportive engineering culture - Real ownership and the ability to influence operational excellence - Work that sits at the intersection of people, performance, and technology - A modern, flexible, globally connected workplace with ambitious goals ## KEY RESPONSIBILITIES ## HARDWARE OPERATIONS & BREAK/FIX - Quickly diagnose and resolve hardware and network issues to maximise uptime; execute structured fault isolation methodologies to drive rapid resolution - Respond to critical hardware alerts via our monitoring and observability platform; contribute to ongoing service improvement to improve monitoring capability and alert quality - Deploy and maintain HPC and AI hardware for uninterrupted operations, including hardware troubleshooting, firmware updates, and component replacement - Execute break/fix procedures for advanced hardware platforms, including GPU module exchange, component-level fault isolation, and firmware-level diagnostics - Execute or support break/fix operations on ultra-high-density compute systems including NVIDIA B300-class chassis or equivalent platforms, including GPU/fan module exchange, chassis-level fault isolation, and busbar connection/disconnection—under the direction of the Lead where qualification is in progress ## AIR COOLING OPERATIONS - Operate, monitor, and maintain high-density air cooling infrastructure in conjunction with our datacentre partner, including CRAC/CRAH units, in-row and containment cooling, and associated airflow management systems - Facilitate in conjunction with our datacentre partner, routine and corrective maintenance on air cooling systems: monitoring supply/return temperatures and airflow rates, maintaining hot-aisle/cold-aisle containment and blanking, and performing scheduled filter and component inspections - Follow and contribute to SOPs for safe working around high-density, air-cooled compute platforms - Monitor thermal performance and raise anomalies before they escalate into incidents ## CAPACITY MANAGEMENT - Contribute to site-level capacity management operations, maintaining accurate records of power, space, and cooling utilisation - Support capacity planning activities by providing accurate as-built data and flagging infrastructure changes to the Lead and relevant teams - Manage on-the-ground assets from point of purchase and delivery through lifecycle management and disposal, owning asset management within Radiant's CMDB system ## INFRASTRUCTURE & FACILITIES - Handle RMAs and support requests within Radiant's Service Level Objectives (SLOs) to meet customer contract SLAs - Contribute to ongoing maintenance, fostering compliance and leveraging strong vendor partnerships - Operate cooling, power distribution (including busbar and PDU infrastructure), and other critical data centre technologies to maintain high operational standards - Develop and maintain datacentre/hardware management SOPs, ensuring continual alignment with Radiant's governance and compliance requirements ## SERVICE & OPERATIONAL EXCELLENCE - Apply ITSM frameworks: Incident, Major Incident, Change Management, and service improvement - Operate and support services 24x7x365 for production environments as part of a structured on-site shift rotation—working 12-hour shifts across a 4-team pattern with two engineers on shift at all times—to meet tight customer SLAs - Prioritise and triage incident and smart-hands workload ensuring tight SLA coverage is maintained with a small on the ground team - Contribute to Incident postmortem analyses, root cause analysis, document learnings, and automate remediations - Communicate technical decisions clearly to stakeholders and customers - Champion a culture of: do, document, automate - Willing to cross train and upskill in Infrastructure/Platform SRE practices - Willing to travel across EMEA to support future datacentre onboarding and train in new technologies Essential Skills & Experience - Degree in Computer Science/Electrical Engineering, or 5+ years of directly relevant industry experience in data centre operations - 3+ years of experience in data centre operations, HPC, or related roles - Passion for hardware and upholding the highest operational standards on the ground - Strong communication skills in English - Proven hands-on experience with HPC/NVIDIA GPU platforms or equivalent high-density compute systems, high-performance storage, and networking - Practical knowledge of high-density air cooling systems—CRAC/CRAH operation, airflow and thermal monitoring, containment management, and associated maintenance—or strong related cooling infrastructure experience - Experience with, or demonstrable willingness and aptitude to train on, ultra-high-density compute platforms such as NVIDIA B300, busbar-based power distribution, or equivalent systems - Experience with and passion for management of compute at massive scale - Comfort operating a newly-onboarded OEM hardware platform, applying transferable knowledge from other high-density GPU/HPC OEM systems to ramp up quickly without direct prior exposure to this specific vendor - Familiarity with structured break/fix practices for complex hardware platforms, including module-level component exchange and firmware fault isolation - Expertise in hardware installation, network configuration, and low-level system maintenance, including firmware management - Knowledge of data centre environment technologies, including cooling and high-density power distribution - Understanding of capacity management principles: power, space, and cooling tracking - Strong understanding of hardware and spares management; ability to handle RMAs within defined SLOs - Understanding of HPC and AI workloads at a high level - Strong problem-solving abilities and resilience in a fast-paced environment - Strong grasp of ITSM and service operation best practices - Excellent communication skills and ability to collaborate with cross-functional, internationally dispersed teams - Comfortable interfacing with internal stakeholders and external customers - Bonus: Vendor-endorsed qualifications from NVIDIA or equivalent OEMs for high-density, air-cooled AI compute systems ## Preferred Qualifications - Knowledge of large scale private cloud deployments and capacity planning. - Qualifications in HVAC management and deployments - Certifications in relevant areas - Hardware, Networking - ITIL Foundation level qualification or equivalent experience ## #LI-AC1 ## About Radiant ## Company Overview - **One-liner**: Radiant is a vertically integrated AI infrastructure company that builds and operates AI factories combining utility-scale powered land, long-term capital, and a proprietary software platform. - **Entity Type**: Private (portfolio company of Brookfield’s AI Infrastructure Fund) - **Headquarters**: London, United Kingdom - **Founded**: 2026 (formed via merger of Brookfield’s Radiant entity with Ori Industries) - **Founders**: Mahdi Yahya (Founder & former CEO of Ori, now President of Radiant) ## Core Business - **Primary industry**: AI Infrastructure / Cloud Computing / Data Centers - **Target customers**: Sovereign governments, large enterprises, telecommunications providers (B2B, Enterprise, Sovereign) - **Mission or purpose statement**: “Building the utility model for AI compute — ubiquitous, always on and offering superior economics. That model will power the intelligence revolution — giving every nation, enterprise and network the foundation to build what comes next.” [radiant.co](https://radiant.co/about) ## Products & Services - **Radiant AI Cloud**: On-demand AI compute platform offering pre-configured bare metal offerings, GPU instances, and AI-as-a-Service offerings including Inference, Fine-Tuning, Model Registry, Kubernetes, and high-performance Storage. - **AI Factories (Sovereign/Enterprise Deployments)**: Purpose-built, vertically integrated AI data centers built on the NVIDIA DSX reference design, offering utility-grade economics under long-term contracts for sovereign governments and select enterprises. - **Powered Land Portfolio**: Access to over 5 GW live and 45 GW of renewable generation capacity globally — enabling rapid deployment of massive AI compute clusters at 20% below-market power costs through a mix of hydro, wind, geothermal, biomass, and dispatchable natural gas. [radiant.co](https://radiant.co/) - **Proprietary Software Platform**: Dependency-free, lightweight architecture built on engineering first principles, featuring intelligent scheduling, automated node management, secure multi-tenancy, and a distributed control panel. Scales consistently from 10K to 100K+ GPUs. ## Market Standing - **Valuation/Market Cap**: Not publicly disclosed as a private entity. Backed by Brookfield with access to a $100 billion investment program for AI Infrastructure (Brookfield AI Infrastructure Fund). - **Key Metric**: Total funding — Backed by Brookfield, with “more than $100 billion in deployable capital” available to the fund. Radiant is the first compute deployment vehicle and second seed investment for Brookfield’s AI Infrastructure Fund. [radiant.co](https://radiant.co/press-release-launch) - **Notable Investors/Partners**: Brookfield (global alternative asset manager), NVIDIA (Cloud Partner — using NVIDIA Blackwell, GB200 NVL72, and upcoming Rubin architectures) - **Growth Signals**: Recently launched (February 24, 2026) via merger of Radiant and Ori Industries. Targeting the multi-trillion market for integrated AI factories over the next decade. Rapidly expanding Ori’s integrated AI Cloud assets with the latest NVIDIA platforms. ## Competitive Advantages - **Vertically Integrated Model**: Uniquely combines capital (Brookfield), powered land (5 GW live, 45 GW renewable capacity), proprietary software (built from first principles), and compute (NVIDIA partnership) into a single platform — “from silicon to service.” - **Cost Advantage**: 20% below-market power costs through a balanced energy mix. - **Capital Advantage**: Direct pipeline to $100B+ investment program, enabling massive, long-duration projects that competitors cannot match. - **NVIDIA DSX Reference Design**: First-mover advantage in deploying at scale using NVIDIA’s latest architectures (Blackwell, Rubin) with full design and operational ownership. - **Sovereign Focus**: Positioned as a partner for nations seeking AI infrastructure independence — a growing geopolitical priority. ## Strategic Focus - **Scale AI Factories Globally**: Rapidly deploy AI compute capacity using the NVIDIA DSX reference design for sovereign governments, enterprises, and telecoms under long-term contracts. - **Expand the Ori Global AI Cloud**: Continue to grow the on-demand AI cloud for customers needing rapid deployment and flexible capacity. - **Maintain Vertical Integration**: Control every layer from capital and energy to software and hardware, ensuring operational autonomy and superior economics. - **Long-term Planning**: Aligned with the long-term demands of the AI economy — thinking in decades, not quarters. ## Why Work Here - **Culture**: “Outcome oriented” — the company values people who get things done, dream bigger, think in longer timelines, and build the assets that make every other ambition possible. [radiant.co](https://radiant.co/about) - **Remote/Hybrid Policy**: Flexible work options, including remote and hybrid arrangements. - **Compensation**: Competitive compensation with performance-based incentives and meaningful stock options. - **Benefits**: Learning and development support, family-friendly policies with paid parental leave, generous vacation and company holidays. - **Engineering Culture**: Led by CTO Chris Galloway (previously CTO at Ori, with experience at Broadcom/Symantec) and VP Product Joao Coelho (former AWS Solutions Architect, co-founder of open-source Multy). The team is building a proprietary, dependency-free software platform from first principles — appealing to engineers who want deep technical ownership. - **Growth Opportunity**: Early-stage (launched Feb 2026) with massive capital backing — employees can shape the company’s trajectory and work on infrastructure at unprecedented scale (10K to 100K+ GPUs). - **Hiring**: Actively hiring across every department. [radiant.co](https://radiant.co/about) ## Sources 1. [radiant.co - Homepage](https://radiant.co/) 2. [radiant.co - About Page](https://radiant.co/about) 3. [radiant.co - Press Release Launch](https://radiant.co/press-release-launch) 4. [Tech.eu - Radiant and Ori Merge](https://tech.eu/2026/02/24/radiant-and-ori-merge-to-deliver-sovereign-ai-cloud-at-utility-scale/) 5. [radiant.co - Blog: The Age of Abundance in AI](https://radiant.co/blog/the-age-of-abundance-in-ai) ## Other roles at Radiant - [Group Finance Director](https://feeny.ai/job/group-finance-director-radiant-london-0ebp1mf360jf) — London, United Kingdom - [Senior Manager, Data Center Pre-Development & Site Selection](https://feeny.ai/job/senior-manager-data-center-pre-development-site-selection-radiant-united-states-sfck8wnve53s) — United States - [Senior Manager, Regional Data Center Development, US](https://feeny.ai/job/senior-manager-regional-data-center-development-us-radiant-united-states-3ew35eabb396) — United States - [Senior Manager, Data Center EHSQ](https://feeny.ai/job/senior-manager-data-center-ehsq-radiant-london-21ftcvmcjejn) — London, United Kingdom - [Senior Manager, Data Center Commissioning](https://feeny.ai/job/senior-manager-data-center-commissioning-radiant-london-b0hkevr6zz0f) — London, United Kingdom - [Senior Manager, Data Center Design](https://feeny.ai/job/senior-manager-data-center-design-radiant-london-df24ayzf6cq6) — London, United Kingdom - [Senior Manager, Data Center ESG & Sustainability](https://feeny.ai/job/senior-manager-data-center-esg-sustainability-radiant-london-qvt6tpmyybg8) — London, United Kingdom - [Principal Procurement Lead, Compute and Storage](https://feeny.ai/job/principal-procurement-lead-compute-and-storage-radiant-london-vd7c0g6n67q3) — London, United Kingdom - [Cluster Architect](https://feeny.ai/job/cluster-architect-radiant-london-jekhn3m8e2fb) — London, United Kingdom - [Manager – FP&A](https://feeny.ai/job/manager-fp-a-radiant-london-rwb6fef3tdvv) — London, United Kingdom