--- title: 'Platform Site Reliability Engineer at Radiant' canonical: 'https://feeny.ai/job/platform-site-reliability-engineer-radiant-gloucestershire-xak4gfth6thv' type: 'job' last_seen: '2026-09-13' --- # Platform Site Reliability Engineer at Radiant - **Company:** Radiant - **Location:** Gloucestershire - **Employment:** full-time - **Work type:** hybrid - **Posted:** 2026-05-06 - **Last confirmed live:** 2026-09-13 - **Apply:** https://jobs.ashbyhq.com/radiant/084e5679-435e-4bd4-9a36-a17f5d47d574/application **Skills:** Kubernetes, Linux Administration, Ubuntu, Bash, Python, Ansible, Prometheus, Grafana, TCP/IP, DNS, DHCP, VLANs, Routing, Switching, Infrastructure Scripting, ITSM, AI workloads orchestration > Radiant is seeking a Senior Platform Site Reliability Engineer to deploy and manage Kubernetes clusters for AI-centric workloads. The role involves optimizing Linux systems, building automation scripts, maintaining observability stacks, and supporting 24x7 production environments while mentoring junior engineers. ## Job description ## About Radiant Radiant is redefining how AI infrastructure is built. We design and operate AI-native cloud platforms engineered for sovereignty, performance, and scale. Our infrastructure powers GPU-native workloads, multi-tenant control planes, and high-performance AI systems designed for the most demanding environments. We are not building a generic cloud. We are building purpose-built AI infrastructure - from powered land, to compute, to software . As we scale our platform and expand our engineering organisation, we are looking for leaders who can build strong teams, uphold high standards, and deliver reliably at pace. ## Role Responsibilities - Deploy and Manage Kubernetes Clusters, deployed at scale to support AI centric workloads, across both our bare metal clusters and via trusted partner infrastructure - Develop Kubernetes Manifests and Operators: Facilitate application deployments and maintain Kubernetes-native services for networking, storage, security, identity and infrastructure management - Optimize Linux system configuration including kernel, driver, filesystem and services to support workloads running via our orchestration layer - Build and maintain automation scripts and infrastructure as code to support platform lifecycle, as well as simplifying troubleshooting for Incident resolution and provision of tooling for our support organisation - Apply ITSM frameworks: Incident, Major Incident, Change Management, and service improvement. - Maintain and enhance Radiant’s observability stack: Prometheus, Grafana, and custom monitoring integrations - Operate and support services in 24x7 production environments, including on-call rotation - Contribute to Incident postmortem analyses, root cause analysis, document learnings, and automate remediations - Mentor junior engineers and act as an Operational requirements consultant to other departments - Communicate technical decisions clearly to non-technical stakeholders and customers - Uphold a culture of: do, document, automate - Willingness to cross train with Platform Engineering/Platform SRE to fully support both our infrastructure and platform stacks. - Willingness to cross train with HPC Engineering, supported by NVIDIA to enhance our HPC supportability offering ## Requirements - 5+ Years Proven experience in globally scaled, performance-intensive environments operating to a 24/7 support model in an SRE or equivalent role - 3+ years experience in both running, deploying and optimising orchestration platforms with a strong emphasis on Kubernetes - Expert-level Linux administration, especially Ubuntu distributions - Proficiency in system tuning, disk I/O optimization, and hardware-level performance tweaks - Strong networking fundamentals: TCP/IP, DNS, DHCP, VLANs, routing, switching - Strong experience with API interrogation - Strong experience with infrastructure scripting and automation (Bash, Python, Ansible) - Deep understanding of observability principles and tools (Prometheus, Grafana preferred) - Strong grasp of ITSM and service operation best practices - Excellent communication and mentorship skills - Comfortable interfacing with internal stakeholders and external customers - Bonus: Knowledge of running AI workloads via orchestration platforms Bonus Requirements - Bachelor or Masters Level degree in Computer Science, Engineering or related field, or equivalent experience. - LPIC Certifications - ITIL Foundation level qualification or equivalent experience - Certified Kubernetes Administrator (CKA) Qualities we look for: - You approach problems with a systems mindset - balancing practical execution with long-term scalability - You elevate the team, setting high standards for technical quality and engineering excellence. - You hold yourself and others accountable - giving direct feedback and expecting the same - You take initiative, owning challenges end-to-end and proactively driving solutions. - You invest in others, mentoring to build both capability and confidence. Why should you join us? What sets us apart is our blend of modern technology, competitive benefits, and an open, welcoming work culture that enables our people to thrive. Here are just some of the great things you can expect from us: - 25 days of annual leave - A culture that emphasises results over hierarchy, process & ego: we place great emphasis on the quality, ingenuity and creativity of work. - Open communication, regular feedback: we value smooth collaboration, direct and actionable feedback, and believe that leading with empathy and a growth mindset makes us better together. - Learning Time: we all have dedicated learning time to focus on new skills, projects or interests that lay outside of your day-to-day job. - Health & Wellbeing: we want everyone to feel healthy and happy, so we offer private medical insurance via Bupa. - Cycle to Work Scheme: we're committed to building a sustainable business, so we encourage cycling to work. - Gympass subscription to a variety of gyms and wellbeing apps - Participation in the company shares program - Enhanced parental pay & leave Diversity, Equality, Inclusion and Belonging We are an equal opportunity employer and we strive to reduce unconscious bias throughout our hiring process. All applicants will be considered for employment without attention to ethnicity, religion, sexual orientation, gender identity, family or parental status, national origin, veteran, neurodiversity status or disability status. To ensure our recruitment processes provide an equal opportunity for all applicants to succeed, we encourage you to let us know if there are any adjustments that we can make. ## About Radiant ## Company Overview - **One-liner**: Radiant is a vertically integrated AI infrastructure company that builds and operates AI factories combining utility-scale powered land, long-term capital, and a proprietary software platform. - **Entity Type**: Private (portfolio company of Brookfield’s AI Infrastructure Fund) - **Headquarters**: London, United Kingdom - **Founded**: 2026 (formed via merger of Brookfield’s Radiant entity with Ori Industries) - **Founders**: Mahdi Yahya (Founder & former CEO of Ori, now President of Radiant) ## Core Business - **Primary industry**: AI Infrastructure / Cloud Computing / Data Centers - **Target customers**: Sovereign governments, large enterprises, telecommunications providers (B2B, Enterprise, Sovereign) - **Mission or purpose statement**: “Building the utility model for AI compute — ubiquitous, always on and offering superior economics. That model will power the intelligence revolution — giving every nation, enterprise and network the foundation to build what comes next.” [radiant.co](https://radiant.co/about) ## Products & Services - **Radiant AI Cloud**: On-demand AI compute platform offering pre-configured bare metal offerings, GPU instances, and AI-as-a-Service offerings including Inference, Fine-Tuning, Model Registry, Kubernetes, and high-performance Storage. - **AI Factories (Sovereign/Enterprise Deployments)**: Purpose-built, vertically integrated AI data centers built on the NVIDIA DSX reference design, offering utility-grade economics under long-term contracts for sovereign governments and select enterprises. - **Powered Land Portfolio**: Access to over 5 GW live and 45 GW of renewable generation capacity globally — enabling rapid deployment of massive AI compute clusters at 20% below-market power costs through a mix of hydro, wind, geothermal, biomass, and dispatchable natural gas. [radiant.co](https://radiant.co/) - **Proprietary Software Platform**: Dependency-free, lightweight architecture built on engineering first principles, featuring intelligent scheduling, automated node management, secure multi-tenancy, and a distributed control panel. Scales consistently from 10K to 100K+ GPUs. ## Market Standing - **Valuation/Market Cap**: Not publicly disclosed as a private entity. Backed by Brookfield with access to a $100 billion investment program for AI Infrastructure (Brookfield AI Infrastructure Fund). - **Key Metric**: Total funding — Backed by Brookfield, with “more than $100 billion in deployable capital” available to the fund. Radiant is the first compute deployment vehicle and second seed investment for Brookfield’s AI Infrastructure Fund. [radiant.co](https://radiant.co/press-release-launch) - **Notable Investors/Partners**: Brookfield (global alternative asset manager), NVIDIA (Cloud Partner — using NVIDIA Blackwell, GB200 NVL72, and upcoming Rubin architectures) - **Growth Signals**: Recently launched (February 24, 2026) via merger of Radiant and Ori Industries. Targeting the multi-trillion market for integrated AI factories over the next decade. Rapidly expanding Ori’s integrated AI Cloud assets with the latest NVIDIA platforms. ## Competitive Advantages - **Vertically Integrated Model**: Uniquely combines capital (Brookfield), powered land (5 GW live, 45 GW renewable capacity), proprietary software (built from first principles), and compute (NVIDIA partnership) into a single platform — “from silicon to service.” - **Cost Advantage**: 20% below-market power costs through a balanced energy mix. - **Capital Advantage**: Direct pipeline to $100B+ investment program, enabling massive, long-duration projects that competitors cannot match. - **NVIDIA DSX Reference Design**: First-mover advantage in deploying at scale using NVIDIA’s latest architectures (Blackwell, Rubin) with full design and operational ownership. - **Sovereign Focus**: Positioned as a partner for nations seeking AI infrastructure independence — a growing geopolitical priority. ## Strategic Focus - **Scale AI Factories Globally**: Rapidly deploy AI compute capacity using the NVIDIA DSX reference design for sovereign governments, enterprises, and telecoms under long-term contracts. - **Expand the Ori Global AI Cloud**: Continue to grow the on-demand AI cloud for customers needing rapid deployment and flexible capacity. - **Maintain Vertical Integration**: Control every layer from capital and energy to software and hardware, ensuring operational autonomy and superior economics. - **Long-term Planning**: Aligned with the long-term demands of the AI economy — thinking in decades, not quarters. ## Why Work Here - **Culture**: “Outcome oriented” — the company values people who get things done, dream bigger, think in longer timelines, and build the assets that make every other ambition possible. [radiant.co](https://radiant.co/about) - **Remote/Hybrid Policy**: Flexible work options, including remote and hybrid arrangements. - **Compensation**: Competitive compensation with performance-based incentives and meaningful stock options. - **Benefits**: Learning and development support, family-friendly policies with paid parental leave, generous vacation and company holidays. - **Engineering Culture**: Led by CTO Chris Galloway (previously CTO at Ori, with experience at Broadcom/Symantec) and VP Product Joao Coelho (former AWS Solutions Architect, co-founder of open-source Multy). The team is building a proprietary, dependency-free software platform from first principles — appealing to engineers who want deep technical ownership. - **Growth Opportunity**: Early-stage (launched Feb 2026) with massive capital backing — employees can shape the company’s trajectory and work on infrastructure at unprecedented scale (10K to 100K+ GPUs). - **Hiring**: Actively hiring across every department. [radiant.co](https://radiant.co/about) ## Sources 1. [radiant.co - Homepage](https://radiant.co/) 2. [radiant.co - About Page](https://radiant.co/about) 3. [radiant.co - Press Release Launch](https://radiant.co/press-release-launch) 4. [Tech.eu - Radiant and Ori Merge](https://tech.eu/2026/02/24/radiant-and-ori-merge-to-deliver-sovereign-ai-cloud-at-utility-scale/) 5. [radiant.co - Blog: The Age of Abundance in AI](https://radiant.co/blog/the-age-of-abundance-in-ai) ## Other roles at Radiant - [Senior Technical Writer](https://feeny.ai/job/senior-technical-writer-radiant-london-30044erwnn3f) — London, United Kingdom - [Senior/Principal Product Manager - Operations](https://feeny.ai/job/senior-principal-product-manager-operations-radiant-london-b5tk7jntn8sr) — London, United Kingdom - [Senior Product Manager - Storage & Networking](https://feeny.ai/job/senior-product-manager-storage-networking-radiant-london-8mezk7m8maxf) — London, United Kingdom - [Group Finance Director](https://feeny.ai/job/group-finance-director-radiant-london-0ebp1mf360jf) — London, United Kingdom - [Senior Manager, Data Center Pre-Development & Site Selection](https://feeny.ai/job/senior-manager-data-center-pre-development-site-selection-radiant-united-states-sfck8wnve53s) — United States - [Senior Manager, Regional Data Center Development, US](https://feeny.ai/job/senior-manager-regional-data-center-development-us-radiant-united-states-3ew35eabb396) — United States - [Senior Manager, Data Center EHSQ](https://feeny.ai/job/senior-manager-data-center-ehsq-radiant-london-21ftcvmcjejn) — London, United Kingdom - [Senior Manager, Data Center Commissioning](https://feeny.ai/job/senior-manager-data-center-commissioning-radiant-london-b0hkevr6zz0f) — London, United Kingdom - [Senior Manager, Data Center Design](https://feeny.ai/job/senior-manager-data-center-design-radiant-london-df24ayzf6cq6) — London, United Kingdom - [Senior Manager, Data Center ESG & Sustainability](https://feeny.ai/job/senior-manager-data-center-esg-sustainability-radiant-london-qvt6tpmyybg8) — London, United Kingdom