--- title: 'Lead Site Reliability Engineer at Intellum, Inc.' canonical: 'https://feeny.ai/job/lead-site-reliability-engineer-intellum-inc-united-states-hb4tnz461176' type: 'job' last_seen: '2026-09-04' --- # Lead Site Reliability Engineer at Intellum, Inc. - **Company:** Intellum, Inc. - **Location:** United States - **Work type:** remote - **Posted:** 2026-09-01 - **Last confirmed live:** 2026-09-04 - **Apply:** https://job-boards.greenhouse.io/intelluminc/jobs/5411278008 ## Job description ## About us Intellum is the leader in corporate education technology and powers the largest, most successful customer, partner, and employee learning programs in the world. Large brands and fast-moving companies like Google, Meta, Amazon, Walmart, Xero, Atlassian, Mailchimp, Airbnb, Stripe, and TikTok rely on Intellum to engage and educate the audiences they touch. We have always been a “remote first” company and are proud to have team members located all over the world. We value Curiosity, Creativity, Perseverance, and Kindness and strive to demonstrate these core values every day. Our culture is very important to us. We invest in our people in fun and exciting ways, including personal development budgets and an annual all-company retreat that is focused less on work and more on human connections. We are in growth mode, and our “smart growth” approach ensures that we will continue to scale our company effectively. The Lead Systems Engineer is a senior individual contributor responsible for the reliability, scalability, and modernization of Intellum's platform infrastructure. Intellum serves large enterprise customers with demanding availability expectations, and this role will help shape the technical direction for how our platform runs, deploys, and scales. This is a highly hands-on role with significant ownership across infrastructure architecture, cloud environments, deployment systems, observability, and platform reliability. The Lead Systems Engineer will also provide technical leadership across the Systems Engineering function through architecture guidance, mentorship, knowledge sharing, and strong operational standards. A key focus of this role is continuing to modernize the platform toward portable, container-orchestrated infrastructure, improving deployment and observability capabilities, and maintaining an architecture that can operate effectively across multiple cloud providers. ## Responsibilities - Own and drive key infrastructure modernization initiatives, including the continued evolution from legacy compute environments toward modern, container-orchestrated infrastructure while maintaining reliable service for enterprise customers. - Design and maintain infrastructure as code across multiple cloud providers, ensuring infrastructure decisions support portability, maintainability, and long-term scalability. - Improve the reliability and maturity of Intellum's CI/CD systems and deployment tooling so releases are efficient, observable, and recoverable. - Provide technical leadership across the Systems Engineering team through mentorship, architecture guidance, knowledge sharing, and support for strong engineering practices. - Establish and evolve SLI and SLO practices, along with the monitoring, alerting, and load-testing capabilities needed to support platform reliability. - Participate in and provide leadership during platform incidents, including troubleshooting, root cause analysis, and follow-through on corrective actions. - Drive visibility into cloud infrastructure costs and incorporate cost considerations into architecture and infrastructure decisions. - Improve developer experience by evolving the infrastructure and tooling engineers depend on, including development environments, deployment workflows, and production feedback loops. - Partner closely with Security and Engineering teams on access controls, infrastructure hardening, compliance requirements, and secure infrastructure practices. - Identify operational and infrastructure risks early, recommend priorities, and help drive the technical roadmap for the Systems Engineering function. - Contribute to the continued development of the Systems Engineering team and function, including mentoring engineers and helping build strong technical practices as the organization evolves. - Perform other duties as assigned. Required Skills - 8+ years of hands-on experience in infrastructure, DevOps, platform engineering, site reliability engineering, or a related discipline, including experience building and operating production systems. - Deep hands-on experience designing, operating, and troubleshooting highly available production infrastructure. - Production experience across more than one major cloud provider, with depth in at least one of AWS or Google Cloud and working fluency in the other. - Significant experience with container orchestration and Kubernetes in production environments, including cluster operations, workload configuration, reliability, and troubleshooting. - Experience modernizing production infrastructure, including migrations from VM-based or legacy environments toward containerized or cloud-native architectures. - Strong infrastructure-as-code experience using Terraform or comparable tooling, with an emphasis on repeatability and automation. - Experience building, operating, or significantly improving CI/CD systems and deployment infrastructure. - Strong incident response and troubleshooting capabilities, including experience diagnosing complex distributed-system failures and contributing to effective post-incident review. - Strong Linux administration skills and scripting or programming ability in Ruby, Python, or a comparable language. - Experience working in a SaaS environment where reliability, availability, and production stability are critical. - Ability to collaborate effectively with distributed teams across US and European time zones and participate in an on-call rotation. - Strong communication skills and the ability to provide technical direction, mentor other engineers, and influence infrastructure decisions across teams. ## Preferred Qualifications - Experience operating production infrastructure across both AWS and Google Cloud simultaneously. - Prior experience leading or managing engineers, whether through formal people management, technical leadership, or mentorship. - Experience developing engineers and helping build strong, high-performing technical teams. - Experience with cloud cost management or FinOps practices at meaningful scale. - Experience managing deployment platforms such as Spinnaker, Jenkins, or comparable tooling. - Experience operating a Ruby on Rails enterprise application or comparable production codebase. - Familiarity with SOC 2 or similar compliance frameworks and customer-facing security requirements. - Working knowledge of AI-assisted development tooling and its infrastructure implications. - Prior people leadership or management experience in a player-coach capacity, balancing hands-on technical contribution with mentorship, team guidance, and development of engineers. - Background in learning management systems, learning technologies, or adult education platforms. Education - Bachelor's degree in a related field or equivalent practical experience. Equivalent experience is genuinely accepted for this role. ## BENEFITS - Medical - 100% of employee premiums for selected individual plans - Dental - 100% of employee premiums covered - Vision - 100% of employee premiums covered - LinkedIn Learning - 401(k) plus matching (US Based Only) - Flexible PTO - Calm subscription - Annual Company Retreat Intellum is an equal-opportunity employer. We're committed to building an inclusive team that celebrates diversity in people, perspectives, and backgrounds regardless of race, color, national origin, gender, sexual orientation, age, religion, disability, citizenship, veteran status, or any other protected status. We encourage you to apply for an open position and if you have questions about whether or not your job experience and skill set meet the requirements for a specific role, reach out to us directly at careers@intellum.com. If you are an individual applying from CA, NY, CO, CT, MD, NV, or RI, please reach out to careers@intellum.com to inquire about specific pay ranges. ## About Intellum, Inc. ## Company Overview - **One-liner**: Intellum provides an AI-first enterprise learning management system (LMS) that helps organizations accelerate content creation, automate management tasks, and improve learning outcomes through Education-Led Growth. - **Entity Type**: Private (funded; $25M total funding, Series unknown) - **Headquarters**: Atlanta, Georgia, United States - **Founded**: 2000 - **Founders**: Not publicly disclosed ## Core Business - **Primary industry**: E-Learning Providers / Enterprise Learning Management - **Target customers**: B2B, large enterprises and Fortune 500 companies (e.g., Facebook, Google) - **Mission or purpose**: “Help organizations connect knowledge to results” and rebuild the LMS with AI at the core. ## Products & Services - **[Intellum LMS](https://www.intellum.com/)**: AI-native learning management system with built-in AI agents (Creator Agent, Manager Agent, Learner Agent) that automate content creation, management, and personalization. Includes course authoring (Evolve), virtual events, certifications, and analytics. 99.9% uptime SLA, enterprise-grade security, SOC 2 Type II certified. - **[Evolve Authoring Tool](https://www.intellum.com/)**: Next-gen course authoring tool integrated natively into Intellum LMS, enabling rapid creation of engaging content with measurable effectiveness. - **[Intellum AI](https://www.intellum.com/)**: Suite of AI agents that reduce content development tasks by up to 70% and provide personalized learning paths. ## Market Standing - **Valuation/Market Cap**: Not disclosed (privately held) - **Key Metric**: Annual Revenue $25.7M (as of most recent data); Total Funding $25M (Guidepost Growth Equity, August 2023) - **Notable Investors/Partners**: Guidepost Growth Equity (lead investor); clients include Facebook (Blueprint), Google (Retail Training, Academy for Ads), and other Fortune 500 companies. - **Growth Signals**: 111 employees across 10 countries; operates remote-first; acquired Appitierre (2019) and Intellum UK Limited; continuously innovating with AI-native rebuild in 2025-2026. ## Competitive Advantages - **AI-first architecture**: AI is embedded natively into every workflow, not bolted on, enabling deep automation and personalization. - **25+ years of enterprise LMS experience**: Proven track record with Fortune 500 deployments and high reliability (99.9% uptime). - **Integrated authoring (Evolve)**: Reduces toolchain complexity and speeds up content creation. - **Remote-first, global team**: Attracts diverse talent and operates efficiently across time zones. ## Strategic Focus - Continue to deepen AI capabilities across the platform (Creator, Manager, Learner Agents). - Expand enterprise customer base and maintain leadership in customer education. - Grow the engineering and product teams to accelerate AI-native development. ## Why Work Here - **Culture**: Remote-first from inception; values include “learn and grow together,” “open and honest,” “go after big dreams,” “take ownership,” and “we are in this together.” Annual all-employee retreat (“Camptellum”) focused on fun and connection. - **Remote/Hybrid/Office**: Fully remote-first; team spans 10+ countries; no mandatory office requirement. - **Notable Perks & Benefits** (from [Intellum Careers](https://www.intellum.com/company/careers)): - 100% paid premiums for select medical, dental, and vision plans (US employees) - UK private medical insurance - HSA/FSA with employer HSA contribution - Free Calm subscription for employee + up to 5 dependents - Flexible PTO - Paid parental leave: 12 weeks (birthing parent), 4 weeks (non-birthing) - 401(k) with company match (US) / generous pension contribution (UK) - Life insurance & long-term disability - Home office stipend ($300/year) - LinkedIn Learning access - Leadership & career development programs - **Engineering Culture**: Small, collaborative team working on hard problems; “incredibly smart, kind people” who care about building something meaningful. Open roles include Lead Site Reliability Engineer, Senior Product Manager, and Product Marketing Director. ## Sources 1. [Intellum Careers Page](https://www.intellum.com/company/careers) 2. [Intellum Homepage](https://www.intellum.com/) 3. [Intellum About Us](https://www.intellum.com/company/about-us) 4. [LinkedIn Company Profile](https://www.linkedin.com/company/intellum) 5. [Greenhouse Jobs Board](https://job-boards.greenhouse.io/intelluminc) ## Other roles at Intellum, Inc. - [Join our Talent Community!](https://feeny.ai/job/join-our-talent-community-intellum-inc-united-states-fvaz9jeyc0jv) — United States - [Lead Site Reliability Engineer](https://feeny.ai/job/lead-site-reliability-engineer-mattermost-united-states-je123mpv22h1) — United States - [Lead Site Reliability Engineer](https://feeny.ai/job/lead-site-reliability-engineer-heidi-melbourne-gcszcagy808g) — Melbourne, Australia - [Lead Site Reliability Engineer](https://feeny.ai/job/lead-site-reliability-engineer-zeta-hyderabad-hcf5k1c23h3w) — Hyderabad, India - [Lead Site Reliability Engineer](https://feeny.ai/job/lead-site-reliability-engineer-zeta-hyderabad-g22fjg6fxdj9) — Hyderabad, India - [Lead Site Reliability Engineer](https://feeny.ai/job/lead-site-reliability-engineer-zeta-global-bengaluru-5nhy52e2fapg) — Bengaluru, India - [Lead Site Reliability Engineer](https://feeny.ai/job/lead-site-reliability-engineer-movable-ink-new-york-ny-4fkkqh26awhb) — New York NY, United States - [Lead Site Reliability Engineer](https://feeny.ai/job/lead-site-reliability-engineer-kontakt-io-new-york-eghqp7mc6n0e) — New York, NY - [Lead Site Reliability Engineer](https://feeny.ai/job/lead-site-reliability-engineer-nice-southampton-qf2xbb55vfg6) — Southampton, United Kingdom - [Lead Site Reliability Engineer](https://feeny.ai/job/lead-site-reliability-engineer-alloy-new-york-e7sknsrxh6ye) — New York, NY