--- title: 'Senior Infrastructure Engineer - AI/ML Platform at OpenTeams' canonical: 'https://feeny.ai/job/senior-infrastructure-engineer-ai-ml-platform-openteams-united-states-pzcrjf5jzc8f' type: 'job' last_seen: '2026-09-12' --- # Senior Infrastructure Engineer - AI/ML Platform at OpenTeams - **Company:** OpenTeams - **Location:** United States - **Work type:** remote - **Posted:** 2026-09-01 - **Last confirmed live:** 2026-09-12 - **Apply:** https://job-boards.greenhouse.io/openteams/jobs/4729695005 ## Job description ## Who We Are Every organization runs on intelligence: years of accumulated knowledge, decisions, and context. As AI takes on more of that work, companies face a choice: rent that intelligence from vendors who keep the data, the context, and the results, or own it. OpenTeams exists to make ownership possible. Founded by Travis Oliphant, creator of NumPy and SciPy, and built by people with deep roots across the open-source ecosystem, including NumPy, SciPy, PyTorch, and Jupyter, we help enterprises and governments build AI they control, govern, and evolve themselves. If that sounds like your kind of work, we'd like to meet you. Senior Infrastructure Engineer - AI/ML Platform Location: U.S - Remote Work Authorization: U.S. citizenship required Clearance: An active clearance is not required at the time of hire. Candidates must be able to obtain and maintain a Secret security clearance, which includes a federal background investigation. Salary Range: $145,000–$250,000 USD, dependent on experience level and location ## About the Role We’re seeking a Senior Infrastructure Engineer to build and operate the platform underlying secure AI test and evaluation capabilities for Government teams assessing AI systems. You’ll own a Kubernetes-based platform supporting demanding AI/ML workloads, including GPU scheduling, large-scale data movement, reproducible test execution, and multi-tenant isolation. The platform must operate reliably within Government environments that may have restricted networks, accreditation boundaries, limited connectivity, and no assumption of outbound internet access or managed cloud services. You’ll use open-source technologies such as OpenTofu, Terraform, Helm, Argo CD, Kubernetes operators, and Nebari to build reusable and composable infrastructure. You’ll also own platform reliability, observability, capacity planning, upgrade paths, hardened configurations, and the documentation needed to deploy and maintain the platform. This is a fully remote, U.S.-based role working with a distributed team that relies heavily on asynchronous communication. ## Key Responsibilities - Build and operate the Kubernetes platform supporting AI test and evaluation frameworks - Implement GPU scheduling, workload orchestration, resource management, and multi-tenant isolation across evaluation teams - Design infrastructure-as-code, GitOps workflows, and automated deployment pipelines that make the platform reproducible from source - Develop reusable and modular infrastructure components that can be composed into independently owned and operated platforms - Contribute to Nebari and other open-source Kubernetes, infrastructure, and MLOps projects used by the platform - Own platform reliability, including capacity planning, upgrade strategies, failure-mode analysis, backup and recovery considerations, and operational readiness - Design and implement observability, monitoring, logging, tracing, and alerting for large-scale AI/ML workloads - Develop operational runbooks and documentation that enable other engineers to deploy, operate, and troubleshoot the platform - Deploy, configure, and harden infrastructure within secure, restricted, disconnected, or limited-connectivity Government environments - Support security authorization and compliance activities through infrastructure documentation, hardened configurations, control evidence, and repeatable deployment processes - Integrate automated security tooling for container scanning, static and dynamic analysis, artifact signing, and policy enforcement - Collaborate with Government stakeholders, security personnel, software engineers, and ML engineers to translate platform requirements into reliable infrastructure - Provide technical leadership, contribute to engineering standards, and mentor less-experienced team members - Collaborate effectively within a remote and distributed team using asynchronous communication practices Required Skills & Experience - U.S. citizenship and ability to obtain and maintain a Secret security clearance - 6+ years of hands-on infrastructure, platform, DevOps, or site reliability engineering experience supporting production systems - Strong understanding of infrastructure engineering principles, including scalability, reliability, observability, security, and automation - Production experience with Kubernetes, including workload scheduling, resource management, and multi-tenant environments - Experience with automated security tooling, such as container scanning, SAST/DAST, artifact signing, and policy enforcement - Proficiency with infrastructure-as-code tools such as Terraform, OpenTofu, Pulumi, or equivalent technologies - Experience with at least one major cloud platform—AWS, Azure, or Google Cloud—including networking, security, storage, and compute services - Experience implementing monitoring and observability using tools such as OpenTelemetry, Prometheus, Grafana, or equivalent technologies - Strong programming or automation skills using Python, Go, or a comparable language - Experience with CI/CD practices, GitOps workflows, and infrastructure automation - Experience creating maintainable operational documentation, deployment procedures, and runbooks - Experience leading technical initiatives, establishing engineering practices, or mentoring other engineers - Ability to work independently and collaborate effectively within a remote, distributed team - Ability to give and receive constructive technical feedback ## Nice to Have - Experience deploying or operating infrastructure in air-gapped, disconnected, or highly restricted environments - Experience engineering systems subject to the Risk Management Framework, NIST 800-53, NIST 800-171, or comparable security requirements - Experience supporting security personnel in achieving an Authorization to Operate for complex systems in Government or Intelligence Community environments - Experience supporting systems operating at Impact Level 5 or higher - Current DoD 8140 qualifying certification, such as Security+, CISSP, CISM, or an equivalent IAT/IAM Level II or III credential - Experience building MLOps pipelines or infrastructure supporting AI/ML workloads - Experience with GPU scheduling, large-scale data pipelines, or reproducible ML evaluation workloads - Experience with model-serving frameworks such as KServe, vLLM, LLM-D, or equivalent technologies - Familiarity with data sovereignty, privacy, and security requirements for enterprise or Government AI systems - Contributions to open-source Kubernetes, infrastructure, MLOps, or observability projects - Experience with Nebari ## What We Offer - Medical, Dental & Vision – 100% paid for employees, 75% for dependents - 401(k) Match – Up to 5% with full vesting after 2 years - Unlimited PTO – With a required minimum of 15 days off annually - Fully Remote Setup – Includes up to $3,000 equipment reimbursement - Continuous Education –  Includes up to $500 reimbursement - Disability & Life Insurance – 100% employer-paid - HSA & FSA Options – With monthly HSA contributions from OpenTeams Grow With Us At OpenTeams, growth isn’t just about the company—it’s about you. We believe the best careers are built at the edge of your potential. That is where new tools, ideas, and technologies change the world. Here, you’ll work alongside pioneers of AI, solving problems that matter: making AI more transparent, more ethical, and more empowering. As your skills grow, our career framework provides a pathway and recognition of that increased impact. Opportunities aren’t limited by geography. You’ll collaborate with global experts, contribute to open source projects that power the world’s technology, and stretch your skills daily.  That global perspective and diversity makes our solution more universal and robust.  We are committed to continuing to celebrate diversity on our team. Supported people are successful people.  We offer 100% employer paid medical premiums for employees and self-managed PTO with a minimum time off requirement, so that our teams are able to do their best work. We invest  in curiosity, creativity, and ownership. That means you’ll be trusted to boldly innovate, supported to learn fast, and celebrated for successful collaboration. Commitment to diversity, equity, inclusion, and belonging OpenTeams understands that valuing diverse creative practices and forms of knowledge is crucial to and enriches the company’s core mission. We encourage applications from everyone, including members of all equity-seeking communities, such as (but certainly not limited to) women, racialized and Indigenous persons, disabled people, persons of all sexual orientations, gender identities and expressions. We are an equal opportunity employer - all qualified applicants will receive equal consideration for recruitment, interviews, employment, training, compensation, promotion, and related activities. We do not discriminate based on race, religion, gender, gender identity, gender expression, color, national origin, pregnancy, ancestry, domestic partner status, disability, sexual orientation, age, genetic predisposition, medical condition, marital status, citizenship status, military or veteran status, or any other basis covered by applicable laws. OpenTeams will not tolerate discrimination or harassment based on these characteristics or any other unlawful behavior, conduct, or purpose. ## About OpenTeams ## Company Overview - **One-liner**: OpenTeams helps organizations design, build, and operate fully owned AI systems using modular, open-source infrastructure, enabling them to move from renting to owning their AI. - **Entity Type**: Private (Total funding $7.1M across Pre-Seed, Venture Round, and Debt Financing; acquired Quansight’s AI consulting division in 2025) - **Headquarters**: Austin, Texas, United States - **Founded**: 2019 - **Founders**: Travis Oliphant (CEO & Co-Founder), Dharhas Pothina (CTO & Co-Founder), Matt Harward (CRO & Co-Founder) ## Core Business - **Primary industries**: AI infrastructure, open-source software deployment, enterprise consulting, government/defense AI solutions - **Target customers**: Enterprise, Government (including U.S. Department of Defense), Research institutions, BioTech, Energy, Financial services – primarily B2B, with a mix of mid-market and large enterprise - **Mission**: “AI you own” – building accountable, auditable, and fully owned AI systems rather than renting from hyperscalers. ## Products & Services - **[Nebari®](https://openteams.com/)**: An open-source AI hub that serves as the private infrastructure layer for an organization’s models, data, workflows, and institutional knowledge. Fully controlled and owned by the customer. Listed as a JATIC product supporting DoD AI development. - **[Collab Desktop App](https://openteams.com/)**: A free application for publishing, discovering, consuming, and sharing agentic workflows and AI capabilities. Designed to participate in a distributed applied AI economy built on open standards. - **[Enterprise Deployment](https://openteams.com/)**: Full-service consulting and support – includes building the system, training internal teams, and ongoing support. Emphasizes that ownership stays with the customer, not the vendor. - **[Intelligence Hub Connect](https://openteams.com/)**: Secure integration layer that connects partners, vendors, and data providers on the organization’s terms, controlling data flow while maintaining ownership. ## Market Standing - **Valuation/Market Cap**: Not publicly available - **Key Metrics**: - Annual Revenue: ~$11.0M (LinkedIn estimate, as of mid-2025) - Total Funding: $7.13M (Pre-Seed $100K led by Sputnik ATX; Venture Round $3.0M in Nov 2021; Debt Financing $4.0M in June 2023) - **Notable Investors/Partners**: Sputnik ATX (lead investor in pre-seed); strategic partner Quansight (AI consulting division acquired March 2025); partner with the Open-Source AI Foundation (O-SAIF) for government AI safety. - **Growth Signals**: - Headcount: 73 employees, +106.8% year-over-year growth (+47 people in one year) - Operations in 13 countries (US, Canada, UK, Kenya, India, Brazil, etc.) - Appointed Lt. Gen (Ret) Ross Coffman to Board of Directors (June 2025) - Acquired Quansight’s AI consulting division (March 2025), adding ~25 people from Quansight - Listed as a JATIC product for the U.S. Department of Defense next-gen toolchain ## Competitive Advantages - **Founder pedigree**: Travis Oliphant created NumPy and SciPy – foundational tools for modern AI – and previously founded Anaconda and Quansight. This gives OpenTeams deep credibility and institutional knowledge in open-source AI infrastructure. - **Open-source DNA**: OpenTeams builds, maintains, and contributes to open-source projects; customers get transparency and flexibility without vendor lock-in. - **“Owned Intelligence” model**: Unlike hyperscalers or SaaS vendors, OpenTeams ensures organizations own their AI models, data, and infrastructure – no hidden dependencies, no third-party control. - **Forward-deployed engineering model**: “An army of forward-deployed engineers” that builds and maintains solutions inside the customer’s environment, training internal teams and ensuring continuous ownership. ## Strategic Focus - **Scaling government and defense contracts**: Nebari’s inclusion in JATIC and partnership with O-SAIF signal strong push into U.S. government and defense markets. - **Expanding the distributed AI ecosystem**: Through the Collab Desktop App and Applied AI Society, OpenTeams aims to build a global community of applied AI engineers and a networked private ecosystem. - **Post-acquisition integration**: Absorbing Quansight’s consulting talent to deepen enterprise delivery capacity. - **International growth**: Headcount is diversifying across North America, Europe, Africa, and Asia – with active hiring for roles like Senior Solutions Architect in the EU. ## Why Work Here - **Engineering-led culture**: Founded by a legendary open-source developer; strong technical team with 40% of staff in technical roles (engineering, infrastructure, AI). - **Remote and hybrid flexibility**: While the headquarters is in Austin and “on-site workspace” is noted, many roles are posted as remote (e.g., Senior Solutions Architect – UK, Founding Lead Engineer – remote U.S.). The company operates across 13 countries, suggesting a distributed-first approach. - **Growth trajectory**: 106% headcount growth YoY and recent acquisition signal fast scaling – employees can grow into new roles and leadership. - **Meaningful mission**: Work on AI that organizations truly own, not rent – appealing to engineers who value open-source, privacy, and ethical AI. - **Diverse global team**: Employees from six continents; the company actively hires from a variety of talent pools. ## Sources 1. [openteams.com](https://openteams.com/) – Homepage and product overview 2. [openteams.com/about-us](https://openteams.com/about-us/) – Founders, team, beliefs 3. [linkedin.com](https://www.linkedin.com/company/openteams) – LinkedIn company page (employees, revenue, funding, recent news) 4. [builtin.com](https://builtin.com/company/openteams) – Culture, careers, and perks overview 5. [greenhouse.io](http://job-boards.greenhouse.io/openteams) – Current job openings and office location ## Other roles at OpenTeams - [Technical Project Manager - Project Success](https://feeny.ai/job/technical-project-manager-project-success-openteams-united-states-410zt4tkbf2r) — United States - [Senior Engineering Architect](https://feeny.ai/job/senior-engineering-architect-openteams-united-states-1jc57xmcr7dr) — United States - [Program Manager](https://feeny.ai/job/program-manager-openteams-washington-gbfqz8eskb6g) — Washington, DC / Denver, CO / Colorado Springs, CO - [Full-Stack Platform Engineer](https://feeny.ai/job/full-stack-platform-engineer-openteams-washington-8ca7rhvh2nfs) — Washington, DC / Denver, CO / Colorado Springs, CO - [Distributed Systems ML Infrastructure Engineer](https://feeny.ai/job/distributed-systems-ml-infrastructure-engineer-openteams-washington-pcmbae499j9v) — Washington, DC / Denver, CO / Colorado Springs, CO - [Senior AI/ML Test and Evaluation Engineer](https://feeny.ai/job/senior-ai-ml-test-and-evaluation-engineer-openteams-washington-w2jbjf2t874t) — Washington, DC / Denver, CO / Colorado Springs, CO - [Site Reliability Engineer / DevSecOps Engineer](https://feeny.ai/job/site-reliability-engineer-devsecops-engineer-openteams-washington-h3hq515ek7mz) — Washington, DC / Denver, CO / Colorado Springs, CO - [Technical Delivery Lead](https://feeny.ai/job/technical-delivery-lead-openteams-washington-s16fh9f53hjp) — Washington, DC / Denver, CO - [Senior Machine Learning Engineer - Client Facing](https://feeny.ai/job/senior-machine-learning-engineer-client-facing-openteams-remote-2str5wr18sty) - [Engineer/Senior Engineer - Python Packaging](https://feeny.ai/job/engineer-senior-engineer-python-packaging-openteams-remote-vfhq7ngp1ehg)