--- title: 'AI Cluster Architect at Vultr' canonical: 'https://feeny.ai/job/ai-cluster-architect-vultr-united-states-hkq9rap9vxrs' type: 'job' last_seen: '2026-09-10' --- # AI Cluster Architect at Vultr - **Company:** Vultr - **Location:** United States - **Employment:** full-time - **Work type:** remote - **Posted:** 2026-08-06 - **Last confirmed live:** 2026-09-10 - **Apply:** https://jobs.ashbyhq.com/vultr/4b807abb-d447-477a-8d95-6f0e537c81c5 ## Job description ## Who We Are Vultr is on a mission to make high-performance cloud infrastructure easy to use, affordable, and locally accessible for enterprises and AI innovators around the world. With 33 global cloud data center locations, Vultr is trusted by hundreds of thousands of active customers across 185 countries for its flexible, scalable, global Cloud Compute, Cloud GPU, Bare Metal, and Cloud Storage solutions. In December 2024 Vultr announced an equity financing at a $3.5 billion valuation. Founded by David Aninowsky and self-funded for over a decade, Vultr has grown to become the world’s largest privately-held cloud infrastructure company. Vultr Cares - Excellent Medical Benefits w/ 100% company-paid premiums for employee only plan + 100% company-paid dental & vision premiums - 401(k) plan that matches 100% up to 4% with immediate vesting - Professional Development Reimbursement of $2,500 each year - 11 Holidays + Paid Time Off Accrual + Rollover Plan + take your birthday off - Commitment matters to Vultr! Increased PTO at 3 year & 10 year anniversary + 1 month paid sabbatical every 5 years + Anniversary Bonus each year - $500 first year remote office setup + $400 each following year for new equipment - Internet reimbursement up to $75 per month - Gym membership reimbursement up to $50 per month - Company-paid Wellable subscription Join Vultr Vultr is looking for an AI Cluster Architect who will be responsible for creating and refining large-scale GPU cluster architectures within strict power and infrastructure limits. This role focuses heavily on power-aware design: starting from a fixed power envelope, the architect determines the optimal number of GPUs while accounting for the full stack of services needing to be deployed—compute nodes, storage systems, networking fabric, cooling, and facility constraints. This role requires deep experience navigating heterogeneous environments, multiple generations of hardware, and end user requirements. The architect must understand how different GPU SKUs, NICs, switches, and fabrics interact at scale, including their individual and aggregate power and thermal characteristics. They will evaluate multi-plane, rail-optimized, and tiered fabric designs across technologies like InfiniBand, RoCE, and SpectrumX to ensure the networking architecture supports the intended GPU count without overrunning facility limits or switch radix and/or topology constraints. This role balances customer-specific requirements for compute, storage, and service density, ensuring that the final cluster design maintains acceptable levels of GPU and fabric performance, while maximizing the number of usable GPUs within the total power budget. ## Key Responsibilities - Architect large-scale GPU clusters within fixed site power budgets that optimizes for maximum GPU density while reserving necessary headroom for compute services, storage, and networking. - Model and validate power consumption across the full cluster bill of materials (GPUs, CPUs, NICs, switches, fabric components, storage, and facility limits). - Evaluate tradeoffs across multiple fabric networking architectures (InfiniBand, RoCE, SpectrumX) as well as multi-plane, 2-tier/3-tier, and rail-optimized topologies. - Determine network scale limits based on switch radix, link speed, topology, and blocking requirements. - Gather, interpret, and maintain detailed SKU-level power and thermal specifications for GPUs, NICs, switches, DPUs, storage, and server platforms. - Develop power-aware cluster configuration templates and capacity-planning models that can scale across sites with varying constraints and allow for quick iteration and ideation. - Document architecture, design choices, tradeoff analyses, and operational considerations for deployment and lifecycle management. - Provide guidance on future-proofing, including the ability to incorporate next-gen GPUs, NICs, or fabrics. - Collaborate with vendors on novel fabric architectures that enable large-scale cluster deployments (100k+ GPUs) ## Qualifications - 7+ years designing or building large-scale HPC, AI, or hyperscale GPU clusters. - Expert understanding of GPU and accelerator system design, including node topology, PCIe/NVLink/NVSwitch/ROCm, and NIC-to-GPU affinity considerations. - Strong familiarity with InfiniBand, RoCE, and SpectrumX networking, including multi-tier, multi-plane, Clos/dragonfly variants, and large-radix switch design. - Demonstrated experience modeling power draw and thermal characteristics of servers, GPUs, NICs, switches, optics, and storage systems. - Ability to design networks that maintain full non-blocking performance or intentionally introduce over/under-subscription while understanding impacts on workload performance. - Proven ability to gather and analyze vendor SKU-level specifications and incorporate them into scalable cluster architectures. - Experience balancing customer-driven requirements for compute, storage, and service density in combination with overall GPU count. - Strong documentation, communication, and cross-functional collaboration skills. ## Compensation $165,000 - $185,000 This salary can vary based on location, years of experience, background and skill set. Inclusion & Privacy We are an equal opportunity employer and are committed to creating an inclusive environment for all employees. We welcome applications from individuals of all backgrounds and experiences, and we prohibit discrimination based on race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, veteran status, or any other protected status under applicable laws. Vultr will consider qualified applicants with arrest or conviction records in accordance with applicable laws and will not conduct a background check until after an offer of employment has been extended and accepted. We also take your privacy seriously. We handle personal information responsibly and follow applicable laws, including U.S. privacy rules and India’s Digital Personal Data Protection Act, 2023. Your data is used only for legitimate business purposes and is protected with proper security measures. Where allowed by law, applicants may request details about the data we collect, access or delete their information, withdraw consent for its use, and opt out of nonessential communications. For more details, please see our [Privacy Policy](https://www.vultr.com/legal/privacy/). ## About Vultr ## Company Overview - **One-liner**: Vultr provides a global cloud infrastructure platform that simplifies the deployment of high-performance compute, storage, and bare metal solutions for developers and enterprises. - **Entity Type**: Private (Venture-backed) - **Headquarters**: West Palm Beach, Florida, United States - **Founded**: 2014 - **Founders**: Not publicly disclosed (operates under The Constant Company, LLC) ## Core Business - **Primary Industry**: Cloud Infrastructure (IaaS / PaaS) - **Target Customers**: Developers, SMBs, enterprise IT teams, AI innovators – serving customers in over 185 countries - **Mission**: To empower developers and businesses by simplifying the deployment of infrastructure via an advanced, globally distributed cloud platform ## Products & Services - **Cloud Compute**: Virtual machines (SSD VPS) with hourly billing, available in 32 global data center regions. Supports Linux, Windows, and BSD. - **Optimized Cloud Compute**: Configurations with AMD EPYC CPUs for compute-intensive workloads. - **Bare Metal**: Single-tenant dedicated servers, accelerated by NVIDIA GPUs and high-performance AMD/Intel CPUs. - **Kubernetes Engine (VKE)**: Fully managed Kubernetes service for container orchestration. - **Block & Object Storage**: Scalable, high-performance storage solutions. - **Managed Databases**: Managed MySQL, PostgreSQL, Apache Kafka, and Valkey (Redis-compatible). - **Vultr Marketplace**: One-click apps and partner integrations. - **Vultr Agent**: AI-powered assistant and composable infrastructure tool. - **Additional Services**: 1-Click Apps, ISO mounting, private networking, DDoS protection, and global load balancing. ## Market Standing - **Valuation**: $3.5 billion (as of 2024-12, per financing round from LuminArx and AMD Ventures) [vultr.com](https://www.vultr.com/company/about-us/) - **Key Metric**: Annual revenue of $145.5 million; total funding of $662 million (equity + debt) [linkedin.com](https://www.linkedin.com/company/vultr) - **Notable Investors/Partners**: LuminArx Capital, AMD Ventures, Verizon, Dell Technologies. Debt financing of $329 million from six investors in June 2025. [vultr.com](https://www.vultr.com/company/about-us/) - **Growth Signals**: - 1.5 million+ customers served, 80 million instances deployed - 32 data center regions worldwide with continued expansion (e.g., Chicago, Tel Aviv, Warsaw) - Headcount grew 48.7% YoY to 243 employees (as of mid-2025) [linkedin.com](https://www.linkedin.com/company/vultr) - Active job postings up 364% YoY (65 roles open) [linkedin.com](https://www.linkedin.com/company/vultr) - Won Dell Technologies Global Alliances AI Provider of the Year, Americas award - 2024 Stratus Award for Cloud Computing (cloud disruptor category) ## Competitive Advantages - **Global Footprint**: 32 data center regions across six continents, enabling low-latency local deployments. - **No-Lock-in Pricing**: Hourly billing with no long-term contracts; full root/admin access on all instances. - **Broad Compliance Portfolio**: SOC 2, HIPAA, PCI, CSA STAR, ISO/IEC 27001/27017/27018/20000-1. - **Ecosystem & Partnerships**: Vultr Cloud Alliance with Dell, AMD, Verizon, and others. - **Developer Experience**: Powerful API v2, Terraform integration, one-click apps, and AI assistant. ## Strategic Focus - **AI & GPU Infrastructure**: Expanding Cloud GPU offerings for AI/ML workloads. - **Kubernetes & Containerization**: Continued investment in Vultr Kubernetes Engine (VKE). - **Regional Expansion**: Opening new data centers in underserved markets (e.g., Africa, Latin America). - **Developer Experience**: Enhancing API, marketplace, and composable cloud tooling. - **Enterprise Sales**: Growing channel partnerships and direct enterprise engagement. ## Why Work Here - **Remote-First Culture**: Most roles are fully remote; headquarters in West Palm Beach, FL with local team collaboration. [vultr.com/careers](https://www.vultr.com/company/careers/) - **Benefits**: - 100% employer-paid health insurance - 401(k) with up to 4% matching (immediate vesting) - Generous PTO with rollover, 11 holidays, birthday off, 1-month paid sabbatical every 5 years - 8 weeks paid parental leave - Up to $2,500/year continuing education + free Udemy Business access - Monthly internet and fitness reimbursement - Home office setup allowance - **Culture & Growth**: Emphasis on passion, collaboration, and education. Regular virtual events, quarterly team lunches, spot bonuses, and the “High Flying Vultr” award. - **Engineering Focus**: 31% of workforce in technical roles; opportunities to work on global-scale cloud infrastructure. Active hiring for software engineers, SREs, and data center technicians. ## Sources 1. [vultr.com – About Us](https://www.vultr.com/company/about-us/) 2. [vultr.com – Careers](https://www.vultr.com/company/careers/) 3. [vultr.com – Homepage](https://www.vultr.com/) 4. [linkedin.com – Vultr Company Profile](https://www.linkedin.com/company/vultr) 5. [jobs.ashbyhq.com – Vultr Open Positions](https://jobs.ashbyhq.com/vultr) ## Other roles at Vultr - [Procurement Manager](https://feeny.ai/job/procurement-manager-vultr-united-states-6g9k6twnjx1t) — United States - [Senior Software Engineer, Cloud Networking](https://feeny.ai/job/senior-software-engineer-cloud-networking-vultr-united-states-f4jrpwhrkbn2) — United States - [Senior Software Engineer - Tooling & Automation](https://feeny.ai/job/senior-software-engineer-tooling-automation-vultr-chennai-zmfr8qzvy9vm) — Chennai, India - [Data Center Acquisition Manager](https://feeny.ai/job/data-center-acquisition-manager-vultr-united-states-sfe6mp6fa3xw) — United States - [Provisioning Engineer](https://feeny.ai/job/provisioning-engineer-vultr-united-states-f8vjv8qb375x) — United States - [Principal Technical Product Manager, Strategic Accounts](https://feeny.ai/job/principal-technical-product-manager-strategic-accounts-vultr-united-states-n1v695d4zy15) — United States - [Site Selection Project Manager](https://feeny.ai/job/site-selection-project-manager-vultr-united-states-322xzm6szwyz) — United States - [Accounts Payable Analyst](https://feeny.ai/job/accounts-payable-analyst-vultr-united-states-z4d3qfvfcydh) — United States - [Infrastructure Production Engineer](https://feeny.ai/job/infrastructure-production-engineer-vultr-united-states-xajtmkq5x4x9) — United States - [Software Engineer, Storage](https://feeny.ai/job/software-engineer-storage-vultr-united-states-0p0be4mrx7v6) — United States