Andromeda Cluster

Member of the Technical Staff - Systems at Andromeda Cluster (North, United States / San Francisco, CA)

Andromeda Cluster· North, United States / San Francisco, CA·

Role details

Work type
Remote
Employment
Full-Time

Job description

Member of the Technical Staff, Systems Location: North America Remote / San Francisco, CA · Full-Time

About Andromeda

Andromeda is a market and infrastructure platform to buy, sell, and operate compute.

We believe demand for compute will grow exponentially. So fast that a handful of vertically integrated providers won't be able to scale across operations, capital, supply chains, and politics to serve it. The result is a massive wave of fragmentation, with AI factories of every shape and size coming to market to fill this demand. Our job is to enable all of that fragmented compute to flow through one platform, delivering reliable capacity to model builders, research labs, and inference providers when they need it. We believe every spare electron should be made productive for AI and we're building the platform that makes that possible.

We sit at the center of three forces:

  • Companies that need reliable, high-performance compute fast
  • A fragmented global supply of GPUs across hyperscalers, neoclouds, and independent data centers
  • Capital, risk, and operational complexity that most teams are not equipped to manage

When we succeed, trillions of dollars of compute will flow through Andromeda. Builders get capacity when they need it. Providers get a reliable way to monetize, operate, and finance infrastructure at scale. Capital gets an easy way to deploy, hedge, and underwrite.

In five years, Andromeda won't just participate in the AI infrastructure market. We will shape it.

The Role

We are looking for strong engineers with experience and interest in designing and building high performance systems across, but not limited to: storage, networking, virtualization, container runtimes.

Requirements

  • Impressive technical work you can go deep on, with impact in the world. That can take three years or twenty.
  • Experience building high-performance distributed systems at scale
  • Strong low-level Linux foundations: kernel, drivers, filesystems, containers, PCIe
  • Experience with performance engineering
  • Production experience in Rust, C, or Go
  • Ability to participate in on-call rotations and respond to production incidents

Experience with QEMU/KVM, eBPF, RDMA, or BIOS/UEFI internals is nice to have, but not required.

Why You’ll Love It Here

  • High-growth environment: Get in early at a company at the center of the AI infrastructure boom
  • Competitive compensation: + meaningful equity
  • Comprehensive benefits: for you and your dependents, including healthcare, dental, and vision coverage, 401(k), and unlimited PTO

Andromeda Cluster is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. We do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.

Why work at Andromeda Cluster

  • Culture & environment: Small, high-impact team (14 people) working on one of the most critical bottlenecks in AI. Flat structure with direct access to leadership.
  • Remote/hybrid policy: Global remote with a San Francisco office. Many roles listed as "Global Remote / San Francisco, CA".
  • Engineering culture: Deep technical challenges – GPU orchestration, distributed systems, observability, kernel-level performance tuning. Tech stack includes Kubernetes, Slurm, PyTorch, CUDA, Weka, Prometheus, Grafana, and more.
  • Roles available: Compute Trader, Site Reliability Engineer, Software Engineer, Head of Partnerships, Solutions Engineer, Strategic Compute Finance Lead, and more – all focused on AI infrastructure.

Application questions