--- title: 'Senior/Staff Backend Engineer - Distributed System at Zettabyte' canonical: 'https://feeny.ai/job/senior-staff-backend-engineer-distributed-system-zettabyte-united-states-x741rfe1g0j5' type: 'job' last_seen: '2026-09-05' --- # Senior/Staff Backend Engineer - Distributed System at Zettabyte - **Company:** Zettabyte - **Location:** United States - **Employment:** full-time - **Work type:** hybrid - **Posted:** 2025-10-27 - **Last confirmed live:** 2026-09-05 - **Apply:** https://jobs.ashbyhq.com/zettabyte-space/c2737b0b-79aa-4d8a-8a52-dfd7a40312e4 ## Job description ## ABOUT US At Zettabyte, we’re on a mission to make AI compute ubiquitous, seamless, and limitless. We’re building a cloud where AI just works—anywhere, anytime. “AI Power. Everywhere.” Be part of the team designing the infrastructure for the AI-first world. ## WHY THIS ROLE EXISTS We need a Backend Engineer to build the systems that orchestrate GPU clusters for AI workloads. You'll create APIs that handle GPU allocation, memory management, compute scheduling, and multi-tenant isolation—challenges unique to AI infrastructure that go far beyond typical backend engineering. As part of our backend team, you'll solve problems like: How do we efficiently share expensive GPU resources across users? How do we handle GPU memory constraints for large AI models? How do we ensure quality of service when workloads compete for compute? This is an opportunity to build infrastructure where every API call could allocate thousands of dollars worth of compute per hour, where your optimizations directly impact whether AI startups can afford to train their models. ## WHAT YOU’LL DO - Design APIs that abstract complex GPU operations into simple developer experiences - Build scheduling algorithms that maximize GPU utilization while ensuring SLA compliance - Develop resource management systems for GPU lifecycle—provisioning, allocation, scheduling, and release - Create usage tracking and billing systems for GPU-hours, memory usage, and compute utilization - Implement monitoring for GPU-specific metrics, health checks, and automatic failure recovery - Build multi-tenancy systems with resource isolation, quota management, and fair scheduling - Optimize cold starts for model serving and implement efficient model loading strategies - Collaborate with frontend engineers to expose complex infrastructure through intuitive interfaces - Leverage AI-assisted coding tools (GitHub Copilot, Claude Code, Cursor IDE, etc.) to boost productivity and code quality. ## YOU’LL THRIVE HERE IF YOU - 5+ years backend engineering experience with distributed systems - Strong proficiency in Go, Python, or similar backend languages - Experience with resource scheduling, orchestration, and API design (REST, GraphQL, gRPC) - Understanding of hardware constraints and system optimization - Linux systems knowledge and containerization experience (Docker, Kubernetes) - Comfortable working with expensive resources where efficiency directly impacts costs - Excited about solving novel problems in AI infrastructure (not just another CRUD app) - Startup mindset—comfortable with ambiguity and rapid iteration ## BONUS QUALIFICATIONS - GPU or HPC cluster management experience - Understanding of ML/AI workload patterns and requirements - Experience with high-value resource allocation systems - Background in performance optimization for compute-intensive workloads - Familiarity with GPU virtualization and sharing technologies - Experience building billing or metering systems ## DETAILS - We provide Competitive salary and equity based on your experience and skillset; - This is a Hybrid role - 3 days in office, 2 days WFH; Must locate in Palo Alto - Applicants must be authorized to work in the United States without need for visa sponsorship. ## About Zettabyte ## Company Overview - **One-liner**: Zettabyte builds sovereign-grade AI data centers and full-stack GPU software for governments and enterprises, enabling them to own and control their AI infrastructure without vendor lock-in. - **Entity Type**: Private (Venture-backed) - **Headquarters**: Palo Alto, California, United States (also Taipei, Taiwan) - **Founded**: 2024 - **Founders**: Not publicly disclosed ## Core Business - **Primary Industry**: AI Infrastructure / Data Centers / GPU Cloud - **Target Customers**: Sovereign nations (governments), large enterprises, and organizations requiring high-performance AI compute with full data sovereignty. - **Mission**: Build future-proof, supply chain neutral AI data centers that maximize economic value from hardware investments; committed to aiding the United Nations Sustainable Development Goals (UNGM supplier). ## Products & Services - **Sovereign AI Data Centers**: Full ownership model via Build-Operate-Train & Transfer, allowing nations to run AI on their own infrastructure, data, and networks. - **Hybrid AI Data Centers**: Unified environments supporting both CPU-based enterprise workloads and high-performance GPU-driven AI workloads. - **Enterprise AI Data Centers**: Dedicated compute, storage, and networking for secure AI workloads. - **TITAN AI Data Center**: Next-generation AIDC initiative for scalable GPU-based computing worldwide. - **Distributed Cell Tower Computing**: Edge computing applications placed closer to end users. - **zPLATFORM™**: Software layer for full visibility, control, and operation of GPU infrastructure. - **zCLOUD™**: On-demand access to high-performance GPUs (hourly scaling, no upfront commitment). - **GPU Financing**: Specialized financing through banking partners for data center and GPU cluster buildouts. - **Central Command Center**: Monitoring and management of AI infrastructure. - **Sovereign-Grade Networking**: Networking designed for security and neutrality. ## Market Standing - **Valuation/Market Cap**: Not disclosed - **Key Metric**: Total Funding – Raised multiple venture rounds (latest in 2026); total amount not specified. Investors include Headline Asia, Lam Capital, and strategic backing from Wistron, Foxconn, and Pegatron. - **Notable Investors/Partners**: Headline Asia (formerly Infinity Ventures), Lam Capital, Wistron, Foxconn, Pegatron. United Nations Global Marketplace supplier. - **Growth Signals**: 67,000+ GPUs currently under management; ~1.5 GW upcoming deployment scale; 5+ sovereigns working with Zettabyte; 23+ trusted partners; LinkedIn followers grew +4528.6% year-over-year. ## Competitive Advantages - **Supply Chain Neutrality**: Unmatched procurement speed and avoidance of vendor lock-in for sovereign customers. - **Manufacturing Backing**: Supported by Wistron, Foxconn, and Pegatron, which together build over 75% of the world’s AI GPU servers. - **Integrated Software Stack**: Proprietary zPLATFORM and zCLOUD provide measurable, controllable, and scalable AI infrastructure operations. - **Sovereign-Ready**: Full ownership and control over hardware and software, addressing national security and regulatory compliance. ## Strategic Focus - Scaling AI data center deployments globally (1.5 GW pipeline). - Deepening partnerships with sovereign governments to enable national AI independence. - Expanding software capabilities (zPLATFORM, zCLOUD) to drive efficiency and lower TCO for large-scale AI workloads. - Extending edge computing through distributed cell tower solutions. ## Why Work Here - **Culture**: “Growth isn’t just a goal, it is built into how we work.” Emphasis on meaningful challenges and career shaping. - **Work Model**: Hybrid roles in both Palo Alto, USA and Taipei, Taiwan. - **Roles**: Engineering-heavy (GPU Systems, Cloud Native, Distributed Systems, Frontend, Backend, AI Infrastructure, Research Scientists) plus product and people operations. - **Notable Perks**: Opportunity to work on cutting-edge AI infrastructure at scale, with direct impact on sovereign AI capabilities and global GPU deployments. ## Sources 1. [zettabyte.space – Homepage](https://www.zettabyte.space/) 2. [zettabyte.space – About page](https://www.zettabyte.space/about) 3. [zettabyte.space – Careers page](https://www.zettabyte.space/careers) 4. [jobs.ashbyhq.com – Zettabyte openings](https://jobs.ashbyhq.com/zettabyte-space) 5. [LinkedIn – Zettabyte Inc](https://www.linkedin.com/company/zettabyte-inc) ## Other roles at Zettabyte - [Software Engineer (New Grad)](https://feeny.ai/job/software-engineer-new-grad-zettabyte-united-states-s5kn7jd7zfwr) — United States - [Growth Operations Intern](https://feeny.ai/job/growth-operations-intern-zettabyte-united-states-b3e0gs5eb4tw) — United States - [Head of People](https://feeny.ai/job/head-of-people-zettabyte-united-states-ddvrar2axqa7) — United States - [Software Engineering Intern (Summer 2026)](https://feeny.ai/job/software-engineering-intern-summer-2026-zettabyte-united-states-bcxy76qy69tx) — United States - [Senior/Staff Security Engineer](https://feeny.ai/job/senior-staff-security-engineer-zettabyte-united-states-10xdgqndwwtn) — United States - [Cloud Native Engineer](https://feeny.ai/job/cloud-native-engineer-zettabyte-united-states-vt8p8bjhy9zm) — United States - [Outbound Product Manager](https://feeny.ai/job/outbound-product-manager-zettabyte-united-states-s69nmkmkqx3d) — United States - [Frontend Engineer - UI/UX Focus](https://feeny.ai/job/frontend-engineer-ui-ux-focus-zettabyte-united-states-pwyw0yjcdhvf) — United States