--- title: 'Sr Site Reliability Engineer at SigNoz' canonical: 'https://feeny.ai/job/sr-site-reliability-engineer-signoz-india-q51zgjxvm8mq' type: 'job' last_seen: '2026-09-05' --- # Sr Site Reliability Engineer at SigNoz - **Company:** SigNoz - **Location:** India - **Compensation:** INR 500k–INR 1M - **Employment:** full-time - **Work type:** remote - **Posted:** 2026-06-23 - **Last confirmed live:** 2026-09-05 - **Apply:** https://jobs.ashbyhq.com/SigNoz/83d2c104-dde3-454d-bcc5-7cffdb816d55/application **Skills:** Kubernetes, Distributed systems, Performance debugging, Capacity planning, Infrastructure-as-code, CI/CD, ClickHouse, Golang, OpenTelemetry, Kafka, Observability tools > Own the reliability, scalability, and operability of the SigNoz cloud platform, managing petabyte-scale systems, Kubernetes infrastructure, and data pipelines. This hands-on role focuses on incident response, capacity planning, and infrastructure automation in a remote-first environment. ## Job description ## About SigNoz SigNoz is an open-source observability platform that helps modern engineering teams monitor, debug, and optimize their applications with deep visibility into metrics, traces, and logs — all in one place. We're built natively on OpenTelemetry and offer both self-hosted and cloud options, so teams can run observability the way they want, without vendor lock-in. We are growing fast and building core developer infra products. And we are not fooling around: - 27,000+ GitHub stars - 800+ customers - 7,000+ members in our Slack community Role: Sr Site Reliability Engineer (SRE) We're looking for an SRE to own the reliability, scalability, and operability of the SigNoz cloud platform. You'll keep a petabyte-scale observability system fast and dependable — making sure the people who trust us to watch their systems can always trust ours. The platform team handles infra, scalability of SaaS, ingest pipelines, staging environments, automation, and the operational backbone of the product. This is a deeply hands-on role for someone who understands what actually breaks in production at scale — and enjoys fixing it for good. ## What we're looking for - Kubernetes at scale — not just "I've deployed to k8s," but real fluency with the nuances and gotchas: resource tuning, autoscaling behavior, networking, stateful workloads, upgrades, and the failure modes that only show up under load - Working knowledge of ClickHouse — operating it, tuning queries, and understanding its behavior at scale — is a strong plus - Knowledge of Golang is a plus (most of our stack and tooling is in Go) - Familiarity with OpenTelemetry and running large-scale data ingest pipelines is a plus ## What you'll work on You'll work with a high-caliber team across areas like: - Reliability of the SigNoz cloud platform: SLOs/SLIs, error budgets, incident response, and on-call practices that don't burn people out - Scaling the ingest path — making it robust to bursts while maintaining data freshness - SaaS auto-scalability and capacity planning across a petabyte-scale system - Operating and tuning ClickHouse and the data layer for performance and cost - Kubernetes infrastructure: cluster operations, upgrades, multi-tenancy, and the automation that keeps it boring - Observability of SigNoz itself — we dogfood our own product, so you'll help make it world-class - Infrastructure-as-code, CI/CD, and the tooling that lets a small team operate big systems What will make you successful - 5–8 years in SRE, infrastructure, or platform/backend roles operating production systems at scale - Deep, practical Kubernetes experience — you know where the bodies are buried - Strong grasp of distributed systems failure modes, performance debugging, and capacity planning - Comfortable in code (Go preferred) — you automate and fix things, not just configure them - Loves open source — ideally with prior contributions to OSS projects (any size) - Comfortable in a high-ownership, fast-moving, remote-first environment - Strong communication — can write clear runbooks and tech docs and explain trade-offs Nice-to-haves - Past experience on platform/infra/SRE teams of Series B+ startups - Hands-on experience operating ClickHouse, Kafka, or similar high-throughput data systems - Experience in observability (monitoring / logging / tracing) and with OpenTelemetry ## Why you'll love working at SigNoz - Work on a globally used open-source project that engineers actually love - Huge scope and ownership — your work directly shapes how teams adopt SigNoz - Collaborate with a high-caliber team who just can't stop shipping - Remote-first, async-friendly culture - Opportunity to help define the future of open-source observability ## About SigNoz ## Company Overview - **One-liner**: SigNoz is an open-source observability platform that provides logs, metrics, traces, and alerts under a single pane of glass, serving as an alternative to Datadog and New Relic. - **Entity Type**: Private (Seed Stage) - **Headquarters**: San Francisco, California, United States - **Founded**: 2021 - **Founders**: Pranay Prateek (CEO) and Ankit Nayan (CTO) ## Core Business - **Primary industry/industries**: Observability, Application Performance Monitoring (APM), Software Development, DevOps - **Target customers**: B2B; high-growth engineering teams, platform engineers, and DevOps teams at startups to public enterprises - **Mission or purpose statement**: To help developers find signals (actionable insights) from the noise of their observability systems, building more reliable products through a single, open-source platform. ## Products & Services - **SigNox Cloud**: A fully managed, SOC 2 compliant observability platform hosted by SigNoz, live in minutes. - **SigNoz Self-Hosted**: Deployable via Helm chart in your own VPC or air-gapped environment for complete data privacy and compliance (HIPAA, GDPR). - **SigNoz BYOC (Bring Your Own Cloud)**: SigNoz manages the stack inside the customer’s AWS/GCP/Azure account; data never leaves the customer’s VPC. - **Agent Native Observability**: Connect SigNoz to coding agents (e.g., Claude Code, Cursor) to debug production issues directly from the dev environment. - **Noz (AI Teammate)**: An AI-powered feature for natural language observability, allowing users to talk to their observability stack in English. - **Core Modules**: Distributed Tracing, Log Management, Metrics & Dashboards, Alerts, Infrastructure Monitoring, and LLM/AI Observability (token-level tracing, per-model cost attribution). ## Market Standing - **Valuation/Market Cap**: Not disclosed - **Key Metric**: Total funding of **$6.5M**; Annual Revenue of **$3.0M** (estimated) - **Notable Investors/Partners**: **SignalFire** (lead investor in seed round), Y Combinator (backed). Notable partners include the Cloud Native Computing Foundation (CNCF). - **Growth Signals**: - **Headcount**: 49 employees (+52.9% YoY, +18 people) - **GitHub Stars**: 24,000+ - **Community Members**: 4,500+ - **LinkedIn Followers**: 8,174 (+42.4% YoY) - **Data Scale**: Proven track record of handling 10TB+ data ingestion per day. - **Global Reach**: Used by teams across 5 continents; operates in 5 countries (US, India, Brazil, Chile, Afghanistan). ## Competitive Advantages - **Open Source + OpenTelemetry Native**: Built from the ground up for OpenTelemetry, preventing vendor lock-in. Instrumentation remains a company asset. - **Single Pane of Glass**: Unifies logs, metrics, and traces in one highly optimized backend (ClickHouse), enabling correlated debugging. - **Cost Predictability**: Usage-based, transparent pricing with no per-host penalties for auto-scaling and no special pricing for custom metrics. - **Enterprise-Grade**: SOC 2 Type II, fine-grained RBAC, support for high-cardinality data, and 24x7 expert support. - **AI-Native Observability**: Correlates AI/LLM workloads with the entire infrastructure (databases, microservices, applications), unlike siloed LLM-only tools. ## Strategic Focus - **AI & LLM Workload Observability**: Deepening capabilities for monitoring AI applications, including token-level tracing and cost attribution. - **Agent Native Observability**: Integrating with development agents to reduce MTTR (Mean Time to Resolution) by allowing debugging directly from the coding environment. - **Scale and Performance**: Continued optimization of the ClickHouse-based ingestion engine to handle high-cardinality and massive data volumes (10TB+/day). - **Enterprise Adoption**: Expanding enterprise features (RBAC, compliance, self-hosted/BYOC) and support plans to win larger customers. - **Community Growth**: Nurturing the open-source community (24k+ GitHub stars, 4.5k+ community members) to drive adoption and contributions. ## Why Work Here - **Culture**: A developer-first, open-source company obsessed with the "signal vs noise" problem. The team is highly technical and distributed globally. - **Remote/Hybrid Policy**: Remote-friendly, with a significant presence in India (38 employees) and the US (6 employees). The company operates across 5 countries. - **Engineering Culture**: Deep expertise in OpenTelemetry, distributed tracing, log pipeline design, and cost governance. Engineers work on a high-scale, high-impact product used by thousands. - **Growth Trajectory**: Rapid headcount growth (+52.9% YoY) and increasing job postings (+250% YoY), indicating a scaling phase with ample opportunity for impact. - **Notable Perks**: Work with a modern stack (ClickHouse, Go, OpenTelemetry), contribute to a leading open-source project, and solve hard problems at scale. - **Current Open Roles**: Forward Deployed Engineer, Sr Backend Engineer - Platform, Growth Marketing (India/EU/US), Founding DevRel, Founding Designer, and others (7 active postings as of mid-2026). ## Sources 1. [signoz.io](https://signoz.io/) 2. [signoz.io/about-us](https://signoz.io/about-us/) 3. [signoz.io/why-signoz](https://signoz.io/why-signoz/) 4. [linkedin.com/company/signozio](https://www.linkedin.com/company/signozio) 5. [jobs.ashbyhq.com/SigNoz](https://jobs.ashbyhq.com/SigNoz) ## Other roles at SigNoz - [Head of Finance](https://feeny.ai/job/head-of-finance-signoz-india-vbh01msmx4x3) — India - [Founding SDR Lead (US Market)](https://feeny.ai/job/founding-sdr-lead-us-market-signoz-india-bc1vxdjj9yjs) — India - [Senior Software Engineer - Frontend](https://feeny.ai/job/senior-software-engineer-frontend-signoz-india-c9g6vrz577bj) — India - [Head of Demand Gen](https://feeny.ai/job/head-of-demand-gen-signoz-india-e3s0qh830kw9) — India - [AI GTM Engineer](https://feeny.ai/job/ai-gtm-engineer-signoz-india-pw0a3qk4bben) — India - [Sr Product Designer](https://feeny.ai/job/sr-product-designer-signoz-india-7gkw81919mjp) — India - [Sr Backend Engineer - AI](https://feeny.ai/job/sr-backend-engineer-ai-signoz-india-dywh2svf58hx) — India - [Exceptional Engineer](https://feeny.ai/job/exceptional-engineer-signoz-india-4gbe7747dsxw) — India - [Sr Product Manager - IC](https://feeny.ai/job/sr-product-manager-ic-signoz-india-1gn9gkgkjkhs) — India - [Staff Backend Engineer - Core](https://feeny.ai/job/staff-backend-engineer-core-signoz-india-w2z8zrs7wy6j) — India