--- title: 'Staff Site Reliability Engineer at Sumo Logic' canonical: 'https://feeny.ai/job/staff-site-reliability-engineer-sumo-logic-noida-uttar-pradesh-tp6bn0nmwqh2' type: 'job' last_seen: '2026-09-24' --- # Staff Site Reliability Engineer at Sumo Logic - **Company:** Sumo Logic - **Location:** Noida Uttar Pradesh, India - **Posted:** 2026-02-10 - **Last confirmed live:** 2026-09-24 - **Apply:** https://job-boards.greenhouse.io/sumologic/jobs/7580057 ## Job description Title: Staff Site Reliability Engineer, Product Area Focus Location: Noida/ Bangalore (Hybrid) Summary of role Sumo Logic's microservices architecture, hosted on AWS, ingests petabytes of data daily across many geographic regions in support of our planet-scale observability and security products, serving hundreds of millions of queries a day against thousands of petabytes of data. At that scale, every inefficiency — in code, architecture, or infrastructure — compounds into real cost. This role sits within the Product SRE organization, working alongside your global SRE team on your product area's reliability roadmap — with a mandate that leads with code as much as operations. You'll find where Sumo's systems are spending more compute, storage, or engineering time than necessary, and fix it in the code and architecture, not just the infra config, while also carrying the fuller SRE mandate — reliability, security posture, and improving the day-to-day experience of the engineers within your product area. You'll be part of a team that blends SRE and backend software engineering skillsets, partnering closely with product engineering teams across your product area. This is an engineering role — the work is about shipping code, architecture, and system-level changes that improve unit economics, not managing cost dashboards, tagging, or reserved-instance/savings-plan purchasing. ## What you’ll do - Continuously discover cost and efficiency opportunities through production profiling, telemetry, cost data, capacity trends, and system-level analysis — across algorithmic inefficiencies, resource-heavy code paths, and architectural decisions — and turn ambiguous problems into prioritized engineering initiatives. - Apply performance and capacity engineering techniques to understand CPU, memory, storage, network, and I/O behavior under real production workloads, and optimize the resulting resource footprint. - Write production-grade code to implement the optimizations you identify — JVM/GC tuning, algorithmic and resource-efficiency improvements, re-architecting inefficient services — in systems that process petabytes of data daily. - Define and track engineering efficiency metrics such as cost per GB ingested, cost per query, cost per event, resource utilization, or cost per customer workload, and translate the work into measurable impact for engineering and business stakeholders. - Lead complex, cross-team engineering initiatives from problem discovery through design, implementation, rollout, and measurement — influencing teams where you don't have direct ownership — and help establish engineering patterns and practices that make cost and efficiency a continuous part of the development lifecycle. - Partner with engineering teams in your product area to prioritize changes, and with developer infrastructure and Global SRE to align with the broader reliability roadmap. - Participate in the SRE responsibilities for different product areas — SLOs, on-call, incident response, and blameless RCA — using those experiences to identify systemic reliability, performance, and efficiency improvements. ## What you’ll have - B.Tech, M.Tech, or equivalent degree in Computer Science or a related discipline. - 8+ years of industry experience with a demonstrated track record of ownership. - Strong CS fundamentals — comfortable with algorithmic complexity, data-structure performance characteristics, and system design at scale. - Ability to author production-ready code in at least one OO/systems language (Java, Scala, Go, C++, or similar) — depth of engineering ability matters more than which language. - Experience with distributed systems and microservice architectures in production. - Demonstrated track record of independently identifying ambiguous performance, scalability, or cost problems and driving engineering changes that produced measurable improvements in production. - Strong ability to reason quantitatively about system behavior, capacity, performance, and cost, and use production data to validate hypotheses and measure outcomes. - Working fluency with cloud infrastructure (AWS compute, storage, networking) — enough to reason about cost and architectural tradeoffs. - Comfort moving across the stack, from application code to the infrastructure it runs on, to find root causes of inefficiency. ## Nice to have - JVM tuning and GC optimization experience at scale. - Exposure to cost-attribution/FinOps practices or tooling. - Experience with Kubernetes, Terraform, or modern CI/CD tooling. - Prior SRE experience — on-call, SLOs, incident response. - Experience with streaming technologies (Kafka, Kafka Streams) or observability/security platforms. ## Why this role - Visibility — Cost/efficiency work maps directly to metrics the business already tracks. - Scope you define — you identify where the opportunities are rather than executing a fixed backlog. - Real scale — systems ingesting petabytes of data daily and serving hundreds of millions of queries. ## About Us Sumo Logic, Inc. helps make the digital world secure, fast, and reliable by unifying critical security and operational data through its Intelligent Operations Platform. Built to address the increasing complexity of modern cybersecurity and cloud operations challenges, we empower digital teams to move from reaction to readiness—combining agentic AI-powered SIEM and log analytics into a single platform to detect, investigate, and resolve modern challenges. Customers around the world rely on Sumo Logic for trusted insights to protect against security threats, ensure reliability, and gain powerful insights into their digital environments. For more information, visit[www.sumologic.com.](http://www.sumologic.com/) [Sumo Logic Privacy Policy](https://www.sumologic.com/privacy-statement/). Employees will be responsible for complying with applicable federal privacy laws and regulations, as well as organizational policies related to data protection. ## About Sumo Logic ## Company Overview - **One-liner**: Sumo Logic is a cloud-native AI-powered log analytics and security platform that helps organizations ensure their digital experiences are secure, fast, and reliable. - **Entity Type**: Private (acquired by Francisco Partners in 2023) - **Headquarters**: Redwood City, California, USA - **Founded**: 2010 - **Founders**: Christian Beedgen, Kumar Saurabh ## Core Business - **Primary industry**: Cybersecurity and Observability (Cloud Security, Log Management, SIEM) - **Target customers**: B2B, Enterprise, Mid-market, SMB - **Mission**: “Making the digital world secure, fast, and reliable” [sumologic.com](https://www.sumologic.com/company) ## Products & Services - **Sumo Logic Intelligent Operations Platform**: Cloud-native platform unifying log management, monitoring, and security for Dev, Sec, and Ops teams. - **Cloud SIEM**: Security information and event management solution designed for cloud environments, with automated threat detection and investigation. - **Dojo AI**: Multi-agent AI platform for intelligent security operations and incident response, reducing mean time to resolution (MTTR). - **Log Analytics**: Log management and analytics for monitoring, troubleshooting, and root cause analysis. - **Flex Licensing**: Consumption-based pricing model allowing customers to pay only for the data they ingest. ## Market Standing - **Valuation/Market Cap**: Not publicly available (acquired by Francisco Partners in 2023 for approximately $1.7 billion; current valuation not disclosed) - **Key Metric**: Annual Revenue – Not publicly available - **Notable Investors/Partners**: Francisco Partners (acquirer); key compliance certifications include FedRAMP Moderate, ISO 27001, SOC 2 Type II, HIPAA, PCI DSS 3.2, GDPR, CCPA. - **Growth Signals**: Launch of Dojo AI agentic platform; FedRAMP Moderate authorization; global hiring across US, India, Costa Rica, and Germany; reported 376% three-year ROI for customers. ## Competitive Advantages - AI-native platform with specialized agents (Dojo AI) that automate triage, correlation, and response. - Unified observability and security on a single cloud-native platform, eliminating data silos. - Flex Licensing model enables cost-effective ingestion of all log data without budget waste. - Strong compliance posture (FedRAMP, ISO, SOC 2, HIPAA, PCI) trusted by government and regulated industries. - High customer ROI and reduction in MTTR (60% decrease reported on website). ## Strategic Focus - Intelligent SecOps for the AI era – leveraging AI agents to automate detection, investigation, and remediation. - Continuous expansion of cloud-native capabilities and integrations to support modern digital enterprises. - Scaling global presence and remote-first hiring to attract top talent. ## Why Work Here - **Culture**: Values emphasize “Win Together”, “Stay Hungry”, “Be Customer Champions”, and “Own It” – with a focus on collaboration, resilience, customer obsession, and accountability. [sumologic.com](https://www.sumologic.com/company/careers) - **Work Environment**: Many roles are remote (USA, India) with offices in San Francisco, Boston, Denver, Orlando, Tampa, Noida, Bangalore, San Jose (Costa Rica), and Germany. [greenhouse.io](http://job-boards.greenhouse.io/sumologic) - **Engineering & Growth**: Opportunity to work on cutting-edge AI, machine learning, site reliability, and threat research. Employees describe a supportive environment with career development and personal growth. - **Perks**: Not explicitly listed, but global team, flexible work options, and a mission-driven culture. ## Sources 1. [sumologic.com](https://www.sumologic.com/company) – Company overview, mission, and purpose 2. [sumologic.com](https://www.sumologic.com/) – Product details, Dojo AI, Cloud SIEM, Flex Licensing, and compliance 3. [sumologic.com](https://www.sumologic.com/company/careers) – Values, culture, and employee testimonials 4. [greenhouse.io](http://job-boards.greenhouse.io/sumologic) – Current job openings and locations 5. [sumologic.com](https://www.sumologic.com/company/leadership) – Leadership team and executive biographies ## Other roles at Sumo Logic - [Senior Partner Sales Manager](https://feeny.ai/job/senior-partner-sales-manager-sumo-logic-london-england-49mnrway431f) — London England, United Kingdom - [Senior Solutions Engineer](https://feeny.ai/job/senior-solutions-engineer-sumo-logic-melbourne-zvmexmwwx2n3) — Melbourne, Australia - [Senior Solutions Engineer](https://feeny.ai/job/senior-solutions-engineer-sumo-logic-sydney-new-south-wales-yzg4yq74dpjd) — Sydney New South Wales, Australia - [Senior Sales Engineer](https://feeny.ai/job/senior-sales-engineer-sumo-logic-london-england-wxw28cwbezcf) — London England, United Kingdom - [Staff Software Engineer - Testing & Automation](https://feeny.ai/job/staff-software-engineer-testing-automation-sumo-logic-noida-uttar-pradesh-0esnndpsynj7) — Noida Uttar Pradesh, India - [Principal Product Manager](https://feeny.ai/job/principal-product-manager-sumo-logic-bengaluru-prefs0303b69) — Bengaluru, India - [Principal Product Manager](https://feeny.ai/job/principal-product-manager-sumo-logic-noida-uttar-pradesh-129b9ebbkzbt) — Noida Uttar Pradesh, India - [Staff Site Reliability Engineer](https://feeny.ai/job/staff-site-reliability-engineer-sumo-logic-bengaluru-fkbjq7q3jpn8) — Bengaluru, India - [Staff Site Reliability Engineer](https://feeny.ai/job/staff-site-reliability-engineer-lightspeed-commerce-inc-auckland-nxwj1vf8t1xs) — Auckland, New Zealand - [Staff Site Reliability Engineer](https://feeny.ai/job/staff-site-reliability-engineer-replit-united-states-dqkf7f3vdv0d) — United States