--- title: 'Senior Site Reliability and Infrastructure Engineer at Treeswift Inc' canonical: 'https://feeny.ai/job/senior-site-reliability-and-infrastructure-engineer-treeswift-inc-new-york-wnxbd3rgd04d' type: 'job' last_seen: '2026-09-06' --- # Senior Site Reliability and Infrastructure Engineer at Treeswift Inc - **Company:** Treeswift Inc - **Location:** New York, NY - **Employment:** full-time - **Work type:** hybrid - **Posted:** 2026-08-04 - **Last confirmed live:** 2026-09-06 - **Apply:** https://jobs.ashbyhq.com/treeswift/3718c171-cf51-49c5-86a3-c5d8b0d31f4c ## Job description In the face of rising threats, increasing pressure on affordability, and unprecedented demand for power, Treeswift empowers energy companies to modernize their field work to meet the growth and challenges ahead. We build physical AI for the field worker: whether on foot or in a vehicle, our technology is an ironman suit for engineers, linemen, and vegetation crews: same worker, same boots on the ground, now operating at 10x productivity. Our platform is powered by cutting edge hardware, sensors (LiDAR, camera, etc…), AI and software designed to revolutionize work in the toughest environments. Since our first pilot with a utility in June 2024, we've grown fast, now working with three of the five largest utilities in the US. To date, our technology has enabled our customers to reduce wildfire risk, regulatory and outage risk from vegetation, avoid delays and cost overruns in new construction, and accelerate recovery from severe storms. To tackle this challenge, we are bringing together a team of mission-driven experts with deep industry experience in robotics (Penn, Caltech, CMU) and enterprise software development (Palantir, Stripe, Oracle, MongoDB). We have raised funding from leading investors including Penny Pritzker’s Inspired Capital. We are headquartered in midtown Manhattan, with additional offices in San Francisco and Philadelphia. Our growth is only accelerating. We’re looking for deeply curious and highly ambitious people who want to have a real world impact. Come build the future of (field) work with us. ## About the role - You’ll be our first full-time SRE/infrastructure engineer, so we’ll look to you for leadership on how to improve and scale our infrastructure to support each part of the platform. Our data pipeline, machine learning training platform, and web app could all benefit from further productionization. - Help us scale and harden the platform that schedules our pipelines, runs machine learning training, and hosts our web app. We run Apache Airflow on Astronomer with DAGs that orchestrate high-volume processing across AWS and Kubernetes, including machine learning inference inside pipeline tasks. You will build the observability and reliability foundations that let us run this system confidently as customer data volume grows: monitoring, alerting, performance/cost visibility, and clear operational practices. - Stay curious, collaborative, and cross-functional while also taking ownership of problems. We translate complex, real-world requirements from a critical industry into high-quality data products, so understanding the business holistically is key. We take pride in managing complexity and providing high-fidelity data that our customers can use to make better-informed decisions. ## Responsibilities - Partner with the data platform and engineering teams to understand how changes propagate across pipeline execution (Astronomer-hosted Airflow DAGs), containerized workers (Kubernetes), and AWS services (S3, SQS, Lambda, Step Functions, ECS). - Design and implement reliability and observability for high-volume pipeline operations, including: - actionable monitoring/alerting for DAG/task failures and reruns - visibility into operational workflows like flight orchestration (including DLQ/failed-message alerting and notification pathways) - dashboards and SLO/SLI definitions focused on correctness, throughput, and pipeline health - Own CI/CD guardrails for production changes: build/deploy validation and safe rollout mechanics for Astronomer deployments (image builds pushed to ECR, and Airflow configuration updates via Astronomer CLI variable updates) - Make machine learning inference operations more reliable and observable: - instrument inference runs executed inside pipeline runners (model checkpoint resolution, S3 sync behavior, thresholds and fallback behavior, and output correctness) - add operational visibility for inference outcomes (e.g., unknown classification rates, fallback usage, and failure modes) - Create operational tooling and continuously improve systems (‘leave it better than you found it’), including: - runbooks, incident learnings, and engineering standards for debugging at scale - automate away toil in deployment and operations workflows as we learn what hurts most On-call / incident response There is not currently an established on-call rotation for this platform, and the pipelines do not require real-time processing. That said, you’ll still help lead reliability improvements and operational readiness—so the team has faster diagnosis, better alerts, and safer releases when issues do occur. ## What we’re looking for - You are an experienced software engineer where the last 7-10 years required significant time on observability, systems/infrastructure engineering, SRE, or DevOps (ideally in a cloud environment). - Ability to reason about architecture end-to-end and articulate your thoughts with product impact in mind (data movement, execution, failure handling, and operational visibility). - Hands-on experience with infrastructure-as-code (Terraform and similar) and using it to deliver reliable environments. - Experience with container orchestration and debugging in practice (Kubernetes and/or ECS/container-based deployments). - Strong Linux debugging skills and demonstrated ability to investigate production issues with logs/metrics and clear hypotheses. - Empathy and communication: you can collaborate effectively with engineers across teams (especially the data platform team) and explain tradeoffs clearly. Nice-to-haves - Experience working in early-stage or fast-moving environments where ownership and processes evolve quickly. - Experience with Apache Airflow and/or Astronomer. - Experience with AWS, although other cloud providers are fine. (DuploCloud experience is also helpful.) - Experience with geospatial/imagery/lidar/point-cloud style domains. - ML Ops skills (model deployment/inference reliability, packaging, CI/CD for model artifacts, and operational observability for inference pipelines). Work location This is a full-time, hybrid role based out of our Lower Manhattan, NYC office (2 days per week in person, currently pinned to Tuesdays and Wednesdays). ## Benefits - Competitive salary and equity package - Comprehensive medical, dental, and vision coverage for you and your eligible dependents - Life insurance and short- and long-term disability coverage - 16 weeks of fully paid parental leave to support all new parents - Flexible, unlimited paid time off - 401(k) retirement savings plan - Free OneMedical membership - Commuter benefits - Snacks, goodies, and team lunches provided twice a week to keep you fueled and connected with your colleagues. Salary The estimated salary range for this position is $200,000 - 230,000 USD. Total compensation for this position is determined by skills, qualifications, relevant work experience, location, and other factors. This salary estimate excludes the value of any potential bonuses; the value of any benefits offered; and the potential future value of any long-term incentives. This information is provided per the New York City Human Rights Law. Please note that the range provided is applicable only to New York City-based applicants. Base compensation may vary if the work location is outside of New York City. Treeswift  is proud to be an equal opportunity employer. We provide employment opportunities without regard to age, race, color, ancestry, national origin, religion, disability, sex, gender identity or expression, sexual orientation, veteran status, or any other protected status in accordance with applicable law. If you require any accommodations during the recruitment process, whether it be alternate forms of material, accessible meeting rooms, etc., please let us know and we will work with you to meet your needs. ## About Treeswift Inc ## Company Overview - **One-liner**: Treeswift builds physical AI and robotics solutions that augment field crews for utility grid construction, vegetation management, and asset inspection. - **Entity Type**: Private (Series A) - **Headquarters**: Philadelphia, Pennsylvania, United States (also maintains an office in lower Manhattan) - **Founded**: 2020 - **Founders**: Steven Chen (CEO & Co-Founder), Bianca Rahill-Marier (COO & Co-Founder), Michael Shomin (CTO & Co-Founder), Vaibhav Arcot (Co-Founder & Senior Software Engineer) ## Core Business - **Primary industry**: Physical AI / Robotics for Utility Infrastructure & Vegetation Management - **Target customers**: Large electric utilities across the United States (including PG&E and two of the other five largest utilities in the country) - **Mission statement**: Build physical AI that puts people at the center of the work, not on the sidelines — helping the people who build and maintain the physical world get more done. ## Products & Services - **Treeswift Field Data Platform**: A software platform that integrates with utilities’ existing GIS and planning tools. It processes data from backpack and vehicle-mounted sensors (imagery, LiDAR, GPS, voice) to deliver actionable analytics and 360° field context for construction, vegetation management, asset inspections, and disaster response. - **Backpack & Vehicle-Mounted Sensors**: Hardware kits that seamlessly integrate into existing patrols (on foot or by vehicle) to gather high-resolution ground measurements in any environment — from city sidewalks to remote backcountry. - **AI-Powered Analytics Engine**: Computer vision and machine learning tools that translate raw field data into automated measurements, counts, forms, and reports — reducing the manual busywork for highly skilled field crews. ## Market Standing - **Valuation/Market Cap**: Not publicly disclosed - **Key Metric**: Total funding of **$21.7M** across multiple rounds; estimated annual revenue of **$2.1M** - **Notable Investors/Partners**: Inspired Capital (Penny Pritzker), Pathbreaker Ventures, Crosslink Capital, TenOneTen Ventures, Contour Venture Partners, National Science Foundation (grant) - **Growth Signals**: - Headcount grew **100% YoY** (from ~11 to 36 employees) - Active job postings up **550% year-over-year** - Works with PG&E and two of the five largest U.S. utilities - Filed 5 patents (topics include forest ecology, forest management, forest modelling) - Expanding into disaster response and black sky (storm/outage) use cases ## Competitive Advantages - **Niche focus on electric grid labor augmentation**: Addresses a critical and growing need as electricity demand surges from data centers and new industry — utilities must build and maintain more infrastructure while managing wildfire, storm, and compliance risks. - **PhD-level robotics & AI team**: Founders and technical leads hold PhDs from University of Pennsylvania, Caltech, and Carnegie Mellon University, with deep expertise in robot perception, machine learning, and autonomous systems. - **Enterprise DNA**: Leadership team includes former product leaders from Palantir Foundry (built the product suite for business users), plus veterans from Stripe, Oracle, and MongoDB. - **Field operations expertise**: Team includes arborists and vegetation management specialists from the National Park Service and USFS, giving them rare domain credibility with utility crews. - **Patented technology**: 5 patents covering autonomous scanning, forest representation, and tracking/monitoring of geographic regions. ## Strategic Focus - **Scaling across the U.S. utility market**: Expanding partnerships with major utilities beyond PG&E to address the nationwide grid modernization challenge. - **Expanding use cases**: Moving from vegetation management into construction monitoring, asset inspection, and disaster response (both proactive “blue sky” and reactive “black sky” scenarios). - **Growing the engineering and field operations teams**: 13 active job openings spanning machine learning, site reliability, hardware operations, recruiting, finance, and deployment. - **Product development**: Continuing to build AI that works in any environment (urban, remote, all weather conditions) and integrates with any system utilities already use. ## Why Work Here - **Mission-driven work**: Tackle critical infrastructure and climate resilience challenges — helping prevent wildfires, improve grid reliability, and keep electricity affordable. - **Cutting-edge technology**: Work on physical AI combining robotics, computer vision, LiDAR, and machine learning in real-world outdoor environments. - **Strong team culture**: Described as “mission-driven experts” with backgrounds in robotics research (Penn, Caltech, CMU) and enterprise software (Palantir, Stripe, Oracle, MongoDB). - **High growth trajectory**: 100% headcount growth YoY, recently closed Series A, and expanding rapidly with 13 open positions. - **Hybrid/office presence**: Headquarters in Philadelphia (3401 Grays Ferry Ave) with an additional office in lower Manhattan; customer-facing team members located near utility partners across the country. - **Diverse team composition**: 36 employees across technical (30%), general management, consulting, product, project management, sales, HR, finance, and operations. - **Notable perks**: Opportunity to work on patented technology with real-world impact; strong alumni network from top tech companies and universities. ## Sources 1. [treeswift.com](https://www.treeswift.com/) 2. [treeswift.com/team](https://treeswift.com/team) 3. [linkedin.com/company/treeswift](https://www.linkedin.com/company/treeswift) 4. [cbinsights.com/company/treeswift](https://www.cbinsights.com/company/treeswift) 5. [jobs.ashbyhq.com/treeswift](https://jobs.ashbyhq.com/treeswift) ## Other roles at Treeswift Inc - [Mechanical Design Engineer](https://feeny.ai/job/mechanical-design-engineer-treeswift-inc-new-york-ce7ge0wkxmbf) — New York, NY - [Full Stack Software Engineer](https://feeny.ai/job/full-stack-software-engineer-treeswift-inc-new-york-f5gya3gp40k5) — New York, NY - [Robotics Engineer (Perception)](https://feeny.ai/job/robotics-engineer-perception-treeswift-inc-philadelphia-8x2rs4t0ybab) — Philadelphia, PA - [Robotics Engineer (Perception)](https://feeny.ai/job/robotics-engineer-perception-treeswift-inc-new-york-dq0ke5ag5p5c) — New York, NY - [Business Operations Associate](https://feeny.ai/job/business-operations-associate-treeswift-inc-new-york-44een4895rvb) — New York, NY - [Chief of Staff to CEO](https://feeny.ai/job/chief-of-staff-to-ceo-treeswift-inc-new-york-xv9bkw83gnzv) — New York, NY - [Senior Technical Recruiter](https://feeny.ai/job/senior-technical-recruiter-treeswift-inc-new-york-mcwbvqyejac8) — New York, NY - [Deployment Lead](https://feeny.ai/job/deployment-lead-treeswift-inc-new-york-3c4rmr05nmyz) — New York, NY - [Field Operations Technician](https://feeny.ai/job/field-operations-technician-treeswift-inc-new-york-fjhabsvk6pm4) — New York, NY - [Senior Account Executive](https://feeny.ai/job/senior-account-executive-treeswift-inc-remote-rrcqdpgbaby4)