--- title: 'Senior Back End Engineer at Troveo AI' canonical: 'https://feeny.ai/job/senior-back-end-engineer-troveo-ai-remote-8bqtamecpq9z' type: 'job' last_seen: '2026-09-15' --- # Senior Back End Engineer at Troveo AI - **Company:** Troveo AI - **Location:** Remote - **Employment:** full-time - **Work type:** remote - **Posted:** 2026-07-02 - **Last confirmed live:** 2026-09-15 - **Apply:** https://jobs.ashbyhq.com/troveo/e68eae0d-f637-4d63-9603-d9ffbbf284d6 ## Job description ## About Troveo Troveo builds the data platform that AI labs and model builders need to train the next generation of models. We have created the world's largest licensed platform of scarce, proprietary data for AI, spanning video, audio, text, and business workflows. Troveo indexes, enriches, and packages this high-quality data into formats ready for training, fine-tuning, evaluation, and agentic use cases. Backed by top investors, we’re a small, high-impact team solving one of the biggest bottlenecks in AI development. ## About the Role We are hiring a Senior Backend Software Engineer with deep expertise in building scalable, reliable systems. You will design and operate backend services, with strong expertise in Elasticsearch, container orchestration, and systems architecture.This is a high-impact senior role where you will be building and evolving backend systems while driving best practices in distributed architecture and observability. ## Key Responsibilities - Design, develop, and maintain scalable backend services and APIs using microservices architecture - Build and operate production workloads on Kubernetes (deployments, services, ingress, autoscaling, Helm, etc.) - Work extensively with Elasticsearch - indexing, querying, aggregations, performance tuning, and cluster management - Apply distributed systems concepts to design fault-tolerant, highly available, and scalable systems - Optimize application performance and reliability in distributed and containerized environments - Design and implement robust CI/CD pipelines with GitOps and Kubernetes-native deployments - Troubleshoot complex issues across services, ECS and EKS clusters, and Elasticsearch - Improve system observability, monitoring, and alerting in distributed environments - Participate in architectural decisions and mentor engineers on distributed systems and Kubernetes best practices - Collaborate with platform teams to evolve internal tooling and developer experience Job Requirements - 10+ years of professional software engineering experience (strong backend focus) - Strong hands-on experience with Kubernetes (minimum 3+ years running production workloads) - Solid experience working with Elasticsearch (indexing strategies, query optimization, cluster scaling, etc.) - Strong understanding of distributed systems concepts (consistency models, partitioning, replication, consensus, failure handling, CAP theorem, etc.) - Proficiency in at least one backend language: Go, Python, or Java. - Experience with Docker and container orchestration - Experience with cloud-managed Kubernetes (EKS, GKE, or AKS) - Good understanding of microservices architecture and inter-service communication patterns - Experience with relational and/or NoSQL databases - Familiarity with infrastructure-as-code tools (Terraform preferred) ## Preferred Requirements - Experience with GitOps tools (Argo CD / Flux) - Knowledge of service mesh (Istio, Linkerd) - Experience with workflow orchestration tools (Temporal, Airflow, etc.) - Background in building or operating search-heavy systems - Experience with observability stacks (Prometheus, Grafana, OpenTelemetry, ELK stack) - Contributions to open-source projects related to Kubernetes or distributed systems ## Compensation - Base Salary: $175,000 – $195,000 (depending on experience and location) - Equity: Competitive equity package in a well-funded AI startup with significant upside Compensation is location-adjusted for cost of living. We are open to candidates in California, New York, and select other states. ## What We Offer - Comprehensive Health Benefits: Medical, dental, and vision coverage (100% employer-paid for employees) - Flexible PTO & Paid Holidays: Unlimited PTO with encouragement to actually use it - Remote First Policy: Work from anywhere in the US (with occasional team offsites) - Learning & Growth: Annual learning stipend, access to top conferences, and direct mentorship from experienced founders - Equity Ownership: Competitive equity package with clear growth potential as we scale - Modern Tech Stack & Tools: Budget for the best equipment and software - Strong Culture: High-trust, low-ego environment focused on impact, transparency, and work-life balance We believe great work happens when people are supported, challenged, and given ownership. ## Equal Opportunity Employer Troveo is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. We do not discriminate based on race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status. ## About Troveo AI ## Company Overview - **One-liner**: Troveo builds the world’s largest licensed platform of real-world data for AI, helping frontier labs and model builders source, curate, and deploy training-ready datasets. - **Entity Type**: Private (Seed stage; total funding $5.5M) - **Headquarters**: San Francisco, California, United States - **Founded**: 2024 - **Founders**: Marty Pesis (Co-founder & CEO) ## Core Business - **Primary industry/industries**: AI training data infrastructure, data licensing, enterprise AI - **Target customers**: Frontier AI labs, model builders, robotics companies, and large technology companies (B2B, enterprise) - **Mission or purpose**: “Accelerating foundational AI model development by unlocking non-public, real-world data for training, fine-tuning, evaluation, and agentic use cases.” ## Products & Services - **Video Training Data**: Over 8 million hours of licensed footage across 30+ content categories, with custom metadata and full legal compliance (BIPA, CUBI) – used for foundational video models, avatars, and multi-camera tasks. [[troveo.ai]](https://www.troveo.ai/) - **Audio Training Data**: 4 million hours of single/multi-channel audio in dozens of languages for speech recognition, voice assistants, and conversational AI. [[businesswire.com]](https://www.businesswire.com/news/home/20260428319383/en/) - **Text Data**: Billions of words from publishers and rights holders, structured for training, fine-tuning, and evaluation. [[businesswire.com]](https://www.businesswire.com/news/home/20260428319383/en/) - **Agentic Trajectories (Enterprise Workflow Traces)**: Real-world business data from companies capturing actual enterprise workflows for agentic AI. [[businesswire.com]](https://www.businesswire.com/news/home/20260428319383/en/) - **Gameplay Data**: Video game data with time-synced keystroke and character progression metadata for world models. [[businesswire.com]](https://www.businesswire.com/news/home/20260428319383/en/) - **Egocentric Robotics Data**: First-person perspective data from real operating environments for robotics training. [[businesswire.com]](https://www.businesswire.com/news/home/20260428319383/en/) - **Bespoke Data Sourcing**: Custom dataset assembly based on client briefs, with capabilities like golden dataset calibration and rolling deliveries (e.g., 500K+ clips in three weeks). [[troveo.ai]](https://www.troveo.ai/) ## Market Standing - **Valuation/Market Cap**: Not publicly available - **Key Metric**: Total funding of $5.5M (Seed round $4.5M led by Seven Seven Six, Nov 2024; Pre-Seed $1M, May 2024) [[linkedin.com]](https://www.linkedin.com/company/troveo) - **Notable Investors/Partners**: Seven Seven Six (lead investor), plus three other unnamed seed investors. [[linkedin.com]](https://www.linkedin.com/company/troveo) - **Growth Signals**: 47.1% headcount growth YoY (21 employees as of mid-2026); expansion from video into five new data categories (audio, text, enterprise workflows, gaming, robotics); announced over $20 million in payouts to content owners; active relationships with top AI labs and some of the largest technology companies. [[businesswire.com]](https://www.businesswire.com/news/home/20260428319383/en/) [[linkedin.com]](https://www.linkedin.com/company/troveo) ## Competitive Advantages - **Exclusive licensed data**: Troveo’s data is sourced directly from content owners (broadcast archives, studio vaults, enterprise systems) and is non-public – not scraped from the internet – providing a legal moat for customers. - **Compliance-first approach**: Every dataset is verified for compliance with biometric privacy laws (Illinois BIPA, Texas CUBI) and fully rights-cleared. - **Scale and speed**: Library of 8M+ hours video, 4M hours audio; ability to deliver training-ready clips in as little as three weeks with custom metadata. - **End-to-end service**: From curated off-the-shelf datasets to bespoke sourcing, calibration, and rolling deliveries. ## Strategic Focus - **Category expansion**: Rapidly scaling into audio, text, enterprise workflows, gaming, and robotics to capture demand beyond video. - **Building the “data infrastructure” for AI**: Positioning itself as the essential pipeline for frontier labs needing scarce, legally defensible training data. - **Growing content creator payouts**: Reached $20M paid to rights holders, indicating a marketplace model that incentivizes more supply. ## Why Work Here - **Early-stage impact**: Join a 21-person team (2024-founded) with strong growth (+47% headcount YoY) – opportunity to shape data infrastructure for the AI industry. - **Hybrid workspace**: Employees combine remote and on-site work at the San Francisco office. [[builtin.com]](https://builtin.com/company/troveo-ai) - **Open roles signal growth**: Currently hiring for VP of Engineering, Senior ML Engineer, Lead Software Engineer, Software Engineer (Delivery), Sales positions (Account Executive, VP of Sales) – indicating engineering and go-to-market expansion. [[jobs.ashbyhq.com]](https://jobs.ashbyhq.com/troveo) [[builtin.com]](https://builtin.com/company/troveo-ai) - **Culture**: Described as a technology, information and media company with a focus on solving the “biggest bottleneck” in AI (training data). The company values legal compliance, data quality, and real-world authenticity. - **Team composition**: 24% of employees in technical roles, with talent sourced from companies like TikTok, Dialpad, and Turing. [[linkedin.com]](https://www.linkedin.com/company/troveo) ## Sources 1. [Troveo Website](https://www.troveo.ai/) 2. [LinkedIn Company Page](https://www.linkedin.com/company/troveo) 3. [Built In Company Profile](https://builtin.com/company/troveo-ai) 4. [BusinessWire Press Release (Apr 28, 2026)](https://www.businesswire.com/news/home/20260428319383/en/Troveo-Accelerates-AI-Model-Development-Expands-AI-Training-Data-Platform-to-Five-New-Categories-Announces-%2420-Million-in-Payouts) 5. [Troveo Careers Page (Ashby)](https://jobs.ashbyhq.com/troveo) ## Other roles at Troveo AI - [Senior Product Manager](https://feeny.ai/job/senior-product-manager-troveo-ai-remote-baax71nqcvdj) - [Senior Back End Engineer](https://feeny.ai/job/senior-back-end-engineer-altruist-san-francisco-71vav8tssb60) — San Francisco, CA - [Senior Back End Engineer](https://feeny.ai/job/senior-back-end-engineer-altruist-los-angeles-kn394da3s1v1) — Los Angeles, CA - [Senior Back End Engineer](https://feeny.ai/job/senior-back-end-engineer-abacum-spain-yys85px6b7xc) — Spain - [Senior Back End Engineer](https://feeny.ai/job/senior-back-end-engineer-thrill-labs-europe-z5x0ntq34v8b) — Europe - [Senior Back End Engineer](https://feeny.ai/job/senior-back-end-engineer-horizon3-ai-united-states-ewjeg1dadfvh) — United States