--- title: 'Senior Software Engineer | Kimchi at Cast AI' canonical: 'https://feeny.ai/job/senior-software-engineer-kimchi-cast-ai-bulgaria-croatia-estonia-greece-hungary-a0qeg4xcmcs8' type: 'job' last_seen: '2026-09-12' --- # Senior Software Engineer | Kimchi at Cast AI - **Company:** Cast AI - **Location:** Bulgaria / Croatia / Estonia / Greece / Hungary / Latvia / Lithuania / Poland / Romania / Slovakia / Slovenia / Ukraine - **Posted:** 2026-04-30 - **Last confirmed live:** 2026-09-12 - **Apply:** https://cast.ai/careers/apply/?gh_jid=4235255009 ## Job description Why Kimchi? Kimchi is the AI platform inside CAST AI. We started by helping companies run LLMs on their own Kubernetes clusters and now we're providing a managed variant of those same capabilities. Our Infrastructure today Multi-model inference (MiniMax, Kimi, GLM-5, Nemotron, DeepSeek) with intelligent routing, an OpenAI-compatible API and deployment ranging from our GPUs to your own VPC. The inference layer is the foundation and the API is what sits in front of it as the primary channel for broadly and reliably distributing our AI services and powers our own Kimchi harness. We are hiring across multiple teams! As a Senior Software Engineer, you will have the opportunity to work on different key features of our product. All of these are high-agency roles across multiple parts of the tech stack that minimize process friction that would otherwise prevent you from shipping. In every team you will own features end-to-end: design, implementation, testing, production rollout. Most projects ship in 1-4 weeks. You'll work directly with product and other engineering teams on problems that don't have textbook solutions. We are currently hiring Senior Software Engineers for the following teams: API Platform Owns the inference API responsible for delivering our AI services across the world while making sure it's reliable and capable to scale in tandem with our company's growing ambitions, as well as our analytical platform, billing and role-based controls that enable our users to monitor and control their usage with ease - be it as a solo developer or a large enterprise. You'll own our infrastructure, datastores, analytics, observability and CI/CD pipeline. Responsibilities: - Develop with observability in mind, identify bottlenecks and optimize for performance. When p99 latency climbs, you find the cause through query profiles and flame graphs instead of raising the alert threshold. - Design the datastores and distributed systems behind the inference API, and keep them reliable as usage scales from a solo developer to a large enterprise. - Build the billing, usage metering, and role-based controls that let users monitor and govern their own consumption. Catch correctness problems where a wrong number costs you trust. Harness OpenAI and Anthropic ship models. They also ship one harness each – the scaffolding that turns a raw model into something that can plan, execute, recover, and complete work. We ship a different kind of harness: one built for cost-conscious, long-horizon autonomy, running on inference infrastructure we control end-to-end. A decent model with a great harness beats a great model with a bad harness. We've watched this play out. The gap between what today's models can do and what you see them doing is largely a harness gap – and that gap is where we operate. Responsibilities: - Architect planner/executor/evaluator pipelines – planning with a reasoning model, execution with a fast one, evaluation with a third. No self-verification. - Manage agent memory and context – state persistence across sessions, context compaction, tool-call offloading - Own the harness surface - TUI, MCP integrations, telemetry. - Work directly with users - including our own colleagues - to identify, understand and fix pain points Agent Platform The Agent Platform squad builds the infrastructure that lets teams run fleets of AI agents securely, whether in the cloud or self-hosted, so work can move confidently from a developer's laptop to production environments. The platform enforces least-privilege access to resources, requires human approval before agents can touch anything critical, and logs every action agents take for full auditability. The goal is to give engineering and operations teams the confidence to scale up agent usage without sacrificing security, control, or visibility into what their agents are actually doing. Responsibilities: - Design and build distributed systems that run agent fleets at scale across cloud and self-hosted environments, with security at its core - Make every agent action observable, auditable, and governable - Develop a remote-first experience of working with agents ## Requirements - Production experience with Go or Typescript is strongly preferred; candidates without either should demonstrate strong systems programming skills in a comparable language. - Strong debugging, optimization, and performance-tuning skills – including query profiling, index design, and database performance tuning beyond ORM usage. - Hands-on experience with cloud platforms (AWS, GCP, or Azure) and Kubernetes is a strong plus - Observability tooling (Prometheus, Grafana, OpenTelemetry), CI/CD and DevOps practices experience. - Startup mindset: adaptable, proactive, and comfortable with ambiguity. - Strong English skills, both verbal and written. - You've personally driven a complex project end-to-end. - (Agent Platform) Experience in virtualization, networking, security, and Kubernetes internals is a strong plus - (Harness) Experience on working with harnesses and creating your own AI workflows is a strong plus What’s in it for you? - Competitive salary (€6,500 - €9,000 gross, depending on the level of experience). - Enjoy a flexible, remote-first global environment. - Collaborate with a global team of cloud experts and innovators, passionate about pushing the boundaries of Kubernetes technology - Equity options. - Get quick feedback with a fast-paced workflow. Most feature projects are completed in 1 to 4 weeks. - Spend 10% of your work time on personal projects or self-improvement. - Learning budget for professional and personal development - including access to international conferences and courses that elevate your skills. - Annual hackathon to spark new ideas and strengthen team bonds. - Team-building budget and company events to connect with your colleagues. - Equipment budget to ensure you have everything you need. - Extra days off to help maintain a healthy work-life balance. Hiring process - Screening call with Recruiter - Hiring Manager interview - Technical interview (system design) - Live coding - Culture Check interview with an executive As part of our standard hiring process, we would like to inform you that a background check may be conducted at the final stage of recruitment through our third-party provider, Checkr. Please note that Cast AI does not provide any form of visa sponsorship/work permit. #LI-Remote ## About Cast AI ## Company Overview - **One-liner**: Cast AI provides an Application Performance Automation (APA) platform that continuously optimizes Kubernetes infrastructure for cost, performance, and reliability with minimal manual intervention. - **Entity Type**: Private (Series C; total funding ~$185.8M–$199.9M per conflicting reports) - **Headquarters**: Miami, Florida, United States - **Founded**: 2019 - **Founders**: Leon Kuperman (CEO), Laurent Gil (President), Yuri Frayman (CTO) ## Core Business - **Primary industry**: Kubernetes optimization, cloud cost management, and application performance automation. - **Target customers**: Platform, SRE, and FinOps teams at B2B companies running Kubernetes (SMB to enterprise). - **Mission/purpose statement**: “The cloud as promised: fast, reliable and cost-efficient.” [cast.ai/about-us](https://cast.ai/about-us/) ## Products & Services - **Cast AI APA Platform**: A SaaS platform that ingests real-time workload, cost, and SLO signals to autonomously rightsize pods, scale nodes, optimize GPU and Spot instances, and self-heal operational issues. Supports AWS EKS, Azure AKS, Google GKE, and on-prem clusters. Includes cost and performance dashboards, agentic runbooks, and workload-aware predictive models. [cast.ai](https://cast.ai/) ## Market Standing - **Valuation/Market Cap**: Not publicly disclosed. - **Key Metric**: Annual Revenue reported as ~$60M (LinkedIn); Total Funding ~$185.8M (LinkedIn) or ~$199.9M (CB Insights – conflicting reports). - **Notable Investors/Partners**: Specific investors are not named in available sources; funding rounds include Series B ($35M, Nov 2023, 3 investors), a venture round ($20M, Mar 2023), and Series C ($108M, Apr 2025, 8 investors). [linkedin.com/company/cast-ai](https://www.linkedin.com/company/cast-ai) | [cbinsights.com/company/cast-ai](https://www.cbinsights.com/company/cast-ai) - **Growth Signals**: 2100+ customers globally; 40% cloud waste reduction claimed; 270 employees (+19.4% YoY); ranked #1 out of 223 solutions in its category; offices in Miami, New York, Vilnius, London, Tel Aviv, Bengaluru, Dallas. [cast.ai/about-us](https://cast.ai/about-us/) | [linkedin.com/company/cast-ai](https://www.linkedin.com/company/cast-ai) ## Competitive Advantages - **Predictive engine trained on massive dataset** from thousands of clusters and millions of workloads – moves beyond rule-based automation to workload-aware decisions. - **App-aware reliability**: Predicts spot instance interruptions up to 30 minutes before they occur, enabling graceful workload migration. - **Agentic runbooks** that self-heal drift, image issues, policy violations while requiring human approval for changes. - **Works in read-only mode initially** – no infrastructure changes required to start, lowering adoption risk. ## Strategic Focus - Expanding automation to GPU/AI workload optimization alongside traditional Kubernetes rightsizing. - Deepening integrations with cloud providers (AWS, Azure, Google) and extending agentic capabilities. - Scaling global presence with remote-first teams and new offices in key tech hubs. ## Why Work Here - **Remote-first** culture with flexible hours and fully supported remote work; offices available in multiple countries. [cast.ai/careers](https://cast.ai/careers/) - **Benefits**: Stock options from day one, annual hackathon, home office allowance, learning & development budget, comprehensive health insurance, bonus days off, company events. - **Engineering culture**: Values “practice customer obsession,” “develop and hire the best,” “lead by ownership,” and “expect and advocate change.” Emphasis on innovation and fast learning. - **Glassdoor style**: Employer rating 4.3/5 (57 reviews); culture 4.0, compensation 4.2, career 4.2. [linkedin.com/company/cast-ai](https://www.linkedin.com/company/cast-ai) ## Sources 1. [cast.ai/about-us](https://cast.ai/about-us/) 2. [cast.ai](https://cast.ai/) 3. [cast.ai/careers](https://cast.ai/careers/) 4. [linkedin.com/company/cast-ai](https://www.linkedin.com/company/cast-ai) 5. [cbinsights.com/company/cast-ai](https://www.cbinsights.com/company/cast-ai) ## Other roles at Cast AI - [Founding Account Executive | Enterprise | Australia](https://feeny.ai/job/founding-account-executive-enterprise-australia-cast-ai-anz-73xwj497hmhq) — ANZ - [Kimchi Sales Engineer | EMEA (Remote)](https://feeny.ai/job/kimchi-sales-engineer-emea-remote-cast-ai-europe-447dtf2vpwha) — Europe, the Middle East and Africa / European Union - [Sales Engineer | Singapore](https://feeny.ai/job/sales-engineer-singapore-cast-ai-singapore-mm1d5fhe52kz) — Singapore - [Account Executive | Enterprise | Korea](https://feeny.ai/job/account-executive-enterprise-korea-cast-ai-seoul-seoul-south-bz74mpbe0na9) — Seoul Seoul South, South Korea - [Commercial Counsel](https://feeny.ai/job/commercial-counsel-cast-ai-united-states-r7xd90wtxcwe) — United States - [Senior Field and Partner Marketing Manager | APAC](https://feeny.ai/job/senior-field-and-partner-marketing-manager-apac-cast-ai-apac-india-rrk33y985m7b) — APAC / India - [Senior Product Marketing Manager | Database Optimization](https://feeny.ai/job/senior-product-marketing-manager-database-optimization-cast-ai-netherlands-3dtka7netpc6) — Netherlands - [Field and Partner Marketing Manager | EMEA](https://feeny.ai/job/field-and-partner-marketing-manager-emea-cast-ai-europe-y056q4th0asj) — Europe - [Senior People Operations Specialist](https://feeny.ai/job/senior-people-operations-specialist-cast-ai-lithuania-mnyzbzgad0ga) — Lithuania - [Founding Account Executive | Database Optimization](https://feeny.ai/job/founding-account-executive-database-optimization-cast-ai-united-states-8p0n55cd8vam) — United States