--- title: 'Founding Sales Engineer at Featherless AI' canonical: 'https://feeny.ai/job/founding-sales-engineer-featherless-ai-us-bfj59k5ej0kf' type: 'job' last_seen: '2026-09-15' --- # Founding Sales Engineer at Featherless AI - **Company:** Featherless AI - **Location:** US &, Canada - **Compensation:** $150k–$190k - **Employment:** full-time - **Work type:** remote - **Posted:** 2026-09-15 - **Last confirmed live:** 2026-09-15 - **Apply:** https://jobs.ashbyhq.com/featherlessai/10987253-c5e8-4ce4-81d2-a077d1989d99 ## Job description ## About Featherless [Featherless.ai](http://Featherless.ai) is building the world's most reliable open-model inference platform. Backed by AMD, Airbus Ventures, 500 Global, Kickstart Ventures, HF0 Residency, Panache Ventures, and Oakseed Ventures, we're a well-funded team of researchers and engineers on a mission to democratize AI through performance and accessibility. Our cloud provides instant access to 40,000+ open-source AI models and the research innovations powering next-generation efficiency, reliability, and model optimization. ## About the Role We're hiring our first Sales Engineer to pair directly with our founding Account Executive in North America. You are the technical half of a two-person deal team: the AE owns the commercial motion, you own the technical win. Our buyers are developers, ML engineers, and CTOs, and they don't buy on slides. They buy when someone credible sits with them, looks at their workload, and shows them that open models on Featherless are faster, cheaper, and more reliable than what they're running today. That's your job — hands on keyboard, in the customer's stack, from first technical call through production rollout. Because you're the first SE here, you'll build the function as you go: the demo environment, the POC playbook, the benchmark harness, the security answers, the migration guides. Every deal you win should make the next one easier to win. If you like being the person who can actually answer the hard question in the room — and want the technical foundation of a Series A GTM engine to be yours — this is for you. ## What You'll Own Immediately - Partner with the founding AE as the named technical owner on every active opportunity in North America - Run technical discovery: map each prospect's models, workloads, latency and throughput targets, spend, and constraints - Design and drive POCs and benchmarks that prove Featherless against the incumbent — closed-model APIs, another inference provider, or self-managed GPUs - Build the demo environment and reusable technical collateral that the whole GTM team sells with - Own technical objection handling end-to-end: performance, reliability, cost modeling, security, and data handling - Be the field's voice into product and engineering — you'll know what we're losing on before anyone else does Core Responsibilities - Lead technical calls tailored to each buyer's stack: architecture reviews, live demos, and working code against their real use case - Build migration paths off closed-model APIs and onto open weights — model selection, evaluation, prompt and output parity, cutover plan - Run benchmarks and produce the throughput, latency, and cost-per-token analysis that anchors the business case the AE builds - Scope and execute POCs with clear technical success criteria, then hold the customer and us to them - Write the technical sections of proposals, RFP responses, and security questionnaires; keep a reusable answer library so we never write the same answer twice - Support onboarding and first production workloads, then hand off cleanly and stay available for expansion - Feed structured field input to product and engineering — feature gaps, model coverage requests, reliability issues, competitive intel - Build and maintain the technical assets that scale the team: demo apps, notebooks, reference architectures, integration guides, internal enablement for BDRs and future AEs - Represent Featherless technically at conferences, meetups, and developer events - Keep POC and technical-stage detail current in HubSpot so forecasting reflects technical reality, not optimism ## What You Bring - 3–6 years in pre-sales engineering, solutions architecture, or a forward-deployed / customer-facing engineering role — at a GPU cloud, inference provider, MLOps platform, AI/developer tooling company, or cloud infrastructure vendor - Genuinely hands-on: you write Python comfortably and build your own demos rather than requesting them - Working knowledge of modern LLM inference — serving stacks (vLLM, SGLang, TensorRT-LLM or similar), OpenAI-compatible APIs, quantization, LoRA and fine-tune serving, batching and KV cache behavior, and what actually drives tokens/sec and cost - Practical familiarity with the open-model ecosystem: Hugging Face, the major open weight families, and how teams evaluate one model against another - Comfortable with containers, Kubernetes, and cloud networking and security fundamentals - Credible with both audiences in the same meeting — the ML engineer who wants the numbers and the CTO who wants the risk and cost story - Strong written communication: your benchmark writeups and architecture docs should be good enough to forward to a buyer's CEO - Entrepreneurial and self-directed — comfortable being the first SE, with no playbook and no one to escalate the technical answer to - Use AI tools heavily in your own workflow to research, prototype, and move faster Nice to have: experience selling or building on AMD GPUs / ROCm; exposure to enterprise security and compliance review; open-source contributions or public technical writing; experience as the first technical hire on a GTM team. ## Why Featherless - Build the sales engineering function at a company in one of the fastest-moving spaces in AI - Work as a true pair with the founding AE, and directly with the CRO and founders — small team, no layers, immediate impact - Real technical depth: 40,000+ open models and an in-house research team shipping inference and optimization work you'll get to sell - Be part of a small, high-performing team making open weight AI accessible to everyone - Competitive base + variable tied to the team's number, plus equity ## About Featherless AI ## Company Overview - **One-liner**: Featherless AI provides a serverless platform that offers API access to over 40,000 open-weight AI models from a single endpoint, designed for developers and enterprises. - **Entity Type**: Private (Series A) - **Headquarters**: San Francisco, California, United States - **Founded**: 2023 - **Founders**: Eugene Cheah (CEO, Co-Founder) ## Core Business - **Primary industry**: Artificial Intelligence Infrastructure / Serverless LLM Hosting - **Target customers**: B2B, serving developers, AI startups, and enterprises seeking scalable, cost-effective inference for open-source models. - **Mission or purpose**: To democratize access to all AI models by making them available for serverless inference, eliminating the need for server setup and complex infrastructure management. ## Products & Services - **Featherless API**: A unified API gateway providing instant access to over 40,000 open-weight models (e.g., DeepSeek, Llama, Mistral, Qwen, RWKV, GLM, Kimi) without setup or hosting. Pricing is flat-rate with unlimited tokens, starting at $25/month for up to 4 concurrent connections and 32K context, scaling to $200/month for higher tiers. Agent-specific plans ($100/month) include sandbox environments and persistent storage. The service emphasizes low latency, dependable uptime, and predictable costs. ## Market Standing - **Valuation/Market Cap**: Not publicly disclosed. - **Key Metric**: **Total Funding of $25M** — raised $5M in a Seed round (April 2025, led by Airbus Ventures) and $20M in a Series A round (announced ~May 2026, details still emerging). - **Notable Investors/Partners**: Airbus Ventures, Kickstart Ventures, Panache Ventures, BMW i Ventures, AMD Ventures, and 11 other investors. The platform is built by researchers contributing to RWKV, a Linux Foundation project. - **Growth Signals**: The company is on a rapid growth trajectory, with headcount increasing 46.7% year-over-year to 14 employees. The website boasts over 2,100 "stars" for its Discord community. The company has a global presence, operating in 9 countries (including Singapore, Canada, Czechia, UK, Belgium, and Sweden). Website traffic is strong (73,508 monthly visits, growing +19.9% month-over-month), and there are 43 active job postings, a 65.4% quarterly increase in hiring. The Series A announcement signals significant investor confidence. ## Competitive Advantages - **Extensive Model Library & Zero-Friction Access**: A single API key provides access to the entire Hugging Face trending library, including models up to 229B parameters, with no need to manage infrastructure. - **Unlimited-Token, Flat-Rate Pricing**: A strong differentiator in the "per-token" pricing era, offering predictable costs suitable for scaling, with tiers from $25 to $200/month. - **Build for Reliability & Performance**: Architecture designed for real workloads with low latency and dependable uptime, utilizing proprietary GPU orchestration and model load-balancing. - **Open-Source Roots & Community**: Built by researchers contributing to RWKV (a Linux Foundation project), the company is deeply embedded in the open-source AI ecosystem, which fosters trust and community-driven development. ## Strategic Focus - **Scaling the Platform & Enterprise Adoption**: The Series A funding will be used to expand AI infrastructure and grow the platform. The company is actively hiring for senior roles like Founding Account Executives and Business Development Reps, signaling a shift toward aggressive go-to-market and enterprise sales. - **Expanding Global Presence**: With a distributed team across the US, Europe, and Asia, the company is building a global, remote-first workforce. - **Technical Innovation**: Focused on continuous improvement of inference performance and cost-efficiency through their GPU orchestration system and model load-balancing. ## Why Work Here - **High-Growth Stage**: As a Series A startup with strong investor backing, this is an opportunity to join a company experiencing rapid scaling, which offers significant career growth and impact potential. - **Impact & Ownership**: Employees are likely to have high autonomy and a direct impact on the company's trajectory, from building core infrastructure to driving revenue. - **Remote-First & Global Team**: Based on the distributed headcount across 9 countries (US, Singapore, Canada, UK, Belgium, etc.), the company is clearly remote-first, offering flexibility in where you work. Job postings reflect opportunities in the US and Europe (e.g., Paris, Berlin). - **Cutting-Edge Technical Challenge**: The core work involves solving complex problems in AI inference, GPU orchestration, and MLOps, making it a compelling place for engineers and researchers passionate about AI infrastructure. - **Culture & Values**: The company's deep ties to open-source AI communities and its "flat-rate, no-surprises" pricing philosophy likely translate into a transparent, developer-friendly internal culture. The small, highly-skilled team (14 people) suggests a close-knit, high-performing environment. ## Sources 1. [Featherless.ai Website](https://featherless.ai/) 2. [Featherless AI LinkedIn](https://www.linkedin.com/company/feather-serverless-ai) 3. [Featherless AI Docs](https://featherless.ai/docs/overview) 4. [CB Insights Profile](https://www.cbinsights.com/company/recursal-ai) 5. [Featherless AI Jobs](https://jobs.ashbyhq.com/featherlessai) ## Other roles at Featherless AI - [Founding Business Development Rep (AI Cloud US/CA)](https://feeny.ai/job/founding-business-development-rep-ai-cloud-us-ca-featherless-ai-us-973psexhmg2e) — US &, Canada - [Founding Account Executive (AI Cloud)](https://feeny.ai/job/founding-account-executive-ai-cloud-featherless-ai-us-wb7s9fy8rf34) — US &, Canada - [Chief of Staff](https://feeny.ai/job/chief-of-staff-featherless-ai-san-francisco-15ydsmveqwsq) — San Francisco, CA - [Content Marketer](https://feeny.ai/job/content-marketer-featherless-ai-europe-xnsbz1evvkbs) — Europe - [Business Development Rep (AI Cloud)](https://feeny.ai/job/business-development-rep-ai-cloud-featherless-ai-europe-hd5fymsqsh5h) — Europe - [AI Researcher — Training Optimization](https://feeny.ai/job/ai-researcher-training-optimization-featherless-ai-remote-hyc2csn67p2p) - [AI Researcher – Multilingual Data](https://feeny.ai/job/ai-researcher-multilingual-data-featherless-ai-remote-aj24t2jw442p) - [AI Researcher — AI Architecture Research](https://feeny.ai/job/ai-researcher-ai-architecture-research-featherless-ai-remote-dprg8203nt10) - [AI Researcher — Distillation](https://feeny.ai/job/ai-researcher-distillation-featherless-ai-remote-zxs1tx4mwq1f) - [AI Researcher — Inference Optimization](https://feeny.ai/job/ai-researcher-inference-optimization-featherless-ai-remote-5q35t4rfgjve)