--- title: 'Filmmaker / Storyteller at Inference' canonical: 'https://feeny.ai/job/filmmaker-storyteller-inference-san-francisco-ydsw1y7frf1a' type: 'job' last_seen: '2026-09-11' --- # Filmmaker / Storyteller at Inference - **Company:** Inference - **Location:** San Francisco, CA - **Employment:** full-time - **Work type:** onsite - **Posted:** 2025-05-29 - **Last confirmed live:** 2026-09-11 - **Apply:** https://jobs.ashbyhq.com/inference/fc28c0b4-3032-4817-ab9b-389057a924c8 ## Job description Filmmaker / Storyteller [Inference.net](http://Inference.net) is seeking a Filmmaker / Storyteller to join our team and help define the narrative of building the world's largest distributed GPU cluster. This role combines creative vision with eye-catching content production, crafting stories that capture the magic of what we're building while shipping content that resonates with our users. If you live and breathe video content and can find compelling narratives in complex technical work, we want to hear from you! ## About Inference.net We are building a real-time marketplace for AI inference that matches spare GPU capacity inside data centers with demand from developers building AI-powered applications. We currently operate the world's largest distributed GPU cluster, with over 5,000 GPUs, hundreds of individual operators, and millions of gigabytes of VRAM connected to the network at any given moment. We are a small, well-funded team working on difficult, high-impact problems at the intersection of AI and distributed systems. We primarily work in-person from our office in downtown San Francisco. Our investors include A16z CSX and Multicoin. We are high-agency, adaptable, and collaborative. We value creativity alongside technical prowess and humility. We work hard, and deeply enjoy the work that we do. ## About the Role As our in-house filmmaker, you'll be embedded with our engineering team, capturing the journey of next-generation AI infrastructure. You'll own our entire video content strategy from ideation to final cut, shipping weekly content that attracts eyeballs while authentically representing our mission. This is a high-output role that demands both creative excellence and experimental mindset. ## Key Responsibilities - Content Creation & Production: Produce 4+ short-form videos monthly, 1 commercial monthly, and 1 documentary quarterly, handling everything from concept to final delivery - Narrative Development: Extract compelling stories from our team and technology, making complex AI/distributed systems concepts accessible and engaging - Platform Strategy: Ship content weekly across X, YouTube, TikTok, LinkedIn, and emerging platforms, optimizing for each platform's unique audience - Creative Experimentation: Constantly test new formats, styles, and approaches to find what resonates with developers and tech enthusiasts - Brand Storytelling: Define and evolve our company's visual narrative as we scale from startup to industry leader - Team Building: After establishing our content foundation, recruit and manage freelancers, editors, and production talent to scale output ## What We're Looking For - Portfolio of Excellence: Demonstrated ability to create viral, high-quality video content across multiple formats - Storytelling Mastery: Exceptional ability to find and craft narratives, especially from technical or complex subject matter - Technical Production Skills: End-to-end video production capabilities including shooting, editing, motion graphics, and color grading - Platform Native: Deep understanding of what performs on modern platforms, from TikTok trends to YouTube optimization - High Velocity Mindset: Comfort shipping content weekly while maintaining quality standards - Collaborative Spirit: Ability to work closely with engineers and extract authentic stories from technical experts - Startup DNA: Thrives in ambiguous, fast-moving environments with changing priorities - On-Site Commitment: Available to work in-person from our SF office 5 days per week ## Nice to Have - Experience creating content for developer or B2B tech audiences - Background in documentary filmmaking or journalism - Motion graphics and animation skills - Experience managing creative teams or freelancers ## What You'll Create Your work will span from punchy 30-second clips that stop the scroll to thoughtful mini-documentaries about the future of AI infrastructure. Think less corporate video, more cinematic storytelling that happens to feature GPUs and distributed systems. You'll make content that developers share because it's genuinely good, not just informative. ## Compensation We offer competitive compensation, equity in a high-growth startup, and comprehensive benefits. The base salary range for this role is $100,000 - $140,000, plus competitive equity and benefits including: - Full healthcare coverage - Quarterly offsites - Flexible PTO If you're ready to own the visual narrative of AI's next chapter and can ship content that makes infrastructure feel like magic, we'd love to see your work. Please send your portfolio, resume, and a brief note about your favorite piece of content you've created to hiring@inference.net. ## About Inference ## Company Overview - **One-liner**: Inference.net provides a marketplace and infrastructure for AI-native teams to deploy, observe, evaluate, and train custom LLMs at dramatically reduced costs by utilizing otherwise wasted GPU capacity from data centers. - **Entity Type**: Private (Seed stage) - **Headquarters**: San Francisco, California, United States - **Founded**: 2023 - **Founders**: Amarjot Singh (Co-Founder), Ibrahim Ahmed (Co-Founder, CTO) ## Core Business - **Primary Industry**: AI Inference Infrastructure / Software Development - **Target Customers**: B2B; AI-native companies, startups, and enterprises spending over $50k/month on closed-source AI providers; digital banks; decentralized networks; and high-volume AI applications. - **Mission/Purpose**: "We believe efficient markets for AI inference will drive the widespread proliferation of artificial intelligence over the next decades, leading to unprecedented human flourishing on Earth and beyond. We aim to accelerate this process." ## Products & Services - **Inference.net API**: A pay-as-you-go, OpenAI-compatible API for serving open-source, custom, and fine-tuned LLMs. Offers 50-90% discounts compared to providers like OpenAI and Anthropic by aggregating spot compute from underutilized data center GPU capacity. - **Catalyst Deploy**: A deployment platform for hosting LLMs at massive scale across public cloud, private cloud, or hybrid environments, with a claimed 99.99% uptime. - **Catalyst Observe**: An LLM observability tool that traces every request path (prompts, tool calls, responses, downstream providers) and monitors latency, reliability, usage patterns, and quality signals. - **Catalyst Evaluate**: A model evaluation system that scores quality across any model or metric, using production traces to validate new model variants against baseline behavior before deployment. - **Catalyst Train**: Automatic fine-tuning workflows that turn production traces into training datasets. Allows users to train custom frontier-level language models fine-tuned to specific quality, cost, and latency targets in minutes. ## Market Standing - **Valuation/Market Cap**: Not publicly disclosed. - **Total Funding**: $11.8M (Series Seed, announced October 14, 2025). - **Notable Investors**: Led by Multicoin Capital and a16z CSX, with participation from Topology Ventures, Founders, Inc., and a group of angel investors. - **Growth Signals**: 100% headcount growth year-over-year (from 5 to 10 employees). LinkedIn follower growth of +173.6% year-over-year. Has deployed custom models for "some of the fastest-growing AI-native companies in the world," including a digital bank with 120M+ customers and a nutrition tracking app that scaled to 10M+ users. ## Competitive Advantages - **Unique Business Model**: Acts as a spot market for perishable GPU compute, purchasing underutilized data center capacity in small chunks. This creates an inherent cost advantage, passing 50-90% savings to customers. - **Custom Model Economics**: Their approach trains models up to 100x smaller than GPT-5-class systems that match or exceed frontier model performance for specific tasks, running 2-3x faster and costing up to 90% less. - **Differentiation Focus**: Pitching against "renting intelligence" from closed providers, arguing that custom models trained on proprietary data become a moat competitors cannot replicate. - **SOC 2 Type II Compliant**: Full compliance and operational oversight, enabling enterprise adoption. ## Strategic Focus - **Expand R&D**: Using seed funding to push the frontiers of model and infrastructure performance. - **Scale Customer Acquisition**: Targeting companies spending over $50k/month on closed-source AI, offering to cut costs and improve performance within 4 weeks. - **Continuous Improvement Loops**: Building systems that retrain models on fresh production data as use cases evolve, creating "models that get better every cycle." - **Multi-Model Platform**: Supporting integration with both provider-hosted models (OpenAI, Anthropic, Gemini) and open-source models on optimized infrastructure. ## Why Work Here - **High-Impact Role in AI Infrastructure**: Working at the intersection of cutting-edge LLM research and practical infrastructure engineering, directly enabling the economics of AI for other companies. - **Tiny, High-Caliber Team**: Only 10 employees, plus 4 active job openings, suggesting a lean, high-autonomy culture where individuals have outsized impact. - **Strong Backing**: Backed by top-tier investors including Multicoin Capital and a16z, providing stability and resources despite being an early-stage company. - **Office Policy**: On-site / In-Office in San Francisco, CA (HQ in SoMa area). Employees work from a physical office, with typical time on-site being "None" (indicating potential flexibility). - **Culture Signals**: Described as mission-driven ("human flourishing on Earth and beyond"), with a focus on technical excellence and economic efficiency. The company openly shares its philosophy and strategy in blog posts. - **Active Roles**: Looking for Machine Learning Researchers, Fullstack Engineers (Frontend Focus), Senior Software Engineers (Model Performance), and Applied Machine Learning Engineers – all of which touch core product and research. ## Sources 1. [Inference.net Website](https://inference.net/) 2. [Inference.net Company Page](https://inference.net/company/) 3. [LinkedIn Page](https://www.linkedin.com/company/inference-net) 4. [Built In Profile](https://builtin.com/company/inferencenet) 5. [Seed Round Announcement](https://inference.net/blog/seed-round/) ## Other roles at Inference - [Senior Software Engineer - Model Performance](https://feeny.ai/job/senior-software-engineer-model-performance-inference-san-francisco-ey1dbf876pf7) — San Francisco, CA - [Machine Learning Researcher](https://feeny.ai/job/machine-learning-researcher-inference-san-francisco-hm492bpfacsw) — San Francisco, CA - [Applied Machine Learning Engineer](https://feeny.ai/job/applied-machine-learning-engineer-inference-san-francisco-epe003kkzf34) — San Francisco, CA - [Fullstack Engineer - Frontend Focus](https://feeny.ai/job/fullstack-engineer-frontend-focus-inference-san-francisco-1s8ak7cryhnc) — San Francisco, CA