--- title: 'Member of Technical Staff – AI Research Engineer (Image/Video Foundation Models) at GenPeach AI' canonical: 'https://feeny.ai/job/member-of-technical-staff-ai-research-engineer-image-video-foundation-models-wh2r2dx6teq5' type: 'job' last_seen: '2026-09-05' --- # Member of Technical Staff – AI Research Engineer (Image/Video Foundation Models) at GenPeach AI - **Company:** GenPeach AI - **Location:** Zurich, Switzerland - **Employment:** full-time - **Work type:** remote - **Posted:** 2026-03-30 - **Last confirmed live:** 2026-09-05 - **Apply:** https://jobs.ashbyhq.com/genpeach/d2fd37d4-3c2f-4a76-bf4c-77f869649f20 ## Job description ## ABOUT GENPEACH AI GenPeach AI is a product-driven research lab building vertical multimodal foundation models for hyper-realistic human generation in image and video – designed for emotionally resonant, human-centered AI experiences. Our goal is to create tools that supercharge human creativity rather than replace it. We train models from scratch: proprietary datasets at massive scale, novel architectures and training recipes, large GPU clusters, and tight product integration so research ships to users quickly. We are a deeply technical team of around 10 people. We’re advised by Directors from Google DeepMind and backed by leading AI-focused funds and angels from OpenAI, Meta AI, Microsoft AI, Project Prometheus, and Fal. Collectively, our team, advisors, and angels have contributed to models including Meta’s Imagine/MovieGen and foundation-model work behind OpenAI’s Sora, plus Google’s Veo and Gemini. ## ABOUT THE TEAM You’ll join the research team working across image/video generation and multimodal understanding. You’ll work closely with other Research Engineers and Scientists, as well as Founders and help turn research into scalable training runs, strong evaluations, and production-ready systems. ## ABOUT THE ROLE We’re hiring an AI Research Engineer to help build and scale GenPeach’s foundation models end-to-end – from implementing new model ideas and training recipes, to owning the parts of the training stack that determine quality and speed, to pushing models through production constraints. This is a hands-on, high-ownership role. You’ll write research-grade code that becomes production-critical. ## IN THIS ROLE, YOU WILL - Implement and iterate on image/video generative model ideas (architecture, losses, conditioning, sampling, pre-training, distillation, post-training) - Own training performance end-to-end (distributed training, throughput, memory, stability, debugging scaling failure modes) - Build the experimentation loop (evals, ablations, reproducibility tooling, reporting, decision hygiene) - Build and improve VLMs for image/video captioning (data recipes, training strategies, model variants, evaluation) - Run high-iteration research: read papers when useful, implement ideas, validate empirically - Create captioning pipelines that improve generation training and product quality - Partner with inference/product to ship under real constraints (latency, cost, reliability, rollout safety) Build demos and prototypes to showcase capabilities and accelerate iteration YOU MIGHT THRIVE IN THIS ROLE IF YOU - Love the craft of experimentation: fast iteration, clear ablations, strong evals, and honest conclusions - Enjoy debugging messy real-world training runs (not just clean demos) - Can move between research and engineering: write clean code, ship utilities, and improve team velocity - Take ownership beyond your job description when needed (startup reality) - Communicate clearly and collaborate well in a small, senior team ## MINIMUM QUALIFICATIONS - Strong Python and PyTorch skills (4+ years of experience) - Experience implementing and training deep learning models (generative models, VLMs, LLMs, vision/video, or adjacent) - Solid understanding of training dynamics, optimization, and practical debugging - Ability to drive projects end-to-end with minimal supervision ## PREFERRED QUALIFICATIONS - Hands-on experience with diffusion/flow-based image or video generation, or large-scale generative modeling in adjacent domains - Experience with distributed training at scale (multi-node) and performance tuning (throughput/memory) - Experience building evaluation frameworks (offline metrics + human eval + regression tracking) - Strong intuition for data quality and dataset/labeling tradeoffs for training and captioning - Publications are a plus, but shipped impact and strong technical evidence matter more ## WHAT MAKES THIS ROLE UNIQUE - Build frontier image/video models and the VLM captioning systems that power them - Join a lean, senior team that holds a high engineering + research bar - Direct product impact: your training runs become real user-facing capabilities - Benchmark against the best in the world and compete on model quality through what we ship ## HOW WE WORK - You own outcomes end-to-end and are trusted with real responsibility - Direct, low-ego communication and fast feedback loops - Bias toward impact: measure → iterate → ship - Research discipline: clear ablations, reproducibility, and crisp decision-making ## LOGISTICS - Location: Zurich (Switzerland) or Warsaw (Poland) — onsite or hybrid. If you’re elsewhere, we’re open to remote (team/timezone fit considered). - Compensation: competitive salary + meaningful equity (level-dependent) - Interview process: quick screen → 2x technical rounds (practical + systems) → team fit/values ## WHAT WE OFFER - Visa sponsorship (where applicable); we’ll make a strong effort to relocate you to Switzerland or Poland if desired - Remote-friendly: work fully remote, hybrid, or on-site from our hubs - Regular offsites and in-person events to collaborate and connect - Flexible PTO ## About GenPeach AI ## Company Overview - **One-liner**: GenPeach AI is a product-driven research lab building vertical foundation models for hyper-realistic, emotionally resonant human image and video generation. - **Entity Type**: Private (Seed/Angel stage – backed by notable angels and founders from leading AI organizations) - **Headquarters**: Zurich, Switzerland (with a second hub in Warsaw, Poland) - **Founded**: 2025 - **Founders**: Not publicly named individually, but described as proven AI research and engineering leaders who built and shipped SOTA GenAI products at Meta AI and Wanna. ## Core Business - **Primary industry**: Generative AI / Multimodal Foundation Models - **Target customers**: B2B (developers, content creators, enterprises via API and user interfaces) and B2C (creative professionals seeking human-centered AI tools) - **Mission**: "Supercharge human creativity instead of replacing it" – building AI that offers unmatched creative freedom and human-centered experiences. ## Products & Services - **GenPeach Platform (in development)**: A suite of multimodal AI models and interfaces for generating photorealistic images and videos of humans. The stack includes large-scale diffusion models trained on proprietary datasets, accessible via API, web UI, and digital platforms. The company is currently in stealth/early-access waitlist mode. ## Market Standing - **Valuation / Funding**: Not publicly disclosed. The company is backed by angels and founders from OpenAI, Meta AI, Microsoft AI, Project Prometheus, and Fal. - **Key Metric**: Total employees ≈ 10 (as of early 2026), with 435 LinkedIn followers and 16.7% monthly headcount growth. - **Notable Investors/Partners**: Unnamed angels/founders from OpenAI, Meta AI, Microsoft AI, Project Prometheus, Fal. Advisors include Directors from Google DeepMind. - **Growth Signals**: Company emerged from stealth in late 2025; actively hiring for founding-member-level research and engineering roles; plans to expand presence in Switzerland, Poland, and Dubai. ## Competitive Advantages - **Specialized vertical foundation models**: Focus exclusively on generating hyper-realistic humans (image/video), differentiating from general-purpose models (e.g., Sora, Runway). - **Emotionally resonant AI**: Emphasis on multimodal, emotionally aware outputs built on proprietary datasets and novel architectures. - **Elite founding team**: Deep technical expertise from Meta AI, Google DeepMind, Wanna, and top AI startups, giving credibility and research velocity. ## Strategic Focus - **Current priorities**: Scaling the core research team, training next-generation diffusion models on proprietary human-centric data, and launching an integrated product that allows creative freedom without replacing human creators. - **Direction**: Continue pushing the frontier of visual GenAI while building a lean, fast-moving research lab that turns frontier ideas into real products. ## Why Work Here - **Culture**: "Lean, highly ambitious, fast-moving" – values speed, ownership, and a builder’s mindset over hierarchy. - **Remote/Hybrid/On-site**: Hybrid workspace with hubs in Zurich and Warsaw; remote-friendly globally (team members in US, Switzerland, Poland); visa sponsorship available for Switzerland, Poland, or Dubai. - **Perks**: Flexible PTO, regular team offsites, opportunity to work on bleeding-edge research (visual foundation models), direct impact on product and research direction. - **Engineering culture**: Emphasis on designing and training large-scale models (not just finetuning), high autonomy, and collaboration between research and product. ## Sources 1. [genpeach.ai](https://genpeach.ai/) 2. [genpeach.ai/careers](https://genpeach.ai/careers) 3. [builtin.com](https://builtin.com/company/genpeach-ai) 4. [LinkedIn](https://www.linkedin.com/company/genpeach-ai/) 5. [northdata.com (Swiss Commercial Register)](https://www.northdata.com/GenPeach%20AI%20AG,%20Z%C3%BCrich/CHE-380.073.622) ## Other roles at GenPeach AI - [Founder's Associate (Generalist)](https://feeny.ai/job/founder-s-associate-generalist-genpeach-ai-munich-z1x01ze48fs0) — Munich, Germany / Zürich, Switzerland - [Founding GTM Lead](https://feeny.ai/job/founding-gtm-lead-genpeach-ai-new-york-hspy65rtaqhs) — New York, NY - [Wildcard (Exceptional Talent)](https://feeny.ai/job/wildcard-exceptional-talent-genpeach-ai-warsaw-ev46a7syvcqb) — Warsaw, Poland - [Founding Member of Technical Staff – AI Research Scientist (Image/Video Foundation Models)](https://feeny.ai/job/founding-member-of-technical-staff-ai-research-scientist-image-video-foundation-bddkqn1yjtxp) — Zurich, Switzerland