Cohort AI Inc.

Associate Data Engineer at Cohort AI Inc. (Naperville, IL)

Cohort AI Inc.· Naperville, IL·

Role details

Work type
Remote
Employment
Full-Time

Job description

ABOUT THE ROLE

We’re looking for an Associate Data Engineer to join our team and help make data reliable, usable, and impactful.

In this role, you’ll work closely with different teams to bring in new datasets, clean and transform them, and make sure they’re ready for analysis. A big part of your work will involve building and maintaining data pipelines, mapping data into standard formats, and ensuring everything runs smoothly behind the scenes.

This is a great opportunity if you’re early in your data engineering career and enjoy working hands-on with data, solving real-world problems, and learning how large-scale data systems operate—especially in the healthcare space.

RESPONSIBILITIES

  • Build and maintain data pipelines to ingest, transform, and organize data from multiple sources
  • Work with clinical, claims, or similar structured datasets and map them into standardized data models
  • Run data quality checks to ensure accuracy and consistency
  • Use tools like Databricks, BigQuery, Redshift, dbt, and command-line utilities
  • Collaborate with cross-functional teams including data, product, and business stakeholders
  • Help maintain ongoing data refresh processes and troubleshoot pipeline issues when needed

NEEDED

  • Bachelor’s degree in Computer Science or a related field (or equivalent hands-on experience)
  • Strong working knowledge of SQL
  • Familiarity with Python, PySpark, or SparkSQL
  • Experience with modern data platforms like Databricks, Snowflake, or BigQuery
  • Comfort working in a remote, collaborative environment
  • Based in the United States (preferred time zones: Central, Mountain, or Pacific)

Why work at Cohort AI Inc.

  • Remote-first culture: Most roles are fully remote (US or India), with flexible location options.
  • AI-first engineering: Opportunity to build production-grade AI systems and data pipelines from scratch.
  • Tight-knit team: Small (~11 people), flat structure, high ownership.
  • Impact: Shape a product that directly solves a critical pain point for startups.
  • Values: “No fluff. No pitch decks. Just results.” – fast-paced, execution-oriented environment.
  • Tech stack: Python, SQL, Google Cloud, AWS Lambda, Docker, PostgreSQL, Elasticsearch, GitHub Actions, microservices.

Application questions