--- title: 'Senior Software Engineer, Data Acquisition at People Data Labs' canonical: 'https://feeny.ai/job/senior-software-engineer-data-acquisition-people-data-labs-remote-1z71cbn97amb' type: 'job' last_seen: '2026-09-16' --- # Senior Software Engineer, Data Acquisition at People Data Labs - **Company:** People Data Labs - **Location:** Remote - **Employment:** full-time - **Work type:** remote - **Posted:** 2026-07-23 - **Last confirmed live:** 2026-09-16 - **Apply:** https://jobs.ashbyhq.com/people-data-labs/32b381fa-c60d-4bf8-ae95-b9cf4eb75e5a ## Job description Note for all engineering roles: with the rise of fake applicants and AI-enabled candidate fraud, we have built in additional measures throughout the process to identify such candidates and remove them. ## About Us People Data Labs (PDL) is the provider of people and company data. We do the heavy lifting of data collection and standardization so our customers can focus on building and scaling innovative, compliant data solutions. Our sole focus is on building the best data available by integrating thousands of compliantly sourced datasets into a single, developer-friendly source of truth. Leading companies across the world use PDL’s workforce data to enrich recruiting platforms, power AI models, create custom audiences, and more. We are looking for individuals who can balance extreme ownership with a “one-team, one-dream” mindset. Our customers are trying to solve complex problems, and we only help them achieve their goals as a team. Our Data Engineering & Acquisition Team ensures our customers have standardized and high quality data to build upon. You will be crucial in accelerating our efforts to build standalone data products that enable data teams and independent developers to create innovative solutions at massive scale. In this role, you will be working with a team to continuously improve our existing datasets as well as pursuing new ones. If you are looking to be part of a team discovering the next frontier of data-as-a-service (DaaS) with a high level of autonomy and opportunity for direct contributions, this might be the role for you. We like our engineers to be thoughtful, quirky, and willing to fearlessly try new things. Failure is embraced at PDL as long as we continue to learn and grow from it. ## What You Get to Do - Contribute to the architecture and improvement of our data acquisition and processing platform, increasing reliability, throughput, and observability - Use and develop web crawling technologies to capture and catalog data on the internet - Build, operate, and evolve large-scale distributed systems that collect, process, and deliver data from across the web - Design and develop backend services that manage distributed job orchestration, data pipelines, and large-scale asynchronous workloads - Structure and model captured data, ensuring high quality and consistency across datasets - Continuously improve the speed, scalability, and fault-tolerance of our ingestion systems - Partner with data product and engineering teams to design and implement new data products powered by the data you help collect, and enhance and improve upon existing products - Learn and apply domain-specific knowledge in web crawling and data acquisition, with mentorship from experienced teammates and access to existing systems The Technical Chops You’ll Need - 7+ years of professional experience building or operating backend or infrastructure systems at scale - Solid programming experience in Python, Go, Rust, or similar, including experience with async / await, coroutines, or concurrency frameworks - Strong grasp of software architecture and backend fundamentals; you can reason clearly about concurrency, scalability, and fault tolerance - Solid understanding of browser rendering pipeline, web application architecture (auth, cookies, http request / response) - Familiarity with network architecture and debugging (HTTP, DNS, proxies, packet capture and analysis) - Solid understanding of distributed systems concepts: parallelism, asynchronous programming, backpressure, and message-driven design - Experience designing or maintaining resilient data ingestion, API integration, or ETL systems - Proficiency with Linux / Unix command-line tools and system resource management - Familiarity with message queues, orchestration, and distributed task systems (Kafka, SQS, Airflow, etc.) - Experience evaluating and monitoring data quality, ensuring consistency, completeness, and reliability across releases People Thrive Here Who Can - Work independently in a fast-paced, remote-first environment, proactively unblocking themselves and collaborating asynchronously - Communicate clearly and thoughtfully in writing (Slack, docs, design proposals) - Write and maintain technical design documents, including pipeline design, schema design, and data flow diagrams - Scope and break down complex projects into deliverable milestones, and communicate progress, risks, and blockers effectively - Balance pragmatism with craftsmanship, shipping reliable systems while continuously improving them Some Nice To Haves - Degree in a quantitative field such as computer science, mathematics, or engineering - Experience as a [Red Teamer](https://en.wikipedia.org/wiki/Red_team) - Experience working on large-scale data ingestion, crawling, or indexing systems - Experience with Apache Spark, Databricks, or other distributed data platforms - Experience with streaming data systems (Kafka, Pub/Sub, Spark Streaming, etc.) - Proficiency with SQL and data warehousing (Snowflake, Redshift, BigQuery, or similar) - Experience with cloud platforms (AWS preferred, GCP or Azure also great) - Understanding of modern data storage and design patterns (parquet, Delta Lake, partitioning, incremental updates) - Knowledge of modern data design and storage patterns (e.g., incremental updating, partitioning and segmentation, rebuilds and backfills) - Experience building and maintaining data pipelines on modern big-data or cloud platforms (Databricks, Spark, or equivalent) Our Benefits - Stock - Competitive Salaries - Unlimited paid time off - Medical, dental, & vision insurance - Health, fitness, and office stipends - The permanent ability to work wherever and however you want Comp: $160K - $200K People Data Labs does not discriminate on the basis of race, sex, color, religion, age, national origin, marital status, disability, veteran status, genetic information, sexual orientation, gender identity or any other reason prohibited by law in provision of employment opportunities and benefits. Qualified Applicants with arrest or conviction records will be considered for Employment in accordance with the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. Personal Privacy Policy for California Residents https://privacy.peopledatalabs.com/policies?name=applicant-privacy-policy ## About People Data Labs ## Company Overview - **One-liner**: People Data Labs (PDL) is a B2B data provider that collects, standardizes, and refreshes compliant data on people, companies, and job postings, delivered via API or data-sharing feeds. - **Entity Type**: Private (Series B, 2021; total funding $55.5M) - **Headquarters**: San Francisco, California, USA (with an office in New York) - **Founded**: 2015 - **Founders**: Not publicly available ## Core Business - **Primary industries**: Data Provisioning, HR Tech, Sales & Marketing Intelligence, Investment Research, Fraud & Identity - **Target customers**: B2B – companies building data-driven products in HR, sales, marketing, investment, and fraud prevention; primarily mid-market to enterprise. - **Mission**: “Build the next” (tagline) – focus on enabling companies to build innovative and compliant people data solutions. ## Products & Services - **Person Data**: Access to over 2 billion profiles with fresh resume data, accurate contact information, and compliant sourcing. Used for enrichment, recruiting platforms, AI models, and custom audiences. - **Company Data**: Differentiated firmographic dataset with unique headcount insights and strong company attributes. - **Job Posting Data** (Beta): Global job posting dataset refreshed daily, sourced directly from company career pages. - **APIs**: Enrich, search, autocomplete, and clean datasets via flexible APIs. Also integrates via Zapier and Make. - **Data Sharing (Feeds)**: License custom data feeds delivered through AWS, GCP, Snowflake, or Databricks (Delta Share). ## Market Standing - **Valuation / Market Cap**: Not disclosed - **Key Metrics**: - Annual Revenue: $38.7M (LinkedIn estimate) - Total Funding: $55.5M across 5 rounds (Seed, Series A x2, Series B) - **Notable Investors / Partners**: Craft Ventures (led Series B), Founders Fund (led Series A), 8VC, Susa Ventures; partners include Snowflake, Databricks, Zapier, Make. - **Growth Signals**: - 78 employees (3.3% YoY growth) - Rated 4.0/5.0 on BuiltIn (based on 38 reviews) - Ranked best-in-class among 60,000+ products on SourceForge - Recent partnership: Fraud.net teams up with PDL to enhance fraud prevention (July 2024) - Compliant with ISO 27001 and SOC 2 Type 2 ## Competitive Advantages - **Data compliance**: ISO 27001 and SOC 2 Type 2 certifications provide a security moat for enterprise clients. - **Scale & freshness**: Over 2 billion profiles; datasets refreshed daily from thousands of compliant sources. - **Developer-friendly access**: Flexible APIs, native integrations with Snowflake and Databricks, plus low-code tools (Zapier, Make). - **Multi-vertical coverage**: Serves HR tech, sales & marketing, investment research, and fraud & identity from a single source of truth. ## Strategic Focus - **Remote-first operations** – fully distributed team with offices in San Francisco and New York. - **Expanding data partnerships** – emphasis on data sharing via cloud platforms (Snowflake, Databricks). - **Product development** – launched Job Posting Data (Beta) and continues to enhance API capabilities. - **Compliance leadership** – doubling down on legal, security, and compliance to attract regulated industries. ## Why Work Here - **Work model**: Remote-first company; employees work from wherever they work best. - **Culture**: “Data lovers who lead with creativity and own their s**t” – described as a team that has fun together and wants to win together. - **Benefits**: - Medical, dental, vision insurance (multiple plan options) - Unlimited PTO - Monthly employee stipends (for internet, hobbies, etc.) - Learning & Development opportunities - **Engineering culture**: Emphasis on solving fascinating data challenges; tech stack includes modern web technologies (full-stack roles). ## Sources 1. [peopledatalabs.com](https://www.peopledatalabs.com/) 2. [peopledatalabs.com/company](https://www.peopledatalabs.com/company) 3. [builtin.com/company/people-data-labs](https://builtin.com/company/people-data-labs) 4. [linkedin.com/company/peopledatalabs](https://www.linkedin.com/company/peopledatalabs) 5. [jobs.ashbyhq.com/people-data-labs](https://jobs.ashbyhq.com/people-data-labs/0de90399-96ce-40dc-ba63-8b954b44ebca) ## Other roles at People Data Labs - [Revenue Operations Specialist](https://feeny.ai/job/revenue-operations-specialist-people-data-labs-remote-k711ex6n7m1q) - [Senior Software Engineer, Full Stack](https://feeny.ai/job/senior-software-engineer-full-stack-people-data-labs-remote-nyaheyh5k34v)