People Data LabsPL

Senior Data Engineer

People Data Labs is a B2B data provider specializing in unique person profiles to enhance business applications and services.

People Data Labs

Employee count: 51-200

Salary: 190k-220k USD

United States only

Note for all engineering roles: with the rise of fake applicants and AI-enabled candidate fraud, we have built in additional measures throughout the process to identify such candidates and remove them.

About Us

People Data Labs (PDL) is the provider of people and company data. We do the heavy lifting of data collection and standardization so our customers can focus on building and scaling innovative, compliant data solutions. Our sole focus is on building the best data available by integrating thousands of compliantly sourced datasets into a single, developer-friendly source of truth. Leading companies across the world use PDL’s workforce data to enrich recruiting platforms, power AI models, create custom audiences, and more.

We are looking for individuals who can balance extreme ownership with a “one-team, one-dream” mindset. Our customers are trying to solve complex problems, and we only help them achieve their goals as a team. Our Data Engineering Team is the secret sauce behind all that we do and we are looking for the best of the best.

If you are looking to be part of a team discovering the next frontier of data-as-a-service (DaaS) with a high level of autonomy and opportunity for direct contributions, this might be the role for you. We like our engineers to be thoughtful, quirky, and willing to fearlessly try new things. Failure is embraced at PDL as long as we continue to learn and grow from it.

What You Get to Do

  • Build infrastructure for ingestion, transformation, and loading an exponentially increasing volume of data from a variety of sources using Spark, SQL, AWS, and Databricks
  • Building an organic entity resolution framework capable of correctly merging hundreds of billions of individual entities into a number of clean, consumable datasets.
  • Developing CI/CD pipelines and anomaly detection systems capable of continuously improving the quality of data we're pushing into production.
  • Dreaming up solutions to largely undefined data engineering and data science problems.

The Technical Chops You’ll Need

  • 5-7+ years of industry experience with clear examples of strategic technical problem-solving and implementation
  • Strong software development fundamentals.
  • Experience with Python
  • Expertise with Apache Spark (Java, Scala, and/or Python-based)
  • Experience with SQL
  • Experience building scalable data processing systems (e.g., cleaning, transformation) from the ground up.
  • Experience using developer-oriented data pipeline and workflow orchestration (e.g., Airflow (preferred), dbt, dagster or similar)
  • Knowledge of modern data design and storage patterns (e.g., incremental updating, partitioning and segmentation, rebuilds and backfills)
  • Experience working in Databricks (including delta live tables, data lakehouse patterns, etc.)
  • Experience with cloud computing services (AWS (preferred), GCP, Azure or similar)
  • Experience with data warehousing (e.g., Databricks, Snowflake, Redshift, BigQuery, or similar)
  • Understanding of modern data storage formats and tools (e.g., parquet, ORC, Avro, Delta Lake)

People Thrive Here Who Can

  • Balance high ownership and autonomy with a strong ability to collaborate
  • Work effectively remotely (able to be proactive about managing blockers, proactive on reaching out and asking questions, and participating in team activities)
  • Demonstrate strong written communication skills on Slack/Chat and in documents
  • Exhibt experience in writing data design docs (pipeline design, dataflow, schema design)
  • Scope and breakdown projects, communicate and collaborate progress and blockers effectively with your manager, team, and stakeholders

Some Nice To Haves

  • Degree in a quantitative discipline such as computer science, mathematics, statistics, or engineering
  • Experience working with entity data (entity resolution / record linkage)
  • Experience working with data acquisition / data integration
  • Expertise with Python and the Python data stack (e.g., numpy, pandas)
  • Experience with streaming platforms (e.g., Kafka)
  • Experience evaluating data quality and maintaining consistently high data standards across new feature releases (e.g., consistency, accuracy, validity, completeness)

Our Benefits

  • Stock
  • Competitive Salaries
  • Unlimited paid time off
  • Medical, dental, & vision insurance
  • Health, fitness, and office stipends
  • The permanent ability to work wherever and however you want

Comp: $190K - $220K

People Data Labs does not discriminate on the basis of race, sex, color, religion, age, national origin, marital status, disability, veteran status, genetic information, sexual orientation, gender identity or any other reason prohibited by law in provision of employment opportunities and benefits.

Qualified Applicants with arrest or conviction records will be considered for Employment in accordance with the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act.

Personal Privacy Policy for California Residents

https://www.peopledatalabs.com/pdf/privacy-policy-and-notice.pdf

About the job

Apply before

Posted on

Job type

Full Time

Experience level

Senior

Salary

Salary: 190k-220k USD

Location requirements

Hiring timezones

United States +/- 0 hours

About People Data Labs

Learn more about People Data Labs and their company culture.

View company profile

People Data Labs is a B2B data provider that specializes in building comprehensive datasets of unique person profiles. The company leverages a dataset comprising 1.5 billion unique person profiles, which can be utilized for various purposes, including enriching existing person profiles, powering predictive modeling, and facilitating advanced data analysis. This capability makes People Data Labs an essential partner for businesses looking to enhance their products and services with high-quality data.

Through its extensive data offerings, People Data Labs enables organizations to create more efficient and effective solutions in their operations. The company serves industry-leading platforms that rely on accurate and detailed person data to drive their applications. Whether for boosting marketing efforts, enhancing user experience, or supporting data-driven decision-making, People Data Labs provides the foundational data that businesses need to succeed in today's data-centric landscape.

Employee benefits

Learn about the employee benefits and perks provided at People Data Labs.

View benefits

Unlimited pto

Unlimited paid time off.

Equity

Stock options or equity grants.

Medical, dental, and vision

Medical, dental, and vision insurance coverage.

Monthly health & wellness stipend

A monthly allowance for health and wellness expenses.

View People Data Labs's employee benefits
Claim this profilePeople Data Labs logoPL

People Data Labs

View company profile

Similar remote jobs

Here are other jobs you might want to apply for.

View all remote jobs

5 remote jobs at People Data Labs

Explore the variety of open remote roles at People Data Labs, offering flexible work options across multiple disciplines and skill levels.

View all jobs at People Data Labs
People Data Labs logoPL
United States only

Senior Software Engineer, Data Acquisition

People Data Labs

Employee count: 51-200

Salary: 160k-200k USD

People Data Labs logoPL
United States only

Senior Software Engineer, Platform

People Data Labs

Employee count: 51-200

Salary: 160k-180k USD

People Data Labs logoPL
United States only

Customer Success Manager

People Data Labs

Employee count: 51-200

Salary: 100k-130k USD

Remote companies like People Data Labs

Find your next opportunity by exploring profiles of companies that are similar to People Data Labs. Compare culture, benefits, and job openings on Himalayas.

View all companies

DemandScience simplifies B2B marketing with technology and expertise, empowering marketers to successfully grow their business.

In-database machine learning for time-series & anomaly detection.

The enterprise data science management platform trusted by over 20% of the Fortune 100.

Dataplor is a global location intelligence provider, offering accurate and dynamically updated Point of Interest (POI) data for businesses to make informed decisions and identify growth opportunities worldwide.

GetOnData is a premier data service provider committed to empowering businesses with the tools and insights they need to thrive in a data-driven world.

Find your dream job

Sign up now and join over 85,000 remote workers who receive personalized job alerts, curated job matches, and more for free!

Sign up
Himalayas profile for an example user named Frankie Sullivan
People Data Labs hiring Senior Data Engineer • Remote (Work from Home) | Himalayas