Senior Data Engineer Product

2 weeks ago


California, United States People Data Labs, Inc. Full time

Note for all engineering roles: with the rise of fake applicants and AI-enabled candidate fraud, we have built in additional measures throughout the process to identify such candidates and remove them. About Us People Data Labs (PDL) is the provider of people and company data. We do the heavy lifting of data collection and standardization so our customers can focus on building and scaling innovative, compliant data solutions. Our sole focus is on building the best data available by integrating thousands of compliantly sourced datasets into a single, developer-friendly source of truth. Leading companies across the world use PDL’s workforce data to enrich recruiting platforms, power AI models, create custom audiences, and more. We are looking for individuals who can balance extreme ownership with a “one-team, one-dream” mindset. Our customers are trying to solve complex problems, and we only help them achieve their goals as a team. Our Data Engineering Team is the secret sauce behind all that we do and we are looking for the best of the best. If you are looking to be part of a team discovering the next frontier of data-as-a-service (DaaS) with a high level of autonomy and opportunity for direct contributions, this might be the role for you. We like our engineers to be thoughtful, quirky, and willing to fearlessly try new things. Failure is embraced at PDL as long as we continue to learn and grow from it. What You Get to Do Build infrastructure for ingestion, transformation, and loading an exponentially increasing volume of data from a variety of sources using Spark, SQL, AWS, and Databricks Building an organic entity resolution framework capable of correctly merging hundreds of billions of individual entities into a number of clean, consumable datasets. Developing CI/CD pipelines and anomaly detection systems capable of continuously improving the quality of data we're pushing into production. Dreaming up solutions to largely undefined data engineering and data science problems. The Technical Chops You’ll Need 5-7+ years of industry experience with clear examples of strategic technical problem-solving and implementation Strong software development fundamentals. Experience with Python Expertise with Apache Spark (Java, Scala, and/or Python-based) Experience with SQL Experience building scalable data processing systems (e.g., cleaning, transformation) from the ground up. Experience using developer-oriented data pipeline and workflow orchestration (e.g., Airflow (preferred), dbt, dagster or similar) Knowledge of modern data design and storage patterns (e.g., incremental updating, partitioning and segmentation, rebuilds and backfills) Experience working in Databricks (including delta live tables, data lakehouse patterns, etc.) Experience with cloud computing services (AWS (preferred), GCP, Azure or similar) Experience with data warehousing (e.g., Databricks, Snowflake, Redshift, BigQuery, or similar) Understanding of modern data storage formats and tools (e.g., parquet, ORC, Avro, Delta Lake) People Thrive Here Who Can Balance high ownership and autonomy with a strong ability to collaborate Work effectively remotely (able to be proactive about managing blockers, proactive on reaching out and asking questions, and participating in team activities) Demonstrate strong written communication skills on Slack/Chat and in documents Exhibt experience in writing data design docs (pipeline design, dataflow, schema design) Scope and breakdown projects, communicate and collaborate progress and blockers effectively with your manager, team, and stakeholders Some Nice To Haves Degree in a quantitative discipline such as computer science, mathematics, statistics, or engineering Experience working with entity data (entity resolution / record linkage) Experience working with data acquisition / data integration Expertise with Python and the Python data stack (e.g., numpy, pandas) Experience with streaming platforms (e.g., Kafka) Experience evaluating data quality and maintaining consistently high data standards across new feature releases (e.g., consistency, accuracy, validity, completeness) Our Benefits Stock Competitive Salaries Unlimited paid time off Medical, dental, & vision insurance Health, fitness, and office stipends The permanent ability to work wherever and however you want Comp: $190K - $210K People Data Labs does not discriminate on the basis of race, sex, color, religion, age, national origin, marital status, disability, veteran status, genetic information, sexual orientation, gender identity or any other reason prohibited by law in provision of employment opportunities and benefits. Qualified Applicants with arrest or conviction records will be considered for Employment in accordance with the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. Personal Privacy Policy for California Residents https://www.peopledatalabs.com/pdf/privacy-policy-and-notice.pdf #J-18808-Ljbffr



  • California, United States Windfall Data Inc Full time

    Overview As a Senior Backend Engineer on our platform team at Windfall, you will be building the system for ingesting and processing our customer data. It is the brains of everything our customers interact with. Communication and collaboration are at the heart of Windfall, and you will work closely with our product and our other engineering teams. You will...


  • California, United States People Data Labs Full time

    Note for all engineering roles: With the rise of fake applicants and AI-enabled candidate fraud, we have built in additional measures throughout the process to identify such candidates and remove them. About Us People Data Labs (PDL) is the provider of people and company data. We do the heavy lifting of data collection and standardization so our customers...


  • California, United States People Data Labs, Inc. Full time

    A leading data engineering company in California is seeking an experienced Data Engineer to build and maintain infrastructure for data ingestion and processing. You will work with big data technologies including Spark and Databricks, focusing on enhancing data quality using CI/CD pipelines. The ideal candidate has 5-7+ years of experience and a strong...


  • California, United States People Data Labs, Inc. Full time

    A leading data solutions provider is seeking a skilled Software Engineer to enhance API capabilities. This role involves collaborating with cross-functional teams to deliver innovative data products. The ideal candidate has extensive experience with React.js, Typescript, and Python, and enjoys working in a fast-paced, autonomous environment. Competitive...


  • California, United States People Data Labs, Inc. Full time

    Note for all engineering roles: With the rise of fake applicants and AI-enabled candidate fraud, we have built in additional measures throughout the process to identify such candidates and remove them. About Us People Data Labs (PDL) is the provider of people and company data. We do the heavy lifting of data collection and standardization so our customers...


  • California, United States People Data Labs, Inc. Full time

    Note for all engineering roles: With the rise of fake applicants and AI-enabled candidate fraud, we have built in additional measures throughout the process to identify such candidates and remove them. About Us People Data Labs (PDL) is the provider of people and company data. We do the heavy lifting of data collection and standardization so our customers...


  • California, United States Gusto Full time

    Senior Software Engineer, Data Platform Join Gusto as a Senior Software Engineer on the Data Platform team. Gusto builds payroll, health insurance, 401(k), and HR solutions for small businesses, with teams in Denver, San Francisco, and New York. About Gusto At Gusto, we grow the small business economy. We handle payroll, health insurance, 401(k)s, and HR,...

  • Senior Data Engineer

    4 weeks ago


    California, United States CloudDevs Full time

    Overview OUR ORIGIN STORY: SkySlope started in 2011 as an idea born at the kitchen table of our CEO, with just him and two others. Headquartered in Sacramento, California, we have since grown from our previous three offices and now have close to 150 employees across the United States, supporting nearly 300,000 users in 5,000 offices nationwide and in Canada....

  • Senior Data Engineer

    4 weeks ago


    California, United States CloudDevs Full time

    Overview OUR ORIGIN STORY: SkySlope started in 2011 as an idea born at the kitchen table of our CEO, with just him and two others. Headquartered in Sacramento, California, we have since grown from our previous three offices and now have close to 150 employees across the United States, supporting nearly 300,000 users in 5,000 offices nationwide and in Canada....

  • Data Scientist

    2 weeks ago


    California, United States Gamecompanies Full time

    Data Scientist / Senior Data Scientist - Analytical Data Product, CreatorEvery day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators.At Roblox, we’re building the tools and platform that empower our...