Senior Data Engineer
Job Description
About Us
At People Data Labs, we’re committed to democratizing access to high-quality B2B data and leading the emerging DaaS economy. We empower developers, engineers, and data scientists to create innovative, compliant data products at scale with our clean, easy-to-use datasets of resume, company, location, and education data consumed through our suite of APIs.Â
PDL is an innovative, fast-growing, global team backed by world-class investors, including Craft Ventures, Flex Capital, and Founders Fund. We scour the world for people hungry to improve, curious about how things work, and willing to challenge the status quo to build something new and better.
Roles & Responsibilities:
- Build infrastructure for ingestion, transformation, and loading an exponentially increasing volume of data from a variety of sources using Spark, SQL, AWS, and Databricks
- Building an organic entity resolution framework capable of correctly merging hundreds of billions of individual entities into a number of clean, consumable datasets.
- Developing CI/CD pipelines and anomaly detection systems capable of continuously improving the quality of data we're pushing into production.
- Devising solutions to largely-undefined data engineering and data science problems.
- Work with stakeholders in Engineering and Product to assist with data-related technical issues and support their infrastructure needs
About You:
- Strong software development fundamentals.
- Expertise with Python and the Python data stack (e.g., numpy, pandas)
- Expertise with Apache Spark
- Experience with SQL, including writing advanced queries (e.g., window functions, CTEs)
- Experience building scalable data processing systems (e.g., cleaning, transformation)Â from the ground up.
- Experience using developer-oriented data pipeline and workflow orchestration (e.g., Airflow (preferred), Luigi or similar)
- Knowledge of modern data design and storage patterns (e.g., incremental updating, partitioning and segmentation, rebuilds and backfills)
- Must thrive in a fast paced environment and be able to work independently
Nice To Have:
- Bachelor’s degree in a quantitative discipline such as computer science, mathematics, statistics, or engineering or 5+ years industry experience with clear examples of strategic technical problem solving and implementation
- Experience working with entity data (entity resolution / record linkage)
- Experience working with data acquisition
- Experience with cloud computing services (AWS preferred)
- Experience working in Databricks
- Experience with data warehousing (e.g., Redshift, Snowflake)
- Experience with streaming platforms (e.g., Kafka)
- Experience evaluating data quality and maintaining consistently high data standards across new feature releases (e.g., consistency, accuracy, validity, completeness)
- Understanding of modern data storage formats and tools (e.g., parquet, Delta Lake)
- Familiarity with Java or Scala
Our Benefits
- Stock
- Competitive Salaries
- Unlimited paid time off
- Medical, dental, & vision insuranceÂ
- Health, fitness, and office stipends
- The permanent ability to work wherever and however you want
No C2C, 1099, or Contract-to-Hire. Recruiters need not apply.
People Data Labs does not discriminate on the basis of race, sex, color, religion, age, national origin, marital status, disability, veteran status, genetic information, sexual orientation, gender identity or any other reason prohibited by law in provision of employment opportunities and benefits.
#LI-Remote
Date Posted
11/18/2022
Views
5
Similar Jobs
Senior Design Manager (Infrastructure) - Canonical
Views in the last 30 days - 0
Canonical a leading opensource provider seeks a Senior Design Manager to drive innovation in cloud and AI technologies The role offers remote work glo...
View DetailsSenior Product Designer - Org & Security - Typeform
Views in the last 30 days - 0
This job description outlines a role in developing an intelligent contact management system with AI capabilities The position involves designing user ...
View DetailsSenior Business Analyst - Xpansiv
Views in the last 30 days - 0
Xpansiv promotes its role as an energy market innovator with a global platform for environmental commodities The job posting seeks a Business Analyst ...
View DetailsSenior Specialist Senior Accountant Shared Financial Services - Make-A-Wish America
Views in the last 30 days - 0
The text describes Make a Wish Foundations mission to grant childrens wishes and their community efforts It outlines job positions with remotehybrid o...
View DetailsSoftware Engineer Networking Software and Services - xAI
Views in the last 30 days - 0
The text describes xAIs mission to develop AI systems for understanding the universe and advancing human knowledge It outlines a role involving networ...
View DetailsData Scientist - KoBold Metals
Views in the last 30 days - 0
KoBold a leading mineral exploration company using AI seeks a Data Scientist to advance their exploration tech They highlight successful discoveries i...
View Details