\nAbout Us\n\nAt People Data Labs, weโre committed to democratizing access to high-quality B2B data and leading the emerging DaaS economy. We empower developers, engineers, and data scientists to create innovative, compliant data products at scale with our clean, easy-to-use datasets of resume, company, location, and education data consumed through our suite of APIs. \n\nPDL is an innovative, fast-growing, global team backed by world-class investors, including Craft Ventures, Flex Capital, and Founders Fund. We scour the world for people hungry to improve, curious about how things work, and willing to challenge the status quo to build something new and better.\n\nRoles & Responsibilities:\n\n\n* Analyzing data using statistical techniques and tools to identify anomalous data, clean data, and derive meaningful insights and trends. \n\n* Ensuring data integrity, accuracy, and completeness throughout the analysis process.\n\n* Generate, maintain, and update dashboard reports using our business intelligence tool, highlighting key findings and trends.\n\n* Develop and maintain databases, data systems, and data analytics pipelines within database management systems.\n\n* Work with stakeholders in Engineering, Product, and Revenue to assist with data-related technical issues and support their data infrastructure and analytics needs.\n\n* Ensure the integrity and consistency of database schemas, including managing updates, version control, and documenting schema changes to support data analysis and reporting requirements.\n\n* Responsible for assistance and further development of our quality assurance process.\n\n\n\n\nTechnical Requirements\n\n\n* 3-5+ years industry experience with clear examples of strategic and analytical technical problem solving and implementation\n\n* Strong software development and analytics fundamentals\n\n* Expertise in with SQL & Python\n\n* Experience with Apache Spark or PySpark\n\n* Experience with data cleaning and data processing (e.g., cleaning, transformation) using SQL and Python.\n\n* Knowledge of modern data design and storage patterns (e.g., incremental updating, partitioning and segmentation, rebuilds and backfills)\n\n* Experience working in Databricks (including delta live tables, data lakehouse patterns, etc.)\n\n* Experience with cloud computing services (AWS (preferred), GCP, Azure or similar)\n\n* Experience with data warehousing (e.g., Databricks, Snowflake, Redshift, BigQuery, or similar)\n\n* Understanding of modern data storage formats and tools (e.g., parquet, ORC, Avro, Delta Lake)\n\n\n\n\nProfessional Requirements\n\n\n* Must thrive in a fast paced environment and be able to work independently\n\n* Can work effectively remotely (able to be proactive about managing blockers, proactive on reaching out and asking questions, and participating in team activities)\n\n* Strong written communication skills on Slack/Chat and in documents\n\n* You are experienced in writing data design docs (pipeline design, dataflow, schema design)\n\n* You can scope and breakdown projects, communicate and collaborate progress and blockers effectively with your manager, team, and stakeholders\n\n* Experience collaborating with Product and Engineering teams\n\n\n\n\nNice To Haves:\n\n\n* Degree in a quantitative discipline such as computer science, mathematics, statistics, or engineering\n\n* Experience working with business intelligence tools and dashboard creation\n\n* Experience working with data acquisition / data integration\n\n* Expertise with Python and the Python data stack (e.g., numpy, pandas, PySpark)\n\n* Experience evaluating data quality and maintaining consistently high data standards across new feature releases (e.g., consistency, accuracy, validity, completeness)\n\n\n\n\nOur Benefits\n\n\n* Stock\n\n* Competitive Salaries\n\n* Unlimited paid time off\n\n* Medical, dental, & vision insurance \n\n* Health, fitness, and office stipends\n\n* The permanent ability to work wherever and however you want\n\n\n\n\nNo C2C, 1099, or Contract-to-Hire. Recruiters need not apply.\n\nPeople Data Labs does not discriminate on the basis of race, sex, color, religion, age, national origin, marital status, disability, veteran status, genetic information, sexual orientation, gender identity or any other reason prohibited by law in provision of employment opportunities and benefits. \n\n#Salary and compensation\n
No salary data published by company so we estimated salary based on similar jobs related to Design, Python, Education and Cloud jobs that are similar:\n\n
$65,000 — $95,000/year\n
\n\n#Benefits\n
๐ฐ 401(k)\n\n๐ Distributed team\n\nโฐ Async\n\n๐ค Vision insurance\n\n๐ฆท Dental insurance\n\n๐ Medical insurance\n\n๐ Unlimited vacation\n\n๐ Paid time off\n\n๐ 4 day workweek\n\n๐ฐ 401k matching\n\n๐ Company retreats\n\n๐ฌ Coworking budget\n\n๐ Learning budget\n\n๐ช Free gym membership\n\n๐ง Mental wellness budget\n\n๐ฅ Home office budget\n\n๐ฅง Pay in crypto\n\n๐ฅธ Pseudonymous\n\n๐ฐ Profit sharing\n\n๐ฐ Equity compensation\n\nโฌ๏ธ No whiteboard interview\n\n๐ No monitoring system\n\n๐ซ No politics at work\n\n๐ We hire old (and young)\n\n
\n\n#Location\nSan Francisco, California, United States
๐ Please reference you found the job on Remote OK, this helps us get more companies to post here, thanks!
When applying for jobs, you should NEVER have to pay to apply. You should also NEVER have to pay to buy equipment which they then pay you back for later. Also never pay for trainings you have to do. Those are scams! NEVER PAY FOR ANYTHING! Posts that link to pages with "how to work online" are also scams. Don't use them or pay for them. Also always verify you're actually talking to the company in the job post and not an imposter. A good idea is to check the domain name for the site/email and see if it's the actual company's main domain name. Scams in remote work are rampant, be careful! Read more to avoid scams. When clicking on the button to apply above, you will leave Remote OK and go to the job application page for that company outside this site. Remote OK accepts no liability or responsibility as a consequence of any reliance upon information on there (external sites) or here.
People Data Labs is hiring a Remote Senior Data Engineer
\nAbout Us\n\nAt People Data Labs, weโre committed to democratizing access to high-quality B2B data and leading the emerging DaaS economy. We empower developers, engineers, and data scientists to create innovative, compliant data products at scale with our clean, easy-to-use datasets of resume, company, location, and education data consumed through our suite of APIs. \n\nPDL is an innovative, fast-growing, global team backed by world-class investors, including Craft Ventures, Flex Capital, and Founders Fund. We scour the world for people hungry to improve, curious about how things work, and willing to challenge the status quo to build something new and better.\n\nRoles & Responsibilities:\n\n\n* Build infrastructure for ingestion, transformation, and loading an exponentially increasing volume of data from a variety of sources using Spark, SQL, AWS, and Databricks\n\n* Building an organic entity resolution framework capable of correctly merging hundreds of billions of individual entities into a number of clean, consumable datasets.\n\n* Developing CI/CD pipelines and anomaly detection systems capable of continuously improving the quality of data we're pushing into production.\n\n* Devising solutions to largely-undefined data engineering and data science problems.\n\n* Work with stakeholders in Engineering and Product to assist with data-related technical issues and support their infrastructure needs\n\n\n\n\nTechnical Requirements\n\n\n* 5-7+ years industry experience with clear examples of strategic technical problem solving and implementation\n\n* Strong software development fundamentals.\n\n* Experience withPython Expertise with Apache Spark (Java, Scala, and/or Python-based)\n\n* Experience with SQL\n\n* Experience building scalable data processing systems (e.g., cleaning, transformation) from the ground up.\n\n* Experience using developer-oriented data pipeline and workflow orchestration (e.g., Airflow (preferred), dbt, dagster or similar)\n\n* Knowledge of modern data design and storage patterns (e.g., incremental updating, partitioning and segmentation, rebuilds and backfills)\n\n* Experience working in Databricks (including delta live tables, data lakehouse patterns, etc.)\n\n* Experience with cloud computing services (AWS (preferred), GCP, Azure or similar)\n\n* Experience with data warehousing (e.g., Databricks, Snowflake, Redshift, BigQuery, or similar)\n\n* Understanding of modern data storage formats and tools (e.g., parquet, ORC, Avro, Delta Lake)\n\n\n\n\nProfessional Requirements\n\n\n* Must thrive in a fast paced environment and be able to work independently\n\n* Can work effectively remotely (able to be proactive about managing blockers, proactive on reaching out and asking questions, and participating in team activities)\n\n* Strong written communication skills on Slack/Chat and in documents\n\n* You are experienced in writing data design docs (pipeline design, dataflow, schema design)\n\n* You can scope and breakdown projects, communicate and collaborate progress and blockers effectively with your manager, team, and stakeholders\n\n\n\n\nNice To Haves:\n\n\n* Degree in a quantitative discipline such as computer science, mathematics, statistics, or engineering\n\n* Experience working with entity data (entity resolution / record linkage)\n\n* Experience working with data acquisition / data integration\n\n* Expertise with Python and the Python data stack (e.g., numpy, pandas)\n\n* Experience with streaming platforms (e.g., Kafka)\n\n* Experience evaluating data quality and maintaining consistently high data standards across new feature releases (e.g., consistency, accuracy, validity, completeness)\n\n\n\n\nOur Benefits\n\n\n* Stock\n\n* Competitive Salaries\n\n* Unlimited paid time off\n\n* Medical, dental, & vision insurance \n\n* Health, fitness, and office stipends\n\n* The permanent ability to work wherever and however you want\n\n\n\n\nNo C2C, 1099, or Contract-to-Hire. Recruiters need not apply.\n\nPeople Data Labs does not discriminate on the basis of race, sex, color, religion, age, national origin, marital status, disability, veteran status, genetic information, sexual orientation, gender identity or any other reason prohibited by law in provision of employment opportunities and benefits. \n\n#Salary and compensation\n
No salary data published by company so we estimated salary based on similar jobs related to Design, Python, Education, Cloud, Senior and Engineer jobs that are similar:\n\n
$65,000 — $105,000/year\n
\n\n#Benefits\n
๐ฐ 401(k)\n\n๐ Distributed team\n\nโฐ Async\n\n๐ค Vision insurance\n\n๐ฆท Dental insurance\n\n๐ Medical insurance\n\n๐ Unlimited vacation\n\n๐ Paid time off\n\n๐ 4 day workweek\n\n๐ฐ 401k matching\n\n๐ Company retreats\n\n๐ฌ Coworking budget\n\n๐ Learning budget\n\n๐ช Free gym membership\n\n๐ง Mental wellness budget\n\n๐ฅ Home office budget\n\n๐ฅง Pay in crypto\n\n๐ฅธ Pseudonymous\n\n๐ฐ Profit sharing\n\n๐ฐ Equity compensation\n\nโฌ๏ธ No whiteboard interview\n\n๐ No monitoring system\n\n๐ซ No politics at work\n\n๐ We hire old (and young)\n\n
\n\n#Location\nSan Francisco, California, United States
๐ Please reference you found the job on Remote OK, this helps us get more companies to post here, thanks!
When applying for jobs, you should NEVER have to pay to apply. You should also NEVER have to pay to buy equipment which they then pay you back for later. Also never pay for trainings you have to do. Those are scams! NEVER PAY FOR ANYTHING! Posts that link to pages with "how to work online" are also scams. Don't use them or pay for them. Also always verify you're actually talking to the company in the job post and not an imposter. A good idea is to check the domain name for the site/email and see if it's the actual company's main domain name. Scams in remote work are rampant, be careful! Read more to avoid scams. When clicking on the button to apply above, you will leave Remote OK and go to the job application page for that company outside this site. Remote OK accepts no liability or responsibility as a consequence of any reliance upon information on there (external sites) or here.