Transform your resume with AI
Leverage AI rewrites and personalized suggestions to create a compelling resume
Start your free trial now →
About Us
At People Data Labs, weβre committed to democratizing access to high-quality B2B data and leading the emerging DaaS economy. We empower developers, engineers, and data scientists to create innovative, compliant data products at scale with our clean, easy-to-use datasets of resume, company, location, and education data consumed through our suite of APIs.
PDL is an innovative, fast-growing, global team backed by world-class investors, including Craft Ventures, Flex Capital, and Founders Fund. We scour the world for people hungry to improve, curious about how things work, and willing to challenge the status quo to build something new and better.
Roles & Responsibilities:
Analyzing data using statistical techniques and tools to identify anomalous data, clean data, and derive meaningful insights and trends.
Ensuring data integrity, accuracy, and completeness throughout the analysis process.
Generate, maintain, and update dashboard reports using our business intelligence tool, highlighting key findings and trends.
Develop and maintain databases, data systems, and data analytics pipelines within database management systems.
Work with stakeholders in Engineering, Product, and Revenue to assist with data-related technical issues and support their data infrastructure and analytics needs.
Ensure the integrity and consistency of database schemas, including managing updates, version control, and documenting schema changes to support data analysis and reporting requirements.
Responsible for assistance and further development of our quality assurance process.
Technical Requirements
3-5+ years industry experience with clear examples of strategic and analytical technical problem solving and implementation
Strong software development and analytics fundamentals
Expertise in with SQL & Python
Experience with Apache Spark or PySpark
Experience with data cleaning and data processing (e.g., cleaning, transformation) using SQL and Python.
Knowledge of modern data design and storage patterns (e.g., incremental updating, partitioning and segmentation, rebuilds and backfills)
Experience working in Databricks (including delta live tables, data lakehouse patterns, etc.)
Experience with cloud computing services (AWS (preferred), GCP, Azure or similar)
Experience with data warehousing (e.g., Databricks, Snowflake, Redshift, BigQuery, or similar)
Understanding of modern data storage formats and tools (e.g., parquet, ORC, Avro, Delta Lake)
Professional Requirements
Must thrive in a fast paced environment and be able to work independently
Can work effectively remotely (able to be proactive about managing blockers, proactive on reaching out and asking questions, and participating in team activities)
Strong written communication skills on Slack/Chat and in documents
You are experienced in writing data design docs (pipeline design, dataflow, schema design)
You can scope and breakdown projects, communicate and collaborate progress and blockers effectively with your manager, team, and stakeholders
Experience collaborating with Product and Engineering teams
Nice To Haves:
Degree in a quantitative discipline such as computer science, mathematics, statistics, or engineering
Experience working with business intelligence tools and dashboard creation
Experience working with data acquisition / data integration
Expertise with Python and the Python data stack (e.g., numpy, pandas, PySpark)
Experience evaluating data quality and maintaining consistently high data standards across new feature releases (e.g., consistency, accuracy, validity, completeness)
Our Benefits
Stock
Competitive Salaries
Unlimited paid time off
Medical, dental, & vision insurance
Health, fitness, and office stipends
The permanent ability to work wherever and however you want
No C2C, 1099, or Contract-to-Hire. Recruiters need not apply.
People Data Labs does not discriminate on the basis of race, sex, color, religion, age, national origin, marital status, disability, veteran status, genetic information, sexual orientation, gender identity or any other reason prohibited by law in provision of employment opportunities and benefits.
No salary data published by company so we estimated salary based on similar jobs related to Design, Python, Education and Cloud jobs that are similar:
$65,000 β $95,000/year
π° 401(k)
π Distributed team
β° Async
π€ Vision insurance
π¦· Dental insurance
π Medical insurance
π Unlimited vacation
π Paid time off
π 4 day workweek
π° 401k matching
π Company retreats
π¬ Coworking budget
π Learning budget
πͺ Free gym membership
π§ Mental wellness budget
π₯ Home office budget
π₯§ Pay in crypto
π₯Έ Pseudonymous
π° Profit sharing
π° Equity compensation
β¬οΈ No whiteboard interview
π No monitoring system
π« No politics at work
π We hire old (and young)
Location
San Francisco, California, United States