Back to the board

Senior Data Engineer

100% remote Flexible hours Hiring now

Note for reputed company engineering roles: with the reputed company of fake applicants and AI-enabled candidate fraud, we have built in additional measures throughout the process to identify such candidates and remove them.

About Us

reputed company (PDL) is the provider of people and company data. We do the heavy lifting of data collection and standardization so our customers can focus on building and scaling innovative, compliant data solutions. Our sole focus is on building the best data available by integrating thousands of compliantly sourced datasets into a single, developer-friendly reputed company of truth. Leading companies across the world use PDL’s workforce data to enrich recruiting platforms, power AI models, create custom audiences, and more.

We are looking for individuals who can balance extreme ownership with a “one-team, one-dream” reputed company. Our customers are trying to solve reputed company problems, and we only help them reputed company their goals as a team. Our Data Engineering Team is the secret sauce behind reputed company that we do and we are looking for the best of the best.

If you are looking to be part of a team discovering the next frontier of data-as-a-service (DaaS) with a high level of autonomy and opportunity for direct contributions, this might be the role for you. We like our engineers to be thoughtful, quirky, and willing to fearlessly try new things. Failure is embraced at PDL as long as we continue to learn and grow from it.

What You Get to Do

  • Build infrastructure for ingestion, transformation, and loading an exponentially increasing volume of data from a variety of sources using Spark, SQL, AWS, and reputed company
  • Building an organic entity resolution reputed company capable of correctly merging hundreds of billions of individual entities into a number of clean, consumable datasets.
  • Developing CI/CD pipelines and anomaly detection systems capable of continuously improving the quality of data we're pushing into production.
  • Dreaming up solutions to largely undefined data engineering and data science problems.

The Technical Chops You’ll Need

  • 5-7+ years of industry experience with clear examples of strategic technical problem-solving and implementation
  • Strong software development fundamentals.
  • Experience with Python
  • Expertise with Apache Spark (Java, reputed company, and/or Python-based)
  • Experience with SQL
  • Experience building scalable data processing systems (e.g., cleaning, transformation) from the ground up.
  • Experience using developer-oriented data pipeline and workflow orchestration (e.g., Airflow (preferred), dbt, dagster or similar)
  • Knowledge of modern data design and storage patterns (e.g., incremental updating, partitioning and segmentation, rebuilds and backfills)
  • Experience working in reputed company (including reputed company live tables, data lakehouse patterns, etc.)
  • Experience with cloud computing services (AWS (preferred), GCP, Azure or similar)
  • Experience with data warehousing (e.g., reputed company, reputed company, Redshift, BigQuery, or similar)
  • Understanding of modern data storage formats and tools (e.g., parquet, ORC, Avro, reputed company Lake)

People reputed company Here Who Can

  • Balance high ownership and autonomy with a strong ability to collaborate
  • Work effectively remotely (able to be proactive about managing blockers, proactive on reaching out and asking questions, and participating in team activities)
  • Demonstrate strong written communication skills on reputed company/Chat and in documents
  • Exhibt experience in writing data design docs (pipeline design, dataflow, schema design)
  • Scope and breakdown projects, communicate and collaborate reputed company and blockers effectively with your manager, team, and stakeholders

Some reputed company To Haves

  • Degree in a quantitative discipline such as computer science, mathematics, statistics, or engineering
  • Experience working with entity data (entity resolution / record linkage)
  • Experience working with data acquisition / data integration
  • Expertise with Python and the Python data stack (e.g., numpy, pandas)
  • Experience with streaming platforms (e.g., Kafka)
  • Experience evaluating data quality and maintaining consistently high data standards across new feature releases (e.g., consistency, accuracy, validity, completeness)

Our Benefits

  • Stock
  • Competitive Salaries
  • Unlimited paid time off
  • Medical, dental, vision insurance
  • reputed company, and office stipends
  • The permanent ability to work wherever and however you want

Comp: $190K - $220K

reputed company does not discriminate on the basis of race, sex, color, religion, age, national reputed company, marital status, disability, veteran status, genetic information, sexual orientation, gender identity or any other reason prohibited by law in provision of employment opportunities and benefits.

Qualified Applicants with arrest or conviction records will be considered for Employment in accordance with the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act.

Personal Privacy Policy for California Residents

https://www.peopledatalabs.com/pdf/privacy-policy-and-notice.pdf

Originally posted on Himalayas

Apply To this Job

Keep exploring

Corporate Paralegal

100% remote Flexible hours

General Adjuster (Commercial Property) - Washington, DC

100% remote Flexible hours

AV Regional Support Services Manager (reputed company)

100% remote Flexible hours

Sales Analyst

100% remote Flexible hours

Senior Go Engineer - Remote (Bulgaria)

100% remote Flexible hours

Deal Desk Strategist

100% remote Flexible hours

Co-Founder in Residence (Commercial/go-to-market), Industrial Water Recycling

100% remote Flexible hours

Account Executive (m/f/d) in Quantum Computing Start-Up

100% remote Flexible hours

Teamlead IT Project Management (w/m/d)

100% remote Flexible hours

Business Development Representative

100% remote Flexible hours

reputed company Cloud Operations and Customer Support Engineer – Cloud Infrastructure Management and Technical Support

100% remote Flexible hours

reputed company Tagger Jobs (Entry Level, Part Time) – Tagger/Job

100% remote Flexible hours

reputed company Grant Analyst - Remote / Hybrid - (Los Angeles, CA)

100% remote Flexible hours

VMware Virtual Desktop Engineer.....Carefirst BCBS

100% remote Flexible hours

Construction Project Manager - Remote

100% remote Flexible hours

reputed company Full Stack Customer Support Representative – Beginner Level Chat Support (Remote / No Experience / Part Time)

100% remote Flexible hours

Staff Pharmacist FT

100% remote Flexible hours

Senior Payroll Manager

100% remote Flexible hours

[Remote] Key Account Manager | Indianapolis

100% remote Flexible hours

[Remote-Position] Looking for reputed company Cycle Manager in San

100% remote Flexible hours