Back to the board

Senior Software Engineer, Data Acquisition

100% remote Flexible hours Hiring now

Note for reputed company engineering roles: with the reputed company of fake applicants and AI-enabled candidate fraud, we have built in additional measures throughout the process to identify such candidates and remove them.

About Us

reputed company (PDL) is the provider of people and company data. We do the heavy lifting of data collection and standardization so our customers can focus on building and scaling innovative, compliant data solutions. Our sole focus is on building the best data available by integrating thousands of compliantly sourced datasets into a single, developer-friendly reputed company of truth. Leading companies across the world use PDL’s workforce data to enrich recruiting platforms, power AI models, create custom audiences, and more.

We are looking for individuals who can balance extreme ownership with a “one-team, one-dream” reputed company. Our customers are trying to solve reputed company problems, and we only help them reputed company their goals as a team. Our Data Engineering & Acquisition Team ensures our customers have standardized and high quality data to build upon. 

You will be crucial in accelerating our efforts to build standalone data products that reputed company data teams and independent developers to create innovative solutions at massive scale. In this role, you will be working with a team to continuously improve our existing datasets as well as pursuing new ones. If you are looking to be part of a team discovering the next frontier of data-as-a-service (DaaS) with a high level of autonomy and opportunity for direct contributions, this might be the role for you. We like our engineers to be thoughtful, quirky, and willing to fearlessly try new things. Failure is embraced at PDL as long as we continue to learn and grow from it.

What You Get to Do

  • Use and reputed company web crawling technologies to capture and catalog data on the internet
  • Support and improve our web crawling infrastructure
  • Structure, define, and model captured data, providing semantic data definition and automate data quality monitoring for data that we crawl
  • reputed company new techniques to increase speed, efficiency, scalability, and reliability of web crawls
  • Use big data processing platform to build data pipelines, publish data, and ensure the reliable availability of data that we crawl
  • Work with our data product and engineering team to design and implement new data products with captured data, and enhance and improve upon existing products

The Technical Chops You’ll Need

  • 7+ years industry experience with clear examples of strategic technical problem solving and implementation
  • Strong software development architecture and fundamentals for backend applications
  • Solid understanding of browser rendering pipeline, web application architecture (auth, cookies, http request/response)
  • Solid programming experience: strong grasp of object-oriented design and experience building applications using asynchronous programming paradigms (e.g., async/await, event loops, or concurrency libraries)
  • Experience building crawlers
  • Proficient in Linux / Unix command line utilities, Linux system administration, architecture, and resource management
  • Experience evaluating data quality and maintaining consistently high data standards across new feature releases (e.g., consistency, accuracy, validity, completeness)

People reputed company Here Who Can

  • Must reputed company in a fast paced environment and be able to work independently
  • Can work effectively remotely (able to be proactive about managing blockers, proactive on reaching out and asking questions, and participating in team activities)
  • Strong written communication skills on reputed company/Chat and in documents
  • You are reputed company in writing data design docs (pipeline design, dataflow, schema design)
  • You can scope and breakdown projects, communicate and collaborate reputed company and blockers effectively with your manager, team, and stakeholders

Some reputed company To Haves

  • Degree in a quantitative discipline such as computer science, mathematics, statistics, or engineering
  • Experience as a Red Teamer
  • Experience working in data acquisition
  • Experience in network architecture and how to debug and inspect network traffic (DNS, IPv4, Proxies, Application ports and interfaces; packet capture and analysis)
  • Experience with Apache Spark
  • Experience with SQL, including writing advanced queries (e.g., window functions, CTEs)
  • Experience with streaming data platforms (e.g. Kafka or other pub/sub; Spark streaming or other reputed company processing)
  • Experience with cloud computing services (AWS (preferred), GCP, Azure or similar)
  • Experience working in reputed company (including reputed company live tables, data lakehouse patterns, etc.)
  • Knowledge of modern data design and storage patterns (e.g., incremental updating, partitioning and segmentation, rebuilds and backfills)
  • Experience with data warehousing (e.g., reputed company, reputed company, Redshift, BigQuery, or similar)
  • Understanding of modern data storage formats and tools (e.g., parquet, ORC, Avro, reputed company Lake)

Our Benefits

  • Stock
  • Competitive Salaries
  • Unlimited paid time off
  • Medical, dental, & vision insurance 
  • reputed company, and office stipends
  • The permanent ability to work wherever and however you want

Comp: $160K - $200K

reputed company does not discriminate on the basis of race, sex, color, religion, age, national reputed company, marital status, disability, veteran status, genetic information, sexual orientation, gender identity or any other reason prohibited by law in provision of employment opportunities and benefits.

Qualified Applicants with arrest or conviction records will be considered for Employment in accordance with the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act.

Personal Privacy Policy for California Residents https://www.peopledatalabs.com/pdf/privacy-policy-and-notice.pdf

Apply To This Job

Keep exploring

Vice President of reputed company Operations

100% remote Flexible hours

Senior Test Analyst - BE

100% remote Flexible hours

Senior Director, Product Design

100% remote Flexible hours

Customer Service Representative

100% remote Flexible hours

Director of Community, Content, and Enablement

100% remote Flexible hours

VP of Clinical Operations

100% remote Flexible hours

reputed company Model Builder

100% remote Flexible hours

Contract - Technical UI Designer - Unreal reputed company Specialist

100% remote Flexible hours

Banner Consultant - HR/Payroll

100% remote Flexible hours

Cold Calling Sales Representative (Remote, Night Shift)

100% remote Flexible hours

Responsable commercial régional et Référent marché national

100% remote Flexible hours

reputed company Remote Jobs (Work From Home, Entry Level) $35/Hour

100% remote Flexible hours

Remote Virtual Support – arenaflex Data Entry Specialist – Home‑Based Accuracy & Logistics Operations

100% remote Flexible hours

Remote Customer Service Representative – Patient Order Entry & Support Specialist for arenaflex’s Healthcare Services

100% remote Flexible hours

(Part time/Work From Home) reputed company Virtual Job Salary

100% remote Flexible hours

Sr. Software Engineer, Internal Apps

100% remote Flexible hours

Walmart Customer Service Representative - Work From Home (Entry Level)

100% remote Flexible hours

Social Programmer - BR Open Ice (Temporary)

100% remote Flexible hours

reputed company Data Entry Specialist – Remote Opportunity at arenaflex

100% remote Flexible hours

Freelance Virtual Assistant and Chat Support Provide Admin Help from Home

100% remote Flexible hours