# Senior Software Engineer - Crawler

**Company**: ZoomInfo
**Location**: Remote
**Work arrangement**: remote
**Experience**: senior
**Job type**: full-time
**Salary**: $140,000-$220,000 USD
**Category**: Engineering
**Industry**: Technology

**Apply**: https://job-boards.greenhouse.io/zoominfo/jobs/8687939002?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_7e6148b6-c2a

## Description

We're looking for a Senior Software Engineer to join our Web Data team and help build the next generation of ZoomInfo's web crawling and data extraction infrastructure.

As a Senior Software Engineer, you'll contribute to enterprise-scale crawling and extraction platforms that process massive volumes of web data.

Responsibilities:

- Design and implement components of scalable, fault-tolerant web crawling and extraction pipelines

- Write clean, production-grade code in Java and Python

- Build and operate ETL/ELT pipelines for large-scale data extraction and transformation

- Work with cloud infrastructure on GCP and AWS, primarily on GKE

- Improve observability, reliability, and operational excellence across the systems you contribute to

- Partner with product and data science teams to deliver impactful solutions

- Contribute to code reviews, documentation, and knowledge sharing across the team

- Stay current with evolving web technologies, anti-crawling mechanisms, and AI-powered extraction approaches

Must-Have Qualifications:

- 5+ years of professional software engineering experience building production systems

- Strong CS fundamentals: algorithms, data structures, concurrency, distributed systems

- Proficiency in Java and/or Python

- Track record of owning features end-to-end from design through deployment and operation

- Comfortable making sound architectural decisions at the component level

Data Engineering:

- Hands-on experience with cloud data warehouses such as BigQuery or Snowflake

- Experience designing and operating large-scale ETL/ELT pipelines

- Experience with orchestration tools such as Apache Airflow

- Experience with streaming or event-driven systems such as Apache Kafka

Cloud and Infrastructure:

- Production experience on GCP (preferred) or AWS; multi-cloud exposure is a plus

- Hands-on experience with Kubernetes (GKE/EKS) for distributed workloads

- Familiarity with infrastructure-as-code tooling such as Terraform

Background and Mindset:

- Strong communicator who can explain technical decisions clearly

- Comfortable operating in ambiguity and iterating quickly

- Bias toward action and pragmatic problem solving

- Self-starter who thrives in fast-paced, evolving environments

Benefits: ZoomInfo offers comprehensive benefits, holistic mind, body and lifestyle programs designed for overall well-being.

## Skills

### Required
- Java
- Python
- cloud data warehouses
- ETL/ELT pipelines
- Apache Airflow
- Apache Kafka
- Kubernetes
- Terraform
- distributed systems

### Nice to have
- web crawling at scale
- proxy infrastructure
- rotation strategies
- anti-bot evasion techniques
- SERP extraction
- AI/LLM-based extraction approaches

---

Source: [Apply at job-boards.greenhouse.io](https://job-boards.greenhouse.io/zoominfo/jobs/8687939002?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
