# Senior DevOps Engineer

**Company**: ZoomInfo
**Location**: Bengaluru, Karnataka
**Work arrangement**: hybrid
**Experience**: senior
**Job type**: full-time
**Category**: Engineering
**Industry**: Technology

**Apply**: https://job-boards.greenhouse.io/zoominfo/jobs/8562647002?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_dd24dc09-005

## Description

ZoomInfo is where careers accelerate. We move fast, think boldly, and empower you to do the best work of your life.

The Infrastructure Engineering group (IE) is responsible for innovation in infrastructure and automation for ZoomInfo Engineering. We're using that to build world-class, multi-provider cloud infrastructures, based on the best technologies available.

We are looking for a Senior Infrastructure Engineer to lead the evolution of our cloud ecosystem. In this role, you won't just be 'keeping the lights on',you will be architecting the foundation that powers our next generation of AI-driven products.

**Responsibilities:**

- Cloud Architecture & Orchestration: Design and manage multi-cloud environments (AWS/GCP) using Terraform to ensure infrastructure is versioned, reproducible, and scalable.

- Kubernetes Mastery: Oversee production-grade Kubernetes clusters, focusing on cluster health, resource optimization, and seamless application deployment.

- AI & Data Infrastructure: Build and maintain the specialized infrastructure required for AI workloads.

- Database Management: Maintain a diverse data layer, including Vector search databases (for AI retrieval), PostgreSQL, MySQL, and MongoDB.

- Observability & Reliability: Implement deep-stack monitoring and alerting using Datadog and Prometheus to ensure proactive issue detection and resolution.

- Automation-First Mindset: Maintain and evolve an active codebase in Python, Go, or Bash to automate repetitive tasks.

- Networking & Traffic: Manage complex cloud networking topologies, including VPCs, Load Balancing, Service Meshes, and Caching layers (e.g., Redis) to minimize latency.

- Incident Response: Lead the debugging of complex, distributed systems issues, performing root cause analysis to prevent recurrence, as part of an on-call rotation.

**Requirements:**

- 7+ years of experience in Infrastructure, DevOps, or Site Reliability Engineering, with at least 5 years focused on cloud-native environments.

- Solid understanding of Linux administration and performance tuning.

- Expert-level experience with Terraform (or OpenTofu) and managing state at scale.

- Proven track record of managing Kubernetes in a production environment (EKS, GKE, or self-managed).

- Strong proficiency in Python or Go.

- Hands-on experience managing relational and non-relational databases.

- Experience leveraging and securing PaaS offerings to speed up development cycles.

- Practical experience using LLMs (GitHub Copilot, ChatGPT, Claude) to increase personal and team productivity.

- Demonstrated ability to learn new technologies quickly and independently.

- Strong technical, organizational, and interpersonal skills.

- Strong written and verbal communication skills.

## Skills

### Required
- Terraform
- Kubernetes
- Python
- Go
- Linux
- Datadog
- Prometheus
- AWS
- GCP
- PostgreSQL
- MySQL
- MongoDB
- Redis

---

Source: [Apply at job-boards.greenhouse.io](https://job-boards.greenhouse.io/zoominfo/jobs/8562647002?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
