# Staff TDI Site Reliability Engineer, Okta Federal

**Company**: Okta Federal, Inc.
**Location**: Washington, DC
**Work arrangement**: onsite
**Experience**: staff
**Job type**: full-time
**Salary**: $174,000-$239,000 USD
**Category**: Engineering
**Industry**: Technology

**Apply**: https://job-boards.greenhouse.io/okta/jobs/8073066?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_6418b3bc-7a7

## Description

Okta Federal, Inc. is looking for an experienced Staff TDI Site Reliability Engineer to help build, improve, and maintain cloud platform services that support the most sensitive national security missions.

The Technology, Data & Intelligence (TDI) team delivers systems, tools, and services that power internal operations across the company. As a Staff Site Reliability Engineer, you will design and implement complex cloud-based engineering enablement systems, ensuring compliance with strict government requirements.

**Responsibilities**

- Operate and maintain enterprise-grade solutions within air-gapped environments.

- Build, run, and monitor development tools, pipelines, and infrastructure with a security-first mindset.

- Operate autonomously within secure facilities.

- Maintain Service Level Objectives (SLOs) and Service Level Indicators (SLIs) for workloads with no dependency on external monitoring or SaaS tooling.

- Own runbooks and incident response procedures tailored to limited external escalation paths.

- Participate in Plan of Action & Milestones (POA&M) remediation and support annual/recurring Authority to Operate activities.

- Support and run mission-critical services depended on by product teams.

- Deliver excellent internal customer service and advocate for SRE and DevOps practices across teams.

- Build and operate CI/CD pipelines that function without internet connectivity.

**Requirements**

- 7+ years of experience as an SRE, DevOps Engineer, Cloud Automation Engineer, or Systems Engineer with a track record of delivering complex infrastructure projects at scale.

- Experience with container orchestration and runtime environments, including EKS, ECS Fargate, and general container usage.

- Proficient in infrastructure automation using Terraform and developing automation tools with Python, while leveraging secure software development practices.

- Experience with monitoring tools, especially Splunk, CloudWatch, and the Grafana stack.

- Experience with general networking concepts, such as BGP and IPsec management, and has leveraged AWS networking services, including VPCs, TGWs, and VPC endpoints.

- Security Clearance: Active U.S. TS/SCI with polygraph.

- The selected candidate may be subject to drug testing to the extent required by U.S. Government contracts.

**Additional Requirements**

- U.S. soil status: The employee must be on U.S. soil.

- U.S. Security Clearance status: The employee must be able to obtain and maintain a U.S. security clearance.

**Benefits**

- Annual base salary range: $174,000-$239,000 USD

- Equity (where applicable)

- Bonus

- Health, dental, and vision insurance

- 401(k)

- Flexible spending account

- Paid leave (including PTO and parental leave)

## Skills

### Required
- EKS
- ECS Fargate
- Terraform
- Python
- Splunk
- CloudWatch
- Grafana
- BGP
- IPsec
- AWS networking services
- VPCs
- TGWs
- VPC endpoints

---

Source: [Apply at job-boards.greenhouse.io](https://job-boards.greenhouse.io/okta/jobs/8073066?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
