# Senior Site Reliability Engineer (Capacity) - Platform Infrastructure

**Company**: Elastic
**Location**: Spain
**Experience**: senior
**Job type**: full-time
**Salary**: €76,000-€100,300 EUR
**Category**: Engineering
**Industry**: Technology

**Apply**: https://job-boards.greenhouse.io/elastic/jobs/8070615?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_d719b277-05f

## Description

As a Principal Platform Engineer focused on capacity, you will play a crucial role in managing and optimizing compute resources, ensuring that Elastic Cloud Hosted and Serverless workloads can scale seamlessly.

You’ll collaborate closely with control plane and cross-functional platform engineering teams, addressing cloud scaling and resource allocation challenges.

Our diverse team spans EMEA and NASA, tackling complexities of building scalable systems, making a meaningful impact on serving customers.

Responsibilities:

- Assess current and future capacity requirements based on workload demands to ensure seamless scaling of resources.

- Develop and maintain accurate capacity models that predict resource needs and align with business objectives.

- Collaborate with teams to implement proactive measures that prevent capacity shortages and bottlenecks.

- Implement effective strategies for optimizing resource usage across cloud environments.

- Ensure compute resources are utilized efficiently to enhance performance and support seamless scalability.

- Analyze capacity metrics and trends to guide effective resource allocation decisions.

- Develop insightful reporting tools that provide clear visibility into capacity and performance.

- Operate an autoscaling framework that accommodates various customer workloads seamlessly.

- Optimize infrastructure performance across over 60 regions in Elastic Cloud.

- Collaborate with development teams to implement scaling best practices effectively.

Requirements:

- 5+ years with cloud infrastructure and capacity management

- Knowledge of performance monitoring and optimization techniques

- Understanding of cloud scaling challenges and solutions

- Proficiency with incident investigation and troubleshooting processes

- Experience with compute auto-scaling processes and capacity reservations across major CSPs

- Solid software and platform engineering background

- Worked with major cloud service providers and navigated compute capacity scaling issues

Benefits:

- Competitive pay

- Health coverage for you and your family

- Flexible locations and schedules

- Generous number of vacation days

- Matching up to $2000 for financial donations and service

- Up to 40 hours each year to use toward volunteer projects

- Minimum of 16 weeks of parental leave

## Skills

### Required
- cloud infrastructure
- capacity management
- performance monitoring
- optimization techniques
- cloud scaling
- incident investigation
- troubleshooting
- compute auto-scaling
- capacity reservations

---

Source: [Apply at job-boards.greenhouse.io](https://job-boards.greenhouse.io/elastic/jobs/8070615?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
