Description
Electronic Arts is looking for a Site Reliability Engineer (SRE) to join our GameKit Operations team. You will be part of a newly formed SRE function and help shape the future of how EA builds and operates its development platforms and services.
Main Responsibilities
- Build scalable monitoring and observability systems using Prometheus/Grafana, Datadog, ELK, or similar.
- Build infrastructure and tooling using technologies like Terraform, Ansible, AWS CloudFormation, and CI/CD pipelines (GitLab CI/CD).
- Automate operational processes using Python and Bash to reduce manual toil and improve deployment reliability.
- Operate and improve containerized applications using Kubernetes platforms (EKS, AKS, GKE).
- Contribute to incident response processes and post-mortems, helping teams learn and improve from every incident.
Requirements
- Experience operating cloud platforms, especially AWS and Azure.
- Experience in monitoring, observability, and incident response at scale.
- Hands-on experience with Infrastructure-as-Code and automation.
- Desire to improve processes and team capabilities.
- Comfortable working in dynamic environments and solving problems collaboratively.
- 5+ years of experience building SRE practices from the ground up.
- Experience in on-call rotations or reliability-focused projects.
- Mentored junior engineers and influenced engineering culture through documentation and collaboration.
Benefits EA supports a balanced life with a comprehensive benefits program that emphasizes physical, emotional, financial, and professional well-being, and community wellness. Our packages are customized by location and can include medical insurance, mental health support, retirement savings, paid time off, family leave, free games, and more.
This listing is enriched and indexed by YubHub. To apply, use the employer's original posting:
https://jobs.ea.com/ko_KR/careers/JobDetail/Site-Reliability-Engineer-12-months-contract/214926