Description
At Twilio, we're shaping the future of communications. We deliver innovative solutions to hundreds of thousands of businesses and empower millions of developers worldwide to craft personalized customer experiences.
As a Software Engineer L2 in Cloud Infrastructure, you will evolve and maintain fundamental Compute infrastructure, collaborating with a passionate team to enhance system capabilities. You will develop scalable cloud-native environments, handle VM orchestration, manage AWS-ASG Auto Scaling groups of EC2 instances, create hardened base AMIs, and secure container images while maintaining critical OS libraries.
Responsibilities:
- Collaborate with Tech Leaders, Architects, and other Engineers to develop solutions for complex problems in distributed computing and infrastructure management.
- Automate solutions for operational issues, such as monitoring, performance, planning, and disaster response.
- Participate in an on-call rotation to support business-critical infrastructure.
- Ensure high-quality implementation by applying Infrastructure as Code industry standards.
- Author and review design documents, runbooks, and other service documentation, and maintain good records of changes in the systems.
- Apply Agile methodologies to continuously deliver value to customers.
- Act as a point of contact for legacy/new Compute system components.
Qualifications:
- 2+ years of experience in AWS Cloud infrastructure management, preferably backend/infrastructure-focused.
- Strong ASG (Auto Scaling Groups) knowledge to design, implement, and support scalable cloud-native environments.
- Experience with hardened base AMIs and AL23.
- Proficiency in one or more programming languages, such as Java or Python.
- Proficient in shell scripting to streamline repetitive tasks and enhance operational efficiency.
- Ability to work independently with multiple global teams, developing, configuring, deploying, and operating the global Twilio Infrastructure Platform.
- Knowledge of container-based applications/services.
Desired Skills:
- Knowledge of deployment tools and frameworks like infrastructure as code and continuous deployment processes.
- Operational experience in complex distributed systems, including experience with SLO/SLAs towards high availability and reliability goals.
- Exposure to File Integrity Monitoring (FIM) tools, specifically Falco, and awareness of compliance frameworks.
- Knowledge of Kubernetes.
- Experience with Claude AI or similar.
Location: This role will be remote, but is not eligible to be hired in San Francisco, CA, Oakland, CA, San Jose, CA, or the surrounding areas.
Travel: Occasional travel may be required to participate in project or team in-person meetings.
What We Offer: Working at Twilio offers many benefits, including competitive pay, generous time off, ample parental and wellness leave, healthcare, a retirement savings program, and much more.
Compensation: The estimated pay ranges for this role are as follows:
- Based in Colorado, Hawaii, Illinois, Maryland, Massachusetts, Minnesota, Vermont, or Washington D.C.: $116,960.00-$146,200.00
- Based in New York, New Jersey, Washington State, or California (outside of the San Francisco Bay area): $123,760.00-$154,700.00
- Based in the San Francisco Bay area, California: $137,520.00-$171,900.00