New The Skills of Tomorrow: how AI-exposed is every skill in 2026? See the data →
EarnIn

Site Reliability Engineer

EarnIn
Apply →
hybrid senior full-time $189,000 - $232,000 Mountain View, US

First indexed 14 Jul 2026

Description

About the Role

As a Site Reliability Engineer at EarnIn, you will play a crucial role in building and maintaining production systems with high resilience, clarity, and confidence. Your primary focus will be on strengthening infrastructure, optimizing tooling, deepening observability, streamlining incident response, and elevating reliability standards.

Why This Role Exists

EarnIn's community members rely on the company's products for reliability and trust. This role aims to enhance the product experience by reducing noisy alerts, unclear runbooks, fragile deployments, and repeated incidents.

Key Responsibilities

  • Design and improve systems with resilience and graceful degradation in mind, planning for capacity and possible failure modes.
  • Define and measure SLOs and SLIs that reflect customer experience and help teams make better reliability tradeoffs.
  • Utilize observability tools such as Datadog, CloudWatch, logs, metrics, traces, and APM to build signal-heavy, noise-light visibility into production systems.
  • Configure and improve alerting and routing through incident management workflows, ensuring pages are actionable, well-routed, and worth human attention.
  • Participate in incident response from detection and triage through communication, resolution, postmortems, and follow-up.
  • Continuously improve the incident lifecycle, focusing on better detection, clearer runbooks, stronger postmortems, and concrete remediations.
  • Construct or optimize infrastructure, reliability tooling, and automation that eliminate toil and ensure operational consistency.
  • Leverage AI-assisted tools to accelerate coding and documentation, speed up root-cause exploration, and improve infrastructure-as-code workflows and operational tasks.
  • Collaborate with product engineering and platform teams to implement, explain, and support reliability practices.

Requirements

  • Bachelor's or master's degree in Computer Science, Engineering, or a related field, or equivalent industry experience.
  • 3+ years of experience in SRE, Software Engineering, Infrastructure Engineering, or a related role.
  • Hands-on coding experience in Python, Go, or similar production-oriented programming languages.
  • Experience operating production systems and contributing to reliability, observability, incident response, infrastructure, or automation improvements.

Benefits

The base salary range for this full-time position is $189,000 - $232,000 plus equity and benefits.

This listing is enriched and indexed by YubHub. To apply, use the employer's original posting: https://job-boards.greenhouse.io/earnin/jobs/8060043