210K–275K USD / year

Senior Site Reliability Engineer

DevOpsRemote — US
Published on 2026-09-18
These details were extracted automatically from the original listing. They may not be complete or fully up to date — it's worth checking the original job post.

About this role

The Senior Site Reliability Engineer will play a key role in ensuring the reliability and performance of Replit's infrastructure. Responsibilities include designing monitoring solutions, automating operational tasks, and leading incident management efforts to maintain high service availability.

About the company

Replit is an online platform that provides integrated coding environments for developers and learners, enabling users to write, run, and collaborate on code in various programming languages. The company caters to a diverse user base, including educators, students, and professional developers.

The team

Site Reliability Engineering team

Stack

PythonGoKubernetesTerraformAnsiblePrometheusGrafanaDatadog

What you'll do

  • Design and Implement Observability Solutions
  • Drive Automation and Infrastructure as Code
  • Establish SLOs and SLIs
  • Lead incident response efforts
  • Identify and resolve performance bottlenecks

What we're looking for

  • 4-8 years of experience in Site Reliability Engineering or similar roles
  • Strong programming skills in Python or Go
  • Deep understanding of distributed systems
  • Experience with container orchestration platforms
  • Proven track record of implementing monitoring solutions

Nice to have

  • Experience with Google Cloud Platform (GCP)
  • Knowledge of modern observability platforms

Benefits

  • Competitive Salary & Equity
  • 401(k) Program with a 4% match
  • Health, Dental, Vision and Life Insurance
  • Flexible Time Off (FTO) + Holidays
  • Monthly Wellness Stipend
View original job post