Skip to main content
ResumeKart
← Back to Jobs

Senior Site Reliability Engineer

ReplitRemote - US
Full-timeSenior
$210k - $275k
per year
👁️ 0 views📝 0 applicationsPosted 9/18/2026Expires 10/22/2026
Tailor Resume for This JobCheck ATS Score

Get alerts for roles like this

More Senior Site Reliability Engineer roles in Remote - US — straight to your inbox. No account needed.

Applying to this role? Tailor your résumé to this job description in one click, then download it clean — no watermark, no subscription.

Job Description

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation.

About the role: Join our Site Reliability Engineering team and help ensure the reliability, scalability, and performance of Replit's infrastructure that serves millions of developers worldwide.

As a Site Reliability Engineer, you will bridge the gap between development and operations, implementing automation and establishing best practices that enable our platform to scale efficiently while maintaining high availability.

We are seeking SREs who are passionate about building and maintaining resilient systems at scale. Your mission will be to design and implement robust monitoring solutions, automate operational tasks, and continuously improve our infrastructure's reliability and performance.

You will: Design and Implement Observability Solutions : Develop comprehensive monitoring and alerting systems using modern observability tools. Create dashboards and metrics that provide real-time visibility into system health and performance.

Implement logging strategies that enable quick problem identification and resolution. Drive Automation and Infrastructure as Code : Architect and implement infrastructure automation solutions using tools like Terraform, Ansible, or Pulumi.

Design and maintain CI/CD pipelines that enable reliable and consistent deployments. Create self-healing systems that can automatically respond to common failure scenarios.

Establish SLOs and SLIs : Work with product and engineering teams to define and implement Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Build systems to track and report on these metrics, ensuring we maintain high reliability standards while balancing innovation speed.

Incident Management and Response : Lead incident response efforts, conducting thorough post-morte

Required Skills

TerraformAnsiblePulumiObservabilityCI/CDInfrastructure as CodeMonitoringLoggingSLOSLIIncident Management

Partner picks for Senior Site Reliability Engineer in Remote - US

Matched to the skills this page calls for and the candidate's location.

Partner
  • edXVerified partner
    Partner course provider

    Courses and programmes from universities and institutions worldwide.

  • UdemyVerified partner
    Partner course provider

    A marketplace of instructor-created courses across technology, business and creative skills.

  • Partner competition

    Hackathon platform used by student and community hackathons across India.

Partners are ResumeKart affiliates or institutes it works with; ResumeKart may earn a commission when a candidate enrols. Placement is decided by relevance, not payment. How ResumeKart earns

The best-paying roles in your field. Every week. Free.

Join 10,000+ professionals getting job alerts and salary insights in their inbox

We respect your privacy. Unsubscribe anytime with one click.