Senior Software Engineer - SRE
Mercury•San Francisco, CA, New York, NY, Portland, OR, or Remote within Canada or United States•🌍 Remote
Full-timeSenior
👁️ 0 views•📝 0 applications•Posted 9/10/2026•Expires 10/10/2026
Get alerts for roles like this
More Senior Software Engineer - SRE roles in San Francisco, CA, New York, NY, Portland, OR, or Remote within Canada or United States — straight to your inbox. No account needed.
Applying to this role? Tailor your résumé to this job description in one click, then download it clean — no watermark, no subscription.
Job Description
When the Tarr Steps, a footbridge assembled of heavy stones in Exmoor National Park in England, washed away in a flood in 1942, the Royal Engineers rebuilt it. Then it washed away again in 1952. So the Royal Engineers heaved in more stones: on and on, like a Sisyphean lesson in absurdity. Of course, they do this because the Tarr Steps are a historical monument, believed to be built during the Bronze Age. Modern bridge building looks a lot different and generally requires less upkeep. While we appreciate the charm of ancient things, Mercury is engineering the future of banking*. "Quaint" and "archaic" are not values we seek in our systems. Rebooting a tumbling server and hoping a flood of requests doesn't wash it away can be an emergency tactic, but we actively seek out durable solutions over monotonous ops work. You'll bring this ethos and teach others how to live it too. Up to this point, the Stability team at Mercury has primarily built platform-level constructs that product teams adopt. However, those teams are asking for us to work more closely with them to mature their implementations and practices. We are creating an SRE team that will rotate through product teams. You will understand the team's domain, identify opportunities for improvements in reliability/observability/performance/preparedness and create self-reinforcing, virtuous cycles. As part of this role, you will: Embed with product teams, helping them improve their operational maturity by setting up and refining practices around on-call, monitoring, alerting, and run books Run regular game day exercises with product teams, helping them feel more prepared to investigate and quickly remediate incidents Jump into application code written in Haskell & TypeScript and implement reliability techniques such as retries, better error handling, better logging, circuit breaking, etc Steer SLOs towards meaningful customer outcomes that product t
Required Skills
HaskellTypeScriptmonitoringalertingon-callrun booksreliabilityobservabilityperformancegame day exerciseserror handlingloggingcircuit breakingSLOs
Prepare to Win This Role
Everything you need to ace the interview and negotiate top-of-band compensation.