21h ago
New
$180,200 – $233,200/yr

Senior Software Engineer, Site Reliability Engineering

CanadaCanadaRemoteFull-timesenior
OtherSite Reliability Engineering
1 views0 saves0 applied

Quick Summary

Overview

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Software Engineer,

Technical Tools
OtherSite Reliability Engineering

This is a senior-level opportunity to help build and operate the reliable, secure, and highly scalable infrastructure behind a high-impact digital platform.
You’ll work across the technology stack, from Linux systems and cloud infrastructure to distributed applications and platform services.
The role combines hands-on engineering with architectural leadership, enabling product and infrastructure teams to deliver resilient systems at scale.
You’ll collaborate closely with engineers across product development, developer experience, and backend infrastructure.
Your work will directly influence system availability, performance, observability, and the overall engineering experience.
You’ll also help shape platform capabilities, improve operational practices, and anticipate future capacity and reliability needs.
This is an ideal environment for an experienced SRE who enjoys solving complex technical problems and creating high-leverage engineering solutions.

  • Design, develop, and maintain software and infrastructure that improve service availability, scalability, performance, and operational efficiency.

  • Establish architectural direction for infrastructure and platform services while providing technical guidance and support to engineering teams.

  • Build and improve tools, processes, and systems for deployment, infrastructure, service, and change management.

  • Troubleshoot and resolve complex production issues across the software development lifecycle, with a focus on minimizing downtime and service disruption.

  • Develop and evolve platform capabilities that enable engineering teams to build, deploy, operate, and observe services more effectively.

  • Conduct capacity planning and demand forecasting to anticipate system growth, identify performance bottlenecks, and proactively address scalability challenges.

  • Instrument, operate, and monitor distributed microservices and cloud-based systems to maintain strong reliability and observability.

  • Participate in a rotating on-call schedule and contribute to incident response, service recovery, and continuous reliability improvements.

  • Partner with cross-functional engineering teams and stakeholders to identify opportunities, balance technical trade-offs, and deliver high-impact platform solutions.

Requirements

~1 min read
  • 5+ years of experience managing infrastructure and systems, ideally within large-scale or distributed production environments.

  • Extensive hands-on expertise with AWS and Linux-based systems.

  • Strong ability to read, write, debug, and maintain production-facing software and systems.

  • Deep understanding of large-scale distributed systems and web technologies, including DNS, TLS, HTTP/S, TCP/IP, and related networking concepts.

  • Demonstrated experience operating, instrumenting, and observing distributed microservices in production cloud environments.

  • Strong problem-solving skills, with the ability to break down complex technical challenges and make thoughtful trade-offs based on business and engineering impact.

  • Experience designing resilient, scalable infrastructure and platform services with a focus on availability, performance, and reliability.

  • Strong communication and collaboration skills, with the ability to work effectively with technical and non-technical stakeholders across different levels of an organization.

  • Ability to operate independently while contributing effectively to a collaborative, cross-functional engineering environment.

  • A proactive mindset and strong ownership of production systems, operational excellence, and continuous improvement.

What We Offer

~2 min read
✓Remote work opportunity from eligible locations in Ontario and British Columbia, Canada.
✓Expected total cash compensation of CAD $180,200–$233,200, depending on location, qualifications, skills, competencies, and experience.
✓Opportunity to work on large-scale, distributed systems with significant impact across the engineering organization.
✓Collaborative environment with exposure to product engineering, developer experience, backend infrastructure, and platform teams.
✓Opportunity to influence architectural direction and shape the evolution of engineering infrastructure and platform capabilities.
✓Participation in meaningful reliability, scalability, observability, and infrastructure initiatives.
✓Commitment to diversity, inclusion, and equal employment opportunity.
✓Reasonable accommodations available throughout the recruitment process for candidates who require them.

Location & Eligibility

Where is the job
Canada
Remote within one country
Who can apply
CA

Listing Details

Posted
October 5, 2026
First seen
October 5, 2026
Last seen
October 5, 2026

Posting Health

Days active
0
Repost count
0
Trust Level
80%
Scored at
October 6, 2026

Signal breakdown

freshnesssource trustcontent trustemployer trust
Newsletter

Stay ahead of the market

Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.

A
B
C
D
Join 12,000+ marketers

No spam. Unsubscribe at any time.

Senior Software Engineer, Site Reliability Engineering$180k–$233k