Site Reliability Engineer - Canada Wide - Remote
Quick Summary
• Implementing the improvements to the reliability, fault tolerance, scalability,
Role Overview:
We're looking for a Site Reliability Engineer to improve the reliability, resilience, and operational readiness of our services. You’ll work closely with engineering teams to improve system design and operational excellence. You’ll help prevent incidents, lead response efforts, and drive improvements through post-mortems. Your mission: ensure our systems are reliable, scalable, and resilient.
Responsibilities will include:
• Implementing the improvements to the reliability, fault tolerance, scalability, and performance of our infrastructure
• Managing incidents using your technical know-how to involve the appropriate teams and automate away manual practices
• Providing support to our critical services by responding to automated alerts through our on-call rotation
• Define and maintain SLIs, SLOs,SLA, and error budgets to guide reliability decisions
• Improve observability across our systems (metrics, logs, tracing) to reduce time to detection and resolution
• Make production issues easier to detect, troubleshoot, and resolve
• Improving monitoring, alerting, dashboards, tracing and runbooks for critical services
• Leading postmortems and follow-up actions to reduce repeat incidents
Who you are:
• You have experience designing and operating scalable, reliable systems in AWS or a similar cloud environment
• You have handled on-call shifts for critical systems
• You are experienced with chaos engineering (i.e. Gremlin)
• You are able to dive in and debug live production systems
• You enjoy working in a growing system, and writing and deploying code without any downtime
• You have experience scripting and/or development (i.e. Linux Shell, Python, Javascript, Java)
• You are a self-starter, taking initiative in an ambiguous space preferably within a start-up environment
Listing Details
- Posted
- March 26, 2026
- First seen
- March 26, 2026
- Last seen
- April 24, 2026
Posting Health
- Days active
- 29
- Repost count
- 0
- Trust Level
- 32%
- Scored at
- April 24, 2026
Signal breakdown
Please let Newton know you found this job on Jobera.
2 other jobs at Newton
View all →Explore open roles at Newton.
Similar Devops Engineer jobs
View all →Browse Similar Jobs
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.
