Cloud Performance Engineering - Site Reliability Engineer ( Remote Canada)
Quick Summary
Working for a company like Smile Digital Health means supporting our mandate for #BetterGlobalHealth. We strive towards this goal every day,
The Cloud Site Reliability Engineer (SRE) is responsible for ensuring the reliability, scalability, and performance of production-grade services deployed across multiple cloud vendors and infrastructure platforms for Smile Digital Health, its clients, and partners.
This role designs and automates performance testing frameworks, integrates them into CI/CD pipelines, and uses observability tools to proactively detect and resolve bottlenecks. Working closely with engineering, product, and security teams, the SRE ensures systems meet strict SLAs for performance and availability while driving continuous optimization across multiple cloud platforms.
- 5+ years of experience with Cloud Service Providers and best practices around implementation and configuration, preferably managing Azure environments supporting SaaS products.
- Experience working across multiple cloud providers (Azure required; AWS and/or Google Cloud Platform considered an asset).
- Strong experience in Cloud Performance Engineering, including performance analysis, capacity planning, scalability testing, and optimization of distributed cloud-native applications.
- Proven experience working with microservices architecture, with a strong focus on Java-based services.
- Experience applying Chaos Engineering practices to evaluate and improve system resiliency.
- Strong experience designing and executing performance testing strategies, including load, stress, spike, and endurance (soak) testing, to validate application scalability and defined latency and error-rate thresholds.
- Hands-on experience with performance testing tools such as JMeter, Gatling, Azure Load Testing, or k6.
- Experience validating application services sustaining 500+ transactions per second (TPS) while meeting defined performance objectives.
- Hands-on experience deploying and managing containerized applications using Docker and Kubernetes, including autoscaling and performance optimization.
- Experience using Terraform to provision and manage cloud infrastructure using Infrastructure as Code (IaC).
- Experience tuning Kafka (partitioning, consumer group sizing, throughput/latency trade-offs) and other messaging/queueing platforms to sustain target transaction rates.
- Hands-on experience implementing and using observability platforms including OpenTelemetry, Prometheus, Grafana, Azure Monitor, Application Insights, and Log Analytics.
- Proven experience with Security and Compliance (SOC 2, HIPAA, ISO 27001) best practices and implementing controls that support high-velocity software delivery teams.
Location & Eligibility
Listing Details
- Posted
- July 10, 2026
- First seen
- July 10, 2026
- Last seen
- July 30, 2026
Posting Health
- Days active
- 0
- Repost count
- 0
- Trust Level
- 80%
- Scored at
- July 10, 2026
Signal breakdown
Please let Smiledigitalhealth know you found this job on Jobera.
4 other jobs at Smiledigitalhealth
View all →Explore open roles at Smiledigitalhealth.
Similar Devops Engineer jobs
View all →Browse Similar Jobs
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.