Quick Summary
Own the availability and performance of production SaaS applications running on Azure (AKS, App Service, Redis, SQL, Service Bus etc), across multiple geographic regions.
Our growing technology company is seeking an experienced Site Reliability Engineer with deep Azure expertise to help maintain the availability, performance, and reliability of our critical SaaS applications. In this role, you will own and drive automation, monitoring, incident response, and infrastructure improvements across our multi-cloud, multi-region environment, working closely with senior engineering and cross-functional teams.
Responsibilities
~1 min read- →
Own the availability and performance of production SaaS applications running on Azure (AKS, App Service, Redis, SQL, Service Bus etc), across multiple geographic regions.
- →
Lead troubleshooting and resolution of cloud infrastructure and application issues, including AKS pod/node failures, deployment rollbacks, ingress and networking issues, and resource/autoscaling problems.
- →
Participate in an on-call rotation (including weekends) and drive incident response from detection through resolution, with a primary focus on customer experience and minimizing impact.
- →
Drive improvements to disaster recovery, failover, and incident management processes across multi-region deployments.
- →
Build and maintain automation scripts and monitoring tools to reduce manual toil and streamline operational tasks.
- →
Author post-incident reviews (RCAs), identify root causes, and drive preventive action items to closure.
- →
Partner with senior engineers and cross-functional teams to implement and improve reliability, observability, and performance best practices.
- →
Contribute to continuous improvement initiatives across infrastructure, tooling, and process.
- →
Communicate clearly with customer-facing stakeholders when incidents require external status updates or written incident summaries.
5+ years of relevant experience in Site Reliability Engineering, DevOps, or Cloud Administration, with demonstrated ownership of production systems.
Hands-on experience administering Azure environments, including AKS (Kubernetes), core Azure services, cloud networking, and cloud security fundamentals.
Solid understanding of monitoring, logging, and alerting practices (e.g., Datadog, Azure Monitor, ELK stack), including hands-on troubleshooting with log analysis and stack traces using Datadog APM.
Familiarity with networking fundamentals: firewalls, load balancers, VPNs, DNS, and routing.
Experience with automation and scripting (PowerShell, Python, or similar).
Practical understanding of backup, redundancy, and disaster recovery strategies in cloud environments, including geo-redundant / multi-region deployments.
Strong ownership mindset across the full incident lifecycle, from detection through post-mortem, with a customer-first approach.
Comfort communicating clearly and professionally in written form, including incident status updates and post-incident summaries.
Experience with AWS Cloud Platform.
Experience with CI/CD tools such as Azure DevOps.
Experience with infrastructure-as-code tools such as Terraform or ARM templates.
Prior experience operating SaaS products with regional tenant architectures (e.g., multiple geo-specific production environments).
What We Offer
~1 min readSpirited - We bring energy and passion to everything we do
Trust - We act with integrity and deliver on our commitments
Respect - We listen, value different perspectives, and work as one team
Ownership - We take initiative and follow through
Nimble - We adapt quickly in a fast-changing environment
Global - We embrace diverse people and ideas to drive better outcomes
We believe weaving these core values into our day-to-day actions, and our process for hiring, evaluating, and promoting employees, helps us cultivate a work environment that embraces collaboration and camaraderie.
We take care of our employees. We offer competitive salaries, a meaningful bonus program, and excellent benefits, including healthcare insurance, as well as pension/retirement matching, comprehensive life insurance, an employee assistance program, time off plans, and paid company holidays.
Delinea is an Equal Opportunity and Affirmative Action employer and prohibits discrimination and harassment of any type with regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.
Upon conditional offer of employment, candidates are required to complete comprehensive criminal background check, verification of education, and verification of employment, per employment policy. In addition, all publicly posted social media sites may be reviewed.
Location & Eligibility
Listing Details
- Posted
- August 18, 2026
- First seen
- August 18, 2026
- Last seen
- August 18, 2026
Posting Health
- Days active
- 0
- Repost count
- 0
- Trust Level
- 52%
- Scored at
- August 18, 2026
Signal breakdown
Please let delinea know you found this job on Jobera.
3 other jobs at delinea
View all →Explore open roles at delinea.
Similar Devops Engineer jobs
View all →Browse Similar Jobs
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.