Sr. Site Reliability Engineer
Quick Summary
Own production services end to end. Accountable for reliability, availability, scalability, performance, and operational health. Define and manage SLIs and SLOs,
Nice to Have
~1 min readThe insurance industry runs on Vertafore. We equip agencies, MGAs, and carriers with the core digital systems, specialized AI, and data-driven foundation to eliminate distribution drag across the insurance lifecycle, spanning sales, servicing, and back-office operations.
Underpinned by unmatched speed and performance power, we are the trusted backbone that’s taking the insurance industry from friction to flow with Distribution Velocity – speed, performance, and trust - to drive growth at scale.
With over 95% of the top agencies and insurers and 50% of industry compliance transactions running through Vertafore, we lead at the intersection of innovation and trust, giving insurance professionals the confidence to transform and win in the AI era.
Our reach is global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India.
We are seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services. This role is accountable for the full-service lifecycle, from design and deployment readiness through production operations, incident response, and continuous improvement. Reliability is a core engineering responsibility, requiring strong software engineering skills and autonomous operation across AWS, hybrid data centers, and customer-hosted environments.
Responsibilities
~1 min read- →
Own production services end to end. Accountable for reliability, availability, scalability, performance, and operational health.
- →
Define and manage SLIs and SLOs, using error budgets to guide delivery decisions.
- →
Influence of service and system design to improve fault tolerance, observability and operational sustainability.
- →
Debug complex production issues across application code, services and infrastructure using software engineering practices.
- →
Perform root cause analysis using logs, metrics, traces, and code-level investigation.
- →
Build automation and self-healing mechanisms to prevent repeat failures.
- →
Execute production changes (patching, certificate management, software releases) with safety, automation, and observability.
- →
Design and operate production observability aligned to service health and customer impact.
- →
Lead and participate in incident response, for high-severity events.
- →
Collaborate with engineering, product, architecture, and operations teams.
- →
Operate with autonomy and sound judgment in reliability decisions.
Requirements
~1 min read-
8+ years of hands-on Site Reliability Engineering or reliability-focused engineering experience with end-to-end service ownership.
-
Proven operation at a senior engineering scope with accountability for reliability outcomes.
-
Strong software engineering skills in C#, .NET, Java, Python, React, or similar technologies.
-
Practical experience applying SRE principles (SLIs, SLOs, error budgets).
-
Hands-on experience with AWS, Kubernetes, CI/CD, infrastructure as code and hybrid environments.
-
Strong knowledge of Linux and Windows systems, application platforms and relational databases.
-
Bachelor’s or master’s degree in computer science or equivalent experience.
-
Participation in an on-call rotation; flexible hours as required.
Location & Eligibility
Listing Details
- First seen
- August 23, 2026
- Last seen
- August 23, 2026
Posting Health
- Days active
- 0
- Repost count
- 1
- Trust Level
- 45%
- Scored at
- August 23, 2026
Signal breakdown
Please let Vertafore Career Center know you found this job on Jobera.
4 other jobs at Vertafore Career Center
View all →Explore open roles at Vertafore Career Center.
Similar Devops Engineer jobs
View all →Browse Similar Jobs
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.