Saviynt
Saviynt5mo ago

Staff Site Reliability Engineer

Bangalore · BengaluruFull-Timelead
OtherDevOps & InfrastructureSite Reliability EngineerStaff Site Reliability EngineerInfrastructure & Cloud
0 views0 saves0 applied

Quick Summary

Overview

Saviynt's AI-powered identity platform manages and governs human and non-human access to all of an organization's applications, data, and business processes.

Technical Tools
OtherDevOps & InfrastructureSite Reliability EngineerStaff Site Reliability EngineerInfrastructure & Cloud
Saviynt's AI-powered identity platform manages and governs human and non-human access to all of an organization's applications, data, and business processes. Customers trust Saviynt to safeguard their digital assets, drive operational efficiency, and reduce compliance costs. Built for the AI age, Saviynt is today helping organizations safely accelerate their deployment and usage of AI. Saviynt is recognized as the leader in identity security, with solutions that protect and empower the world’s leading brands, Fortune 500 companies and government institutions. For more information, please visit www.saviynt.com.

We’re a fast-moving AI Security Company building AI-native infrastructure and
applications powered by LLMs and autonomous agents. Our stack is deeply integrated with AWS, Kubernetes, and OpenAI-based systems, and we’re rethinking reliability in a world where software can reason, adapt, and self-heal.
 
We’re hiring a Staff SRE Engineer to own reliability across our cloud-native and AI-driven platform. You’ll work at the intersection of distributed systems, Kubernetes operations, and LLM-powered automation, building systems that don’t just scale—but think and fix themselves.
  • Own uptime, reliability, and performance of services running on AWS + Kubernetes (EKS).
  • Design and implement self-healing infrastructure using automation and AI agents.
  • Build LLM-powered operational tooling using APIs such as the OpenAI API for:
    • Intelligent alert triage
    • Incident summarization
    • Root cause analysis
    • Runbook automation
    • Manage and scale Kubernetes workloads:
      • Deployments, autoscaling, resource optimization
      • Cluster reliability and cost efficiency
      • Build and evolve observability systems:
        • Metrics (Prometheus), dashboards (Grafana)
        • Logs (ELK / OpenSearch)
        • Tracing (OpenTelemetry)
        • Define and enforce SLOs, SLAs, and error budgets tied to business metrics.
        • Automate infrastructure using Terraform and CI/CD pipelines.
        • Lead incident response, postmortems, and continuous reliability improvements.
        • Introduce chaos engineering practices to proactively test system resilience.
  • 8+ years in SRE / DevOps / Platform Engineering.
  • Strong hands-on experience with:
    • AWS infrastructure at scale
    • Kubernetes (production-grade clusters)
    • Proven ability to debug complex distributed systems under pressure.
    • Strong coding skills (Python or Go)—you build internal platforms and tools.
    • Experience implementing monitoring, alerting, and incident management systems.

Nice to Have

~1 min read
  • Experience working with LLM APIs such as the OpenAI API.
  • Familiarity with agent frameworks like:
    • LangChain
    • AutoGen
    • Built or experimented with:
      • AI agents for DevOps / SRE workflows
      • Retrieval-Augmented Generation (RAG) systems
      • Vector databases (Pinecone, Weaviate, etc.)
      • Exposure to AIOps or intelligent automation systems.

Listing Details

Posted
November 11, 2025
First seen
March 26, 2026
Last seen
April 24, 2026

Posting Health

Days active
29
Repost count
0
Trust Level
33%
Scored at
April 25, 2026

Signal breakdown

freshnesssource trustcontent trustemployer trust
Saviynt
Saviynt
lever

Saviynt is a leading provider of cloud-native identity and governance platform solutions, empowering enterprises to secure their digital transformation, safeguard critical assets, and meet regulatory compliance.

Employees
3k+
Founded
2010
View company profile
Newsletter

Stay ahead of the market

Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.

A
B
C
D
Join 12,000+ marketers

No spam. Unsubscribe at any time.

SaviyntStaff Site Reliability Engineer