Manager Site Reliability Engineering
Quick Summary
10+ years of relevant professional experience, along with a Bachelor's degree in Computer Science or an equivalent combination of education and experience.
This role offers the opportunity to lead a global team of Site Reliability Engineers responsible for critical Identity and Access Management (IAM) services. You will combine people leadership with technical direction, helping teams build reliable, scalable, secure, and highly usable cloud services. You will work closely with product and engineering stakeholders to shape roadmaps and drive meaningful technical improvements. The role operates in a collaborative, inclusive, and Agile environment with a strong focus on innovation and operational excellence. You will help remove technical and organizational blockers while developing engineers and strengthening team performance. Your leadership will directly contribute to the reliability and scalability of services used at significant global scale.
-
Lead, manage, and mentor a global team of Site Reliability Engineers, supporting their technical development, career growth, and day-to-day effectiveness.
-
Remove technical and process blockers, helping engineers resolve complex challenges and maintain momentum toward team objectives.
-
Partner with product, engineering, and other stakeholders to understand priorities, shape roadmaps, and align SRE initiatives with broader business and technical goals.
-
Drive technical and product innovation by applying design thinking, analyzing complex systems, and developing solutions that improve reliability, scalability, security, and usability.
-
Organize and support SRE activities within an Agile environment, establishing clear goals, commitments, accountability, and ownership.
-
Foster strong, collaborative working relationships built on trust, transparency, and shared responsibility.
-
Diagnose system performance and reliability challenges alongside SRE engineers and guide teams toward sustainable solutions.
-
Contribute to internal research, methodologies, and engineering practices while helping share and promote operational best practices across teams.
Requirements
~1 min read-
10+ years of relevant professional experience, along with a Bachelor's degree in Computer Science or an equivalent combination of education and experience.
-
Professional experience in Site Reliability Engineering, software development, or a closely related technical discipline, ideally involving large-scale distributed systems.
-
Proven experience leading, managing, and mentoring engineering teams.
-
Strong knowledge of cloud-native technologies and experience working with cloud deployments at scale.
-
Hands-on familiarity with container runtimes and technologies such as Docker and containers.
-
Experience with programming or scripting languages such as Go, Python, and/or Bash.
-
Solid understanding of modern software development practices and product lifecycles, including build, test, deployment, and iterative delivery environments.
-
Strong problem-solving, communication, collaboration, and stakeholder-management skills.
-
Ability to balance strategic leadership and people management with sufficient technical depth to guide complex reliability and infrastructure initiatives.
What We Offer
~2 min readLocation & Eligibility
Listing Details
- Posted
- October 6, 2026
- First seen
- October 6, 2026
- Last seen
- October 6, 2026
Posting Health
- Days active
- 0
- Repost count
- 0
- Trust Level
- 68%
- Scored at
- October 6, 2026
Signal breakdown
Browse Similar Jobs
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.