boeing
boeing6d ago
New

Lead Site Reliability Engineer

United StatesUnited States·Usa - Berkeleylead
EngineeringDevops Engineer
1 views0 saves0 applied

Quick Summary

Key Responsibilities

Define and lead the Site Reliability Engineering technical strategy for GitLab, CI/CD runners, Jira, Confluence, PostgreSQL, Artifactory, SonarQube,

Requirements Summary

Active Secret U.S. Security Clearance. (A U.S. Security Clearance that has been active in the past 24 months is considered active.

Technical Tools
EngineeringDevops Engineer
Lead Site Reliability Engineer

The Boeing Company

The Boeing Company is looking for a Lead Site Reliability Engineer to join the Air Dominance Site Reliability Engineering team located in Berkeley, MO.

We are seeking a highly talented, motivated, and creative technical leader responsible for the reliability strategy, architecture, operational maturity, and long-term technical direction of mission-critical developer platforms used by Air Dominance engineering teams.

This role will provide technical leadership across GitLab, GitLab CI/CD runners, Jira, Confluence, PostgreSQL, related software delivery tools such as Artifactory and SonarQube, and the supporting infrastructure, automation, monitoring, backup, recovery, and security controls required to operate these services. The selected candidate will define standards, guide architecture decisions, mentor engineers, lead complex technical investigations, and partner with program leadership and stakeholders to ensure developer tooling remains secure, reliable, scalable, and supportable.

Responsibilities

~2 min read
  • Define and lead the Site Reliability Engineering technical strategy for GitLab, CI/CD runners, Jira, Confluence, PostgreSQL, Artifactory, SonarQube, and related developer tooling infrastructure

  • Establish platform reliability architecture, operational standards, SLIs, SLOs, SLAs, KPIs, error budgets, observability patterns, capacity models, backup strategies, and disaster recovery approaches

  • Serve as the senior technical authority for complex reliability, performance, scalability, integration, database, automation, and security-related platform decisions

  • Lead architecture and design reviews for developer tooling infrastructure, CI/CD runner topology, PostgreSQL operations, cloud-based and on-premises infrastructure, monitoring, alerting, access controls, and platform integrations

  • Drive automation, Infrastructure as Code, Ansible, configuration management, and repeatable operational patterns that reduce toil and improve reliability

  • Guide major upgrades, migrations, lifecycle planning, patch strategies, recovery planning, and technical roadmaps for supported platforms

  • Lead the most complex incidents and technical investigations, including root cause analysis, corrective action planning, and systemic reliability improvements

  • Mentor and technically guide SREs in operational excellence, troubleshooting, automation, secure administration, and architectural thinking

  • Partner with program leadership, cybersecurity, infrastructure, software engineering, database, networking, suppliers, customers, and other stakeholders

  • Identify platform risks, technical debt, capacity constraints, single points of failure, compliance concerns, and operational gaps, then drive remediation plans

  • Define, collect, analyze, and refine software delivery and platform reliability metrics for team execution, management visibility, and continuous improvement

  • Lead standards for provisioning, platform scaling, configuration management, monitoring, troubleshooting, and software delivery tool integration

  • Develop and maintain architecture documentation, design patterns, standards, operational readiness criteria, and executive technical briefings

  • Influence support models, maintenance strategies, escalation paths, and investment priorities based on mission impact and operational risk

  • Lead efforts to operationally field higher-quality end-to-end system software more frequently

  • Participate in after-hours support and escalation for urgent or mission-impacting issues as required

Requirements

~2 min read
  • Active Secret U.S. Security Clearance. (A U.S. Security Clearance that has been active in the past 24 months is considered active.)

  • Ability to obtain access to Special Access Programs (SAP)

  • Bachelor's Degree

  • 14+ years of experience with DevOps, Site Reliability Engineering, software engineering, and/or cloud engineering

  • Experience with GitLab, Azure DevOps and CI/CD (Continuous Integration and Continuous Delivery (CI/CD)

  • Experience with technical leadership

  • Experience with designing and implementing scalable computing infrastructure for data solutions, including cloud architectures (AWS, Azure, Google Cloud)

  • 5 years of experience conducting root cause analysis of design failures

  • Bachelor of Science degree from an accredited course of study in engineering, engineering technology (includes manufacturing engineering technology), chemistry, physics, mathematics, data science, or computer science and 14+ years of related work experience or Bachelor’s Degree and 18+ years of directly related work experience or 22+ years of related, relevant experience

  • Experience communicating technical strategy, risk, tradeoffs, and recommendations to senior technical and program leadership

  • Deep experience administering or architecting GitLab, GitLab CI/CD, GitLab runners, or comparable enterprise source control and CI/CD platforms

  • Deep experience administering or architecting Jira, Confluence, or other Atlassian products in an enterprise environment

  • Deep experience with PostgreSQL architecture and operations, including backup and recovery, replication, performance tuning, storage planning, maintenance, and high-availability patterns

  • Experience with AWS, Microsoft Azure, Infrastructure as Code, Ansible, configuration management, containers, Docker, Kubernetes, virtualization, artifact management, secrets management, and secure software delivery practices

  • Experience administering or architecting Artifactory, SonarQube, Jenkins, or similar software delivery tools

  • Experience designing observability platforms, alerting strategies, SLO frameworks, service health dashboards, and operational reporting

  • Experience supporting Air Dominance, classified, air-gapped, or highly regulated engineering environments

  • Experience developing disaster recovery strategy, continuity of operations plans, recovery time objectives, recovery point objectives, and restore validation programs

  • Experience guiding cybersecurity hardening, vulnerability remediation, audit readiness, privileged access controls, and compliance-driven operations

  • Ability to obtain Security+ certification

  • Demonstrated ability to lead through influence across engineering teams, customers, suppliers, cybersecurity, infrastructure, and program stakeholders

  • Strong written and verbal communication skills with the ability to produce architecture documentation, executive briefings, technical roadmaps, and decision records

Not Applicable

This position must meet U.S. export control compliance requirements. To meet U.S. export control compliance requirements, a “U.S. Person” as defined by 22 C.F.R. §120.62 is required. “U.S. Person” includes U.S. Citizen, U.S. National, lawful permanent resident, refugee, or asylee.

Successful candidates for this job must satisfy the Company’s Conflict of Interest (COI) assessment process.

Boeing is a Drug Free Workplace where post offer applicants and employees are subject to testing for marijuana, cocaine, opioids, amphetamines, PCP, and alcohol when criteria is met as outlined in our policies.

To be considered for this position you will be required to complete a technical assessment as part of the selection process.  Failure to complete the assessment will remove you from consideration.

What We Offer

~1 min read

At Boeing, we strive to deliver a Total Rewards package that will attract, engage and retain the top talent. Elements of the Total Rewards package include competitive base pay and variable compensation opportunities. 

The Boeing Company also provides eligible employees with an opportunity to enroll in a variety of benefit programs, generally including health insurance, flexible spending accounts, health savings accounts, retirement savings plans, life and disability insurance programs, and a number of programs that provide for both paid and unpaid time away from work. 

The specific programs and options available to any given employee may vary depending on eligibility factors such as geographic location, date of hire, and the applicability of collective bargaining agreements. 

Pay is based upon candidate experience and qualifications, as well as market and business considerations. 

Bachelor's Degree or Equivalent

This position offers relocation based on candidate eligibility.

This is not a Safety Sensitive Position.

This position requires an active U.S. Secret Security Clearance (U.S. Citizenship Required). (A U.S. Security Clearance that has been active in the past 24 months is considered active)

Employer will not sponsor applicants for employment visa status.

This position is not contingent upon program award

Shift 1 (United States of America)

Location & Eligibility

Where is the job
Usa - Berkeley, United States
On-site at the office
Who can apply
US

Listing Details

Posted
August 6, 2026
First seen
August 12, 2026
Last seen
August 12, 2026

Posting Health

Days active
0
Repost count
0
Trust Level
28%
Scored at
August 12, 2026

Signal breakdown

freshnesssource trustcontent trustemployer trust
Newsletter

Stay ahead of the market

Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.

A
B
C
D
Join 12,000+ marketers

No spam. Unsubscribe at any time.

boeingLead Site Reliability Engineer