Site Reliability Engineer (SRE)
Quick Summary
Expert knowledge of Kubernetes. Expert knowledge of continuous deployment systems such as Buildkite and ArgoCD. Expert knowledge of monitoring technologies such as Prometheus, Grafana, and PagerDuty.
xAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.
ABOUT THE ROLE:
You will work on the team that is responsible for the backend services that power our products such as grok.com and the API. We focus on writing and maintaining highly scalable and reliable services that can efficiently process tens of thousands of queries per second. The services are hosted on a number of Kubernetes clusters (on-prem & cloud).
BASIC QUALIFICATIONS:
- Expert knowledge of Kubernetes.
- Expert knowledge of continuous deployment systems such as Buildkite and ArgoCD.
- Expert knowledge of monitoring technologies such as Prometheus, Grafana, and PagerDuty.
- Expert knowledge of infrastructure as code technologies such as Pulumi or Terraform.
- Familiarity with a systems programming language like Rust, C++ or Go
- Experience with traffic management and HTTP proxies such as nginx and envoy.
xAI is an equal opportunity employer. For details on data processing, view our Recruitment Privacy Notice.
Listing Details
- Posted
- April 17, 2026
- First seen
- March 26, 2026
- Last seen
- April 19, 2026
Posting Health
- Days active
- 23
- Repost count
- 0
- Trust Level
- 74%
- Scored at
- April 19, 2026
Signal breakdown
Please let Xai know you found this job on Jobera.
Similar Site Reliability Engineer jobs
View all →Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.
