Senior SRE – OpenStack & Bare Metal (m/f/d)
Quick Summary
At our company, it’s all about #OneTeam! Join gridscale and help shape the future of the cloud together with OVH. As a leading tech company,
At our company, it’s all about #OneTeam! Join gridscale and help shape the future of the cloud together with OVH.
As a leading tech company, we’ve been working for over two decades to reduce our environmental footprint - with innovative solutions and an open cloud designed to be sustainable from the ground up: #SustainableByDesign.
· OpenStack · Kubernetes · KVM · Linux · Bare Metal · Ansible · Terraform · Go
· FluxCD / ArgoCD · Git· Python · Claude Code · Cursor · Agentic Coding Tooling
You will help us build, operate and industrialize OVHcloud's on-premise cloud platform. As part of a small, senior team, you will work on our OpenStack-based infrastructure and the Kubernetes / GitOps stack powering our customer-facing cloud. AI-assisted engineering is a first-class part of how we work – from spec-driven development and agentic coding workflows to incident response and automation. The platform is in active development, giving you real influence on the architecture, automation strategy and how we adopt AI in platform engineering. As a Senior Engineer, you take ownership of your area and shape your focus based on your strengths, with a clear backbone of automation, compute lifecycle, platform engineering and AI substrate work.
Responsibilities
~1 min read- →
Design and build our OpenStack-based on-premise cloud infrastructure, with the goal of bringing complete cloud environments from bare metal to production through a highly automated deployment process.
- →
Develop and operate Infrastructure as Code with Ansible and Terraform as well as our Kubernetes and GitOps workflows using FluxCD / ArgoCD, supported by LLMs, agentic workflows and automated testing and review.
- →
Own the lifecycle of our compute infrastructure – from bare metal, firmware and provisioning through to hypervisors and virtual compute nodes. This includes patching, migrations, host evacuation, capacity rebalancing and the automation required to keep the platform healthy.
- →
Build and extend our AI substrate and self-healing capabilities, from structured knowledge bases and agentic workflows for incident triage and capacity planning to gradually turning today's manual runbooks into automated processes.
- →
Design tests for non-regression, performance and security, document and package solutions, and continuously improve the platform based on telemetry, operational experience and user feedback.
- →
Act as a technical reference and sparring partner for peers across automation, platform engineering and AI tooling.
What We Offer
~2 min readLocation & Eligibility
Listing Details
- First seen
- August 7, 2026
- Last seen
- October 5, 2026
Posting Health
- Days active
- 58
- Repost count
- 0
- Trust Level
- 19%
- Scored at
- October 5, 2026
Signal breakdown
Browse Similar Jobs
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.