Staff Software Engineer, Runtime Systems
Quick Summary
The Physical AI Engineering Team at CoreWeave is building the software and infrastructure that enables demanding AI, simulation, robotics, and engineering workloads to run reliably at scale.
Medical, dental, and vision insurance - 100% paid for by CoreWeave Company-paid Life Insurance Voluntary supplemental life insurance Short and long-term disability insu
Responsibilities
~2 min readThe Physical AI Engineering Team at CoreWeave is building the software and infrastructure that enables demanding AI, simulation, robotics, and engineering workloads to run reliably at scale.
As these workloads become more complex, the challenge is no longer simply providing compute. We need to make heterogeneous workloads easier to execute, observe, reproduce, and move across different systems without hiding the capabilities or semantics of the infrastructure underneath them.
At CoreWeave, we work hard, have fun, and move fast. We’re in an exciting stage of hyper-growth, and we’re constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values:
- →Be Curious at Your Core
- →Act Like an Owner
- →Empower Employees
- →Deliver Best-in-Class Client Experiences
- →Achieve More Together
We support entrepreneurial thinking, independent judgement, and collaboration. You’ll work alongside some of the best talent in the industry on technically difficult problems at the frontier of AI infrastructure.
The base salary range for this role is 116,000 GBP to 155,000 GBP. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility).
About the Role
~1 min readWe’re seeking a Staff Software Engineer, Runtime Systems to help design and build this layer.
This is a hands-on systems engineering role at the intersection of distributed systems, runtimes, workflow execution, programming language concepts, and large-scale compute infrastructure.
You’ll work on the foundations that allow complex workloads to move from an abstract description into reliable execution across systems such as Kubernetes, Argo, OSMO, and future execution environments.
A major part of the role is deciding where abstraction is useful — and where it creates more complexity. Rather than building another universal workflow engine, you’ll help establish clear contracts between our platform and the systems that execute work, while preserving native capabilities.
You’ll operate across architecture and implementation: defining contracts, writing production software, validating assumptions against real workloads, and working closely with platform and infrastructure teams.
- Runtime Systems & Architecture
- Design and build runtime components for complex AI, simulation, and engineering workloads.
- Define abstractions for workloads, execution environments, dependencies, state, capabilities, and failure.
- Design interfaces between higher-level services and systems such as Kubernetes, Argo, and OSMO.
- Establish clear boundaries around which systems own state, decisions, and side effects.
- Make architectural decisions balancing simplicity, extensibility, performance, and operational reality.
- Execution Models & Contracts
- Define durable contracts between workload definitions, control-plane services, and execution backends.
- Develop typed representations and schemas that allow workloads to be transformed safely across systems.
- Design compatibility and evolution mechanisms for those contracts.
- Build conformance and validation mechanisms that make guarantees executable rather than dependent on documentation.
- Reason deeply about retries, partial failure, idempotency, cancellation, dependencies, and uncertain outcomes.
- Hands-On Systems Engineering
- Write production-quality software for critical runtime and control-plane components.
- Build adapters and integrations for heterogeneous execution environments.
- Diagnose behaviour across application, orchestration, cluster, and infrastructure boundaries.
- Improve the reliability, observability, and debuggability of distributed workload execution.
- Work closely with Go, Kubernetes, and infrastructure engineers to turn architecture into production systems.
- Performance & Experimentation
- Develop rigorous ways to understand workload performance across large-scale GPU infrastructure.
- Design experiments that separate real performance gains from noise, warm-up effects, scheduling behaviour, and stragglers.
- Build repeatable workload and benchmark environments.
- Use evidence from real execution to challenge assumptions and guide platform development.
- Technical Leadership
- Lead ambiguous systems problems where the correct architecture is not yet known.
- Reduce complex problems into smaller contracts and mechanisms that can actually be implemented.
- Challenge unnecessary abstraction and simplify designs where complexity has outgrown its value.
- Influence technical direction across teams without requiring direct authority.
- Mentor engineers and contribute to technical hiring and engineering standards.
- Significant experience building complex systems software, distributed infrastructure, runtimes, workflow systems, or adjacent technology.
- Deep expertise in at least one of:
- Distributed systems
- Runtime systems
- Workflow or execution engines
- Programming languages, compilers, or interpreters
- Cluster scheduling and orchestration
- High-performance or systems software
- Strong software engineering fundamentals and production coding ability.
- Experience designing APIs, protocols, schemas, or contracts between independently evolving systems.
- Strong understanding of distributed-system failure modes, state, authority, retries, concurrency, and side effects.
- Strong technical judgement around when abstraction helps and when it simply moves complexity elsewhere.
- Comfortable entering unfamiliar technical domains and building depth quickly.
- Strong communication skills and experience influencing architectural decisions across teams.
Experience with some of the following would be valuable, but is not required:
- Go, Rust, C/C++, or Python.
- Kubernetes and containerised infrastructure.
- Argo, OSMO, Temporal, Ray, Kubeflow, or similar systems.
- GPU clusters or large-scale AI infrastructure.
- High-performance computing.
- Simulation, robotics, autonomous systems, or Physical AI.
- Programming language or compiler research.
- Performance engineering.
- Cloud infrastructure at scale.
- Platforms designed to be operated by autonomous software or AI agents.
What We Offer
~1 min readThe range we’ve posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location.
In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings for full-time employees; for roles in other locations, benefits vary and are shared during the hiring process. These include:
CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace. All qualified applicants and candidates will receive consideration for employment without regard to race, color, religion, sex, disability, age, sexual orientation, gender identity, national origin, veteran status, or genetic information.
As part of this commitment and consistent with the Americans with Disabilities Act (ADA), CoreWeave will ensure that qualified applicants and candidates with disabilities are provided reasonable accommodations for the hiring process, unless such accommodation would cause an undue hardship. If reasonable accommodation is needed, please contact: careers@coreweave.com.
This position requires access to export controlled information. To conform to U.S. Government export regulations applicable to that information, applicant must either be (A) a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful permanent resident (green card holder), (iii) refugee under 8 U.S.C. § 1157, or (iv) asylee under 8 U.S.C. § 1158, (B) eligible to access the export controlled information without a required export authorization, or (C) eligible and reasonably likely to obtain the required export authorization from the applicable U.S. government agency. CoreWeave may, for legitimate business reasons, decline to pursue any export licensing process.
Location & Eligibility
Listing Details
- Posted
- September 15, 2026
- First seen
- September 15, 2026
- Last seen
- September 15, 2026
Posting Health
- Days active
- 0
- Repost count
- 0
- Trust Level
- 67%
- Scored at
- September 15, 2026
Signal breakdown
Please let CoreWeave know you found this job on Jobera.
3 other jobs at CoreWeave
View all →Explore open roles at CoreWeave.
Similar Software Engineer jobs
View all →Browse Similar Jobs
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.
