Senior Site Reliability Engineer
Quick Summary
Work with large-scale systems (handling millions of requests per second, serving millions of users, across multiple cloud providers). Develop solutions to enhance performance, availability, security,
Coding agents are indisputably useful tools. We provide access to the top 3 models, and were early adopters of evals-first development flows.
Intuition Machines uses AI/ML to build enterprise security products. We apply our research to systems that serve hundreds of millions of people, with a team distributed around the world. You are probably familiar with our best-known product, the hCaptcha security suite. Our approach is simple: low overhead, small teams, and rapid iteration.
As a Senior Site Reliability Engineer, you will focus on engineering solutions related to performance, availability, security, and cost-effectiveness. We consider these non-functional features to be core requirements for us and our customers. You will work at multiple layers of our internet-scale system (infrastructure, data, application logic) and build the solutions.
Using AI: Coding agents are indisputably useful tools. We provide access to the top 3 models, and were early adopters of evals-first development flows. Familiarity with coding using agents is part of all interviews. However, reliability and correctness are critical for us. You will need to read and understand every line of code with your name on it, and it will be reviewed by both people and machines.
Responsibilities
~1 min read- →Work with large-scale systems (handling millions of requests per second, serving millions of users, across multiple cloud providers).
- →Develop solutions to enhance performance, availability, security, and cost-effectiveness.
- →Keep us up, keep us fast, and keep our dev teams productive ensuring that every peer release improves performance across the spectrum including quality, security, uptime, speed-to-deliver, threat detection, and customer engagement.
- →Source improvement ideas, priority and capabilities from customers, the internal community, new and existing system metrics. Make decisions rapidly.
- →Be creative and desire an environment where you can directly create value and be a force to improve the experience for our customers.
- Expert in Kubernetes.
- Expert in monitoring applications, infrastructure and network.
- Background in software engineering with expertise in backend development within Kubernetes-based systems.
- Strong programming skills in one or more of the following languages: Python, JavaScript, Go, C++, Rust.
- Strong understanding and experience in networking, proxies, content delivery networks (Cloudflare)
- Multi cloud experience including virtual networking, load balancing, web application firewall.
- Strong experience with CI/CD.
- Hands-on experience in development and orchestration within high-scale, high-uptime, and high-reliability environments.
- Minimum of six years of hands-on experience in related roles (engineering, DevOps, SRE).
- Familiarity with distributed systems, including queue-first architectures and sharding.
- Demonstrated engineering expertise, including gathering requirements, problem-solving, and making recommendations.
- Preferred: Familiarity with security frameworks, attack vectors, botnets, and impact analysis.
What We Offer
~1 min readLocation & Eligibility
Listing Details
- Posted
- November 8, 2024
- First seen
- September 28, 2026
- Last seen
- September 28, 2026
Posting Health
- Days active
- 0
- Repost count
- 0
- Trust Level
- 25%
- Scored at
- September 28, 2026
Signal breakdown
Browse Similar Jobs
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.