$140K – $225K • Offers Equity • Offers Bonus/yr

Software Engineer (Backend-Focused)

United StatesUnited States·SeattleRemotefull-timemid
Software EngineerSoftware Engineering
0 views0 saves0 applied

Quick Summary

Key Responsibilities

Build and maintain backend services for our LLM gateway — routing, rate limiting, key management, and observability in front of the inference fleet.

Requirements Summary

4+ years of experience in backend engineering fundamentals: distributed systems, API design, and production experience in Go, Rust, or async Python.

Technical Tools
Software EngineerSoftware Engineering

Our mission is to accelerate positive impact in critical industries through AI transformation. We specialize in physics-informed ML and enterprise AI solutions that directly address climate and sustainability challenges.

We’re growing quickly and already work with category-leaders in real estate (CBRE), energy (LevelTen Energy), logistics (Flexe) and utilities.

We’re a public benefit corporation, founded in 2024, and have been profitable from inception.

We work on challenges in clean energy, decarbonization, climate risk, energy systems, and global economics. We’re building our company for long-term success and aim to create the ultimate place to work for those passionate about AI and making a positive impact.

About the Role

~1 min read

We're looking for a Staff or Senior ML Engineer to own the technical backbone of how AZX serves and evaluates models at scale. This is a high-leverage IC role spanning our inference platform — GPU scheduling, autoscaling, and serving infrastructure for vLLM/SGLang across cloud and customer-managed clusters — and the evaluation systems that tell us whether model, prompt, and agent changes that make things better.

You'll create technical direction for how AZX serves models reliably. This role suits someone who wants architectural ownership over hard ML infrastructure problems, paired with the judgment to build the guardrails that let the rest of the team move fast safely.

Responsibilities

~1 min read
  • Build and maintain backend services for our LLM gateway — routing, rate limiting, key management, and observability in front of the inference fleet.

  • Contribute to sandboxing and isolation infrastructure that keeps agent-generated code safe to execute, working alongside our security-focused engineers.

  • Support Kubernetes-based platform services, including operators and autoscaling logic adjacent to our inference platform.

  • Write high-performance backend code in Go, Rust, or async Python (FastAPI/Starlette), working with infrastructure like Envoy and gRPC.

  • Instrument services with OpenTelemetry so behavior, latency, and cost stay observable as the platform scales.

  • Collaborate across the gateway, sandbox, and inference platform teams, flexing across areas as priorities shift.

Requirements

~1 min read
  • 4+ years of experience in backend engineering fundamentals: distributed systems, API design, and production experience in Go, Rust, or async Python.

  • Familiarity with LLM-specific backend concerns (rate limiting, caching, token accounting) is a plus, though not required on day one.

  • Exposure to Kubernetes and containerization; interest in sandboxing or security is a plus.

  • Comfort working across a range of platform concerns rather than one narrow specialty — this role is intentionally broader than our specialist infra profiles.

  • Eagerness to grow into deeper specialization in gateway, sandbox, or inference infrastructure over time.

  • Experience in both startup and enterprise environments

  • Past work in energy, real estate, utilities, climate or related fields

  • Bonus if you have experience and passion in one or more of

  • Additional web frameworks (e.g. Svelte, Vue, Angular)

  • Lower-level languages e.g. C++, Rust

  • Networking paradigms e.g. GraphQL, Websockets

  • ML capabilities e.g. Sk-learn, xgboost, Pytorch/Tensorflow/JAX, Onnx…

  • Additional database types such as graph or vector databases

  • DevOps e.g. CI/CD pipelines, Docker, Kubernetes, Terraform, Pulumi and/or Bicep

  • Generative AI e.g. prompt engineering, RAG, fine-tuning, tooling ecosystem

  • Be part of a fast-growing, profitable, mission-driven company with industry-leading clients tackling the massive opportunity of AI transformation in critical industries.

  • Competitive early-stage startup compensation (based on capabilities, experience, and location)

  • Bonus eligibility

  • Health insurance with meaningful coverage for dependents

  • Flexible paid time off

  • Equity

  • Fully remote culture with a cluster of teammates in Seattle

 
  • Must be willing to travel to Seattle area for final interview and travel 2x/year for company summits

  • We are unable to sponsor or take over sponsorship of employment visas at this time.

  •  

    If this job sounds like a great fit but don’t check ALL of these qualification boxes, we’d still love to hear from you!

    Location & Eligibility

    Where is the job
    Seattle, United States
    Remote within one country
    Who can apply
    US

    Listing Details

    Posted
    August 25, 2026
    First seen
    August 25, 2026
    Last seen
    August 27, 2026

    Posting Health

    Days active
    0
    Repost count
    0
    Trust Level
    72%
    Scored at
    August 25, 2026

    Signal breakdown

    freshnesssource trustcontent trustemployer trust
    Newsletter

    Stay ahead of the market

    Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.

    A
    B
    C
    D
    Join 12,000+ marketers

    No spam. Unsubscribe at any time.

    careers.azx.ioSoftware Engineer (Backend-Focused)$140K – $225K • Offers Equity • Offers Bonus