bespokelabs
New

RL Environments Engineer

United StatesUnited States·Mountain Viewfull-timemid
OtherEngineer
3 views0 saves0 applied

Quick Summary

Overview

About Bespoke Labs Bespoke Labs is an applied AI research lab pioneering data and RL environment curation for training and evaluating agents. Recently, we curated Open Thoughts,

Technical Tools
OtherEngineer

Bespoke Labs is an applied AI research lab pioneering data and RL environment curation for training and evaluating agents.

Recently, we curated Open Thoughts, one of the best open reasoning datasets used by multiple frontier labs, trained SOTA specialized models such as Bespoke-MiniChart-7B and Bespoke-MiniCheck, and taught agents to do multi-turn tool-calling with reinforcement learning.

Bespoke is uniquely positioned to capture a large market share of data and RL environment curation.

 

About the Role

~1 min read

This is a delivery role. We want an engineer who has built the machinery that turns environment ideas into validated agentic coding tasks.

You will not be studying environments in the abstract. You will build the pipelines that produce them, design the complex coding worlds agents train inside, and keep pushing throughput: more environments, higher quality, less manual work per task. We will measure you on the volume and quality of environments you ship, not on papers.

The thing we care about most is whether you have done this before. If you have stood up an environment-generation pipeline, scaled agentic task creation into the hundreds or thousands, and shipped it, we want to talk.

 

Responsibilities

~1 min read
  • You have built pipelines that produce RL environments or agentic tasks, and shipped them. Show us the volume you personally drove.

  • Experience scaling task or environment creation into the hundreds or thousands through automation rather than manual effort.

  • A record of throughput. You judge yourself by what you ship, and you keep making the next unit cheaper to produce.

  • Strong software engineering fundamentals and fluency in Python.

  • Real experience with production software: large codebases, real conventions, build systems, testing, devops, SRE, diagnosis and RCA.

  • A good sense of what frontier coding agents can and cannot do, and where they cut corners.

  • You can build the infrastructure behind scaled production: pipelines, automation, grading and verification systems, sandboxed execution.

  • Experience running workloads at scale on GCP.

  • A tool-builder's instinct. You automate repetitive work and unblock the people around you.

Nice to Have

~1 min read
  • Hands-on experience with RL training systems, post-training, verifiers, or tool-use harnesses.

  • Background in developer tooling, CI/CD sandboxes, or code-execution infrastructure.

  • Experience working in the review of task creation process for a benchmark such as Terminal Bench 3.0

Location: Mountain View, CA.

Compensation: Competitive salary and equity based on experience and background

Benefits: Health coverage, lunch, flexible work arrangements, and the opportunity to shape how the AI community evaluates and trains agents

We encourage applications from candidates with diverse research backgrounds. If you're passionate about understanding agent behavior and creating systematic approaches to environment design, we'd love to hear from you.

 

Location & Eligibility

Where is the job
Mountain View, United States
On-site at the office
Who can apply
US

Listing Details

Posted
August 21, 2026
First seen
August 21, 2026
Last seen
August 21, 2026

Posting Health

Days active
0
Repost count
0
Trust Level
52%
Scored at
August 21, 2026

Signal breakdown

freshnesssource trustcontent trustemployer trust
Newsletter

Stay ahead of the market

Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.

A
B
C
D
Join 12,000+ marketers

No spam. Unsubscribe at any time.

bespokelabsRL Environments Engineer