RL Environment Engineers: Building Tool Gyms For Knowledge Work

Remote solely

About this role

What We're Researching

We're hiring an RL environment engineer to help build tool-based environments, commonly known as tool gyms, for knowledge work applications. This paid engagement focuses on creating robust simulation frameworks where models can learn to interact with complex software tools. The resulting environments will directly support advanced reinforcement learning research and training pipelines.

How It Works

You will spend approximately 20 hours per week developing and testing new tool environments. This involves writing code to simulate various knowledge work tasks, ensuring realistic and stable agent interactions. You will also participate in remote video check-ins to discuss architecture decisions and troubleshoot implementation blockers. Throughout the project, you will iterate on environment designs based on model performance and technical feedback.

Who This Is For

We are looking for specialized machine learning professionals with hands-on experience in reinforcement learning environments. We welcome RL engineers, AI researchers, simulation developers, and machine learning infrastructure specialists. Ideal candidates have previously built or maintained custom gyms and are comfortable committing to a part-time weekly schedule.

What You'll Do

  • Design and implement tool-based environments for knowledge work simulations.

  • Write clean and modular code to support reinforcement learning training pipelines.

  • Troubleshoot and refine environment mechanics based on testing feedback.

  • Collaborate asynchronously and participate in remote progress check-ins.

Who Should Apply

  • Professional experience as a machine learning or reinforcement learning engineer

  • Hands-on background building custom RL environments or tool gyms

  • Ability to commit to approximately 20 hours of work per week

  • Comfortable discussing technical architecture and implementation strategies

Compensation

$120 per hour

Ready to participate?

Start your paid interview now

About Terac

Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.

Learn more at terac.com or on YouTube at @jointerac.

Company at a glance

Terac is an AI-native two-sided marketplace headquartered in downtown San Francisco that connects enterprises with specialized expertise to high-quality professionals on demand. The company's mission is to fundamentally change how labor is allocated globally by using AI-driven matchmaking to connect experts with work opportunities based on thousands of signals about capability, performance, preferences, and potential. Currently placing thousands of experts per month across 50+ projects for 25+ clients, Terac serves leading AI research companies, market research firms, data companies training AI models, and research organizations. With a panel of 100,000+ professionals and ambitions to scale to 10 million, the company has achieved an eight-figure annual run rate and expects to exceed $10 million in revenue in the next year. Having raised $9 million in seed funding, Terac is preparing to launch a Series A round while rapidly expanding its team from 4-5 people to approximately 15 members, with plans to reach 30 by year end.

Founded2024
Team Size11-50
WorkspaceRemote solely
IndustryInternet Marketplace Platforms
Location
United States
Websiteterac.com
LinkedInLinkedIn

Tired of cold applications?

Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.

Know someone who'd be great for this?