Senior Full-Stack Engineers (React): AI Evaluation Environments

Remote solely

About this role

What We're Researching

We are hiring senior full-stack engineers to help build and refine next-generation coding environments that evaluate AI agents. This work directly supports the development of robust grading systems that accurately measure how effectively an AI agent solves complex coding tasks. The goal is to create secure evaluation suites that cannot be bypassed or easily cheated by the models.

How It Works

Throughout this long-term engagement, you will work remotely with real website clones to identify bugs and write technical specifications. You will build comprehensive automated test suites designed to grade AI-generated code. This involves analyzing how an agent interacts with a given task and ensuring the grading logic is watertight. The process requires deep technical scrutiny and creative problem-solving to account for unpredictable AI behavior.

Who This Is For

We welcome senior full-stack software engineers with extensive hands-on experience in React. You should be highly comfortable writing technical specs and building automated testing frameworks for complex web environments. Candidates with a strong background in web security, QA automation, or AI evaluation are highly encouraged to apply.

What You'll Do

  • Interact with real website clones to identify bugs and workflow gaps

  • Write detailed technical specifications for complex AI coding tasks

  • Build robust automated test suites to securely grade AI agent performance

  • Ensure grading logic is watertight and resistant to AI bypassing or shortcuts

Who Should Apply

  • Senior-level experience in full-stack software engineering

  • Deep professional expertise building web applications with React

  • Strong background in automated testing and technical specification writing

  • Available for an ongoing, high-commitment technical engagement

Compensation

$75 per hour

Ready to participate?

Start your paid interview now

About Terac

Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.

Learn more at terac.com or on YouTube at @jointerac.

Company at a glance

Terac is an AI-native two-sided marketplace headquartered in downtown San Francisco that connects enterprises with specialized expertise to high-quality professionals on demand. The company's mission is to fundamentally change how labor is allocated globally by using AI-driven matchmaking to connect experts with work opportunities based on thousands of signals about capability, performance, preferences, and potential. Currently placing thousands of experts per month across 50+ projects for 25+ clients, Terac serves leading AI research companies, market research firms, data companies training AI models, and research organizations. With a panel of 100,000+ professionals and ambitions to scale to 10 million, the company has achieved an eight-figure annual run rate and expects to exceed $10 million in revenue in the next year. Having raised $9 million in seed funding, Terac is preparing to launch a Series A round while rapidly expanding its team from 4-5 people to approximately 15 members, with plans to reach 30 by year end.

Founded2024
Team Size11-50
WorkspaceRemote solely
IndustryInternet Marketplace Platforms
Location
United States
Websiteterac.com
LinkedInLinkedIn

Tired of cold applications?

Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.

Know someone who'd be great for this?