Clera home
·Dashboard

Jobs at Elorian AI (Now Hiring) — 2 open

Elorian AI

Reinforcement Learning Infrastructure Engineer

Palo Alto, California, United States · Hybrid

$275k–$475k/yr

Mid levelVisa sponsorship

About Us We are a well-funded, early-stage AI lab focused on building the next generation of frontier multimodal AI models. Founded by former DeepMind researchers, including Andrew Dai, who was previously a leader on Gem…

Skills: Distributed Systems, Reinforcement Learning, Python, PyTorch, JAX

Elorian AI

Inference Infrastructure Engineer, Serving

Palo Alto, California, United States · Hybrid

$275k–$475k/yr

Mid levelVisa sponsorship

About Us We are a well-funded, early-stage AI lab focused on building the next generation of frontier multimodal AI models. Founded by former DeepMind researchers, including Andrew Dai, who was previously a leader on Gem…

Skills: Inference Serving Systems, Quantization, Batching, Speculative Decoding, KV Cache Management

Reinforcement Learning Infrastructure Engineer

Elorian AI

Palo Alto, California, United States • Hybrid

Apply
Mid levelVisa sponsorship

Tired of cold applications?

Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.

  • $275k–$475k/yr
  • Full-time
  • Health Insurance, Dental Insurance, Vision Insurance, Unlimited PTO, Paid Parental Leave, Relocation Support
  • Visa sponsorship available
  • Posted 4d ago
  • ~40 hrs/week

Responsibilities

Design and optimize the core infrastructure for large-scale reinforcement learning and post-training workloads. Build actor-learner architectures and monitoring tools to ensure high throughput and reliability for multimodal AI models.

Requirements

Requires 3+ years of distributed systems experience with a focus on RL training pipelines and multi-node GPU orchestration. Proficiency in Python, PyTorch or JAX, and experience with async training infrastructure is essential.

Full job description

About Us

We are a well-funded, early-stage AI lab focused on building the next generation of frontier multimodal AI models. Founded by former DeepMind researchers, including Andrew Dai, who was previously a leader on Gemini. Our team currently consists of 20 world-class scientists and engineers. We recently raised $55M in seed funding from Striker Ventures, Menlo Ventures, Altimeter Capital, and NVIDIA. We are tackling some of the hardest problems in artificial intelligence, and we are growing fast.

The Role

We're looking for an infrastructure engineer to design and build the core systems behind how we train our models with reinforcement learning (RL).

You'll own the training infrastructure end to end, from rollout and reward pipelines to orchestration, reliability, and observability. The work spans both the algorithmic side of RL and the systems reality of running distributed training at scale, and you'll partner closely with our research team to keep RL training fast, stable, and dependable for the multimodal, visual reasoning models at the center of our work.


What You Will Do

  • Design, build, and optimize the infrastructure that powers our large-scale RL and post-training workloads

  • Improve the reliability, scalability, and throughput of distributed RL training pipelines

  • Build actor-learner architectures and orchestrate environment rollouts at scale

  • Develop monitoring and observability tools that ensure high uptime, debuggability, and reproducibility across RL systems

  • Collaborate with researchers to translate algorithmic ideas into production-grade training pipelines

  • Improve GPU utilization and training throughput across the cluster

What We're Looking For

Minimum qualifications:

  • 3+ years of distributed systems experience, including building or optimizing large-scale RL training pipelines (PPO, GRPO, or similar on-policy methods)

  • Experience with actor-learner architectures and environment rollout orchestration at scale

  • Strong Python skills, plus PyTorch or JAX

  • Experience with async training infrastructure, replay buffers, or simulation-based environment frameworks

  • Multi-node GPU orchestration experience (Ray, SLURM, or Kubernetes)

  • A track record of improving training throughput and GPU utilization at scale

  • Strong engineering skills; ability to contribute performant, maintainable code and debug in complex codebases

Preferred qualifications (strong candidates may have some, not all):

  • Experience with multimodal or agentic RL environments

  • Experience with RLHF or reward modeling pipelines

  • A self-directed builder who moves quickly and works across teams in an early-stage setting

Logistics

Location: This role is based on-site in Palo Alto, California.

Compensation: Depending on background, skills, and experience, the expected annual base salary range for this position is $275,000 - $475,000 USD, plus equity and benefits.

Visa sponsorship: We sponsor work visas. We can't promise every case will succeed, but for the right person we'll work through the process with you.

Benefits: We offer health, dental, and vision benefits, unlimited PTO, paid parental leave, and relocation support as needed.

Elorian AI is an equal opportunity employer. We are committed to building a diverse team and inclusive environment.

Related keywords

Reinforcement LearningDistributed SystemsPPOGRPOPyTorchJAXRaySLURMKubernetesGPU UtilizationMultimodal AIRLHFReward ModelingActor-Learner ArchitecturePythonPost-training

About Elorian AI

LinkedInVisit site

Building Next Generation Multimodal Intelligence

Industry
Software Development
Company size
11-50 employees
Founded
2025
Headquarters
Palo Alto, California
LinkedIn followers
1,494

We are an early-stage AI lab focused on building the next generation of frontier multimodal AI models. Founded by former Google Brain & DeepMind researchers, including Andrew Dai, who was previously a leader on Gemini. Our team consists of world-class scientists and engineers. We recently raised $55M in funding from Striker Ventures, Menlo Ventures, and Altimeter with participation from 49 Palms and others with prominent AI researchers including Jeff Dean. We are tackling some of the hardest problems in artificial intelligence, and we are growing fast.

Offices: Palo Alto, California 94301, US

View all jobs at Elorian AI

About Elorian AI

LinkedInVisit site

Building Next Generation Multimodal Intelligence

Industry
Software Development
Company size
11-50 employees
Founded
2025
Headquarters
Palo Alto, California
LinkedIn followers
1,494

We are an early-stage AI lab focused on building the next generation of frontier multimodal AI models. Founded by former Google Brain & DeepMind researchers, including Andrew Dai, who was previously a leader on Gemini. Our team consists of world-class scientists and engineers. We recently raised $55M in funding from Striker Ventures, Menlo Ventures, and Altimeter with participation from 49 Palms and others with prominent AI researchers including Jeff Dean. We are tackling some of the hardest problems in artificial intelligence, and we are growing fast.

Offices: Palo Alto, California 94301, US

View all jobs at Elorian AI

Similar companies hiring

Amazon (10791)Bosch (3578)Google (3474)Prolific (3433)AgileEngine (2991)Transport AI (1739)Booz Allen Hamilton (1546)Speechify (1529)Microsoft (1527)Salesforce (1012)BJAK (956)Cisco (938)
Clera home

Your AI-talent agent. Connecting talents with dream jobs.

Earn $5,000

Tools

  • Salary Calculator
  • Resume Review
  • Startup Map

Explore

  • Jobs
  • Discover Jobs
  • Companies
  • Referral

Platform

  • Pricing
  • Integrations
  • Partners
  • Acquihire

Clera

  • Manifesto
  • Engineering
  • We are hiring!
  • FAQs
  • Blog
  • Press

Tools

  • Salary Calculator
  • Resume Review
  • Startup Map

Explore

  • Jobs
  • Discover Jobs
  • Companies
  • Referral

Platform

  • Pricing
  • Integrations
  • Partners
  • Acquihire

Clera

  • Manifesto
  • Engineering
  • We are hiring!
  • FAQs
  • Blog
  • Press

© 2026 Clera Labs, Inc.

PrivacyTermsBug Bounty