Hiring.Camp

Research Engineer, Reinforcement Learning

techire ai

·

Jan 6, 2026

Location
San Francisco, CA
Workplace
Onsite
Type
Full-time
Department
IT
Closing date
Today
Source
Vincere

Description

Want to build the large-scale RL environments frontier labs use to train agents that can truly reason and act?

This team are creating complex reinforcement learning environments — simulations where advanced agents learn to plan, adapt, and solve multi-step problems that stretch beyond standard benchmarks. The focus isn’t on training the models themselves, but on building the worlds that make meaningful learning and evaluation possible — the foundation for more capable, aligned systems.

You’ll work end-to-end across environment design, reward dynamics, and scalable simulation — developing the feedback loops that define what “good” looks like for intelligent behaviour. It’s open-ended, research-driven work where the task definition, data, and reward structure are often the hardest and most important problems to solve.

You’ll collaborate closely with researchers tackling unsolved challenges in reinforcement learning and agent behaviour, shaping experiments, scaling infrastructure, and refining how agents learn in the loop.

It suits someone with strong ML and RL experience, deep intuition for agent dynamics, and the curiosity to explore problems that don’t come with clear instructions.

On-site in San Francisco. Compensation up to $300 K base (negotiable, depending on experience) plus equity.

If you want to help build the environments that teach the next generation of AI systems how to think, act, and adapt — we’d love to hear from you.

All applicants will receive a response.

Skills

GoMachine LearningComputer Vision

Similar Jobs

8

Research Engineer (Reinforcement Learning)

Livekit·NAMER +2·Remote

1mo ago

Staff Reinforcement Learning Research Engineer

Bostondynamics·Waltham Office, US

1mo ago

Research Engineer, Chip Design RL (Reinforcement Learning)

Anthropic·San Francisco, CA +1·Onsite

2mo ago

Research Engineer, Code RL (Reinforcement Learning)

Anthropic·San Francisco, CA +1·Onsite

3mo ago

Research Engineer, Machine Learning (Reinforcement Learning)

Anthropic·London, UK +1·Onsite

7mo ago

AI Research Engineer - Reinforcement Learning

Helsing·Berlin, Munich +4

1y+ ago

Research Engineer, Machine Learning (Reinforcement Learning)

Anthropic·San Francisco, CA +2·Onsite

1y+ ago

Research Engineer - Reinforcement Learning

Primeintellect·San Francisco +1·Remote

1y+ ago