Senior Staff Research Engineer – Reinforcement Learning for AI Agents
Xpengmotors
·Mar 19, 2026
- Salary
- $244k – $413k
- Location
- Santa Clara, CA · Mountain View, California, United States
- Department
- AI Infrastructure Team
- Seniority
- Senior
- Education
- PhD
- Source
- Greenhouse
Description
Key Responsibilities:
-
Reinforcement learning methods for LLM-driven agents and decision systems.
-
Policy optimization for long-horizon reasoning and planning.
-
Learning from human or AI feedback (RLHF / RLAIF).
-
Agent training pipelines built on top of our agent infrastructure platform.
-
Evaluation and benchmarking systems for agent capabilities.
-
Learning loops that integrate real-world and simulation data.
-
Contribute to AI systems that continuously improve after deployment.
Basic Qualifications
-
MS or PhD in Computer Science, AI, Machine Learning, Robotics, or a related field.
-
Strong background in reinforcement learning or machine learning.
-
Experience implementing RL algorithms such as PPO, Actor-Critic, or policy gradient methods.
-
Strong programming skills in Python with PyTorch or JAX.
-
Experience building ML training systems or infrastructure.
Preferred Qualifications
-
Experience with RLHF or preference learning.
-
Experience with LLM agents or tool-using AI systems.
-
Multi-agent systems or long-horizon planning.
-
Simulation environments for RL.
-
Publications in NeurIPS, ICML, ICLR, ACL, or related venues.
-
A fun, supportive and engaging environment.
-
Opportunity to make significant impact on transportation revolution by the means of advancing autonomous driving.
-
Opportunity to work on cutting edge technologies with the top talent in the field.
-
Competitive compensation package.
-
Snacks, lunches and fun activities.