Hiring.Camp

Senior AI Research Engineer

Emergent

·

May 17, 2026

Location
Bengaluru, Karnataka, IN
Type
Full-time
Department
Engineering
Seniority
Senior
Source
Y Combinator

Description

Emergent builds autonomous coding agents that replace traditional software development by generating, testing, and deploying production applications directly from plain-language intent. Our systems run in production at global scale and are used to build millions of real applications.

Since our public launch, we've crossed $130M in Annualised Revenue and grown to over 10M users across 190+ countries , who have built 12M+ applications on Emergent. We're backed by Creaegis, Khosla Ventures, SoftBank, Lightspeed, Together, Y Combinator, Google, Claypond and Sentinel Global.

We're solving the hard part of AI-driven software creation: correctness, reliability, security, and scale in real production systems. The team is built by repeat founders, Olympiad medalists, IIT & IIM alumni, and leaders from Google, Amazon, and Dropbox.

We're hiring builders who want ownership, speed, and impact at global scale.

The Role:
We're looking for Senior AI Research Engineer to characterize, measure, and advance the capabilities of our coding agents. You will turn ambiguous notions of "agent quality" into clear, defensible metrics that the team, leadership, and the field can rely on, and you will use those metrics to drive both incremental wins and moonshots in agent performance.

This is a deep-work role at the intersection of agent behavior, evaluation research, and applied training. You will define what good looks like for long-horizon coding agents, build the evaluation dataset and methodology that produces those signals, mine production data for failure modes most teams never see, and run targeted training, fine-tuning, RL, memory, and prompt-optimization experiments that translate research advances into shipped improvements. You will operate with strong independence, make hard calls in inherently subjective and probabilistic systems, and own outcomes end-to-end. If you treat models as objects of study rather than black boxes, take pride in moving benchmark numbers with rigor, and want to apply the frontier of agent research at the scale of millions of real applications, this is your role.

What You'll Do:

  • Architect the next version of the Emergent agent: shape the core architecture and make the foundational design choices that define how the agent thinks, learns, and improves over time
  • Characterize agent behavior at depth: develop a deep, evidence-grounded understanding of how the agent succeeds and fails across the full range of real-world usage, and convert that understanding into rigorous, quantitative measurement
  • Design and ship evaluations across reasoning, planning, tool use, code correctness, long-horizon execution, security, and agent reliability: define the metric, build the dataset, validate against known signals, and ship dashboards that make regressions impossible to miss
  • Drive step-function gains: take on the ambitious bets that meaningfully advance the state of the art, the 10-point leaps on hard capabilities, not incremental polish. Pick the problems where the upside is large and the path is uncertain
  • Climb public benchmarks: move the needle on SWE-bench Pro, Terminal-Bench, and other industry-standard benchmarks the field uses to grade coding agents
  • Run training and post-training experiments, including supervised fine-tuning, RLHF/RLAIF, DPO, distillation, reward modeling, prompt optimization, and judge-model calibration, against production-grounded objectives
  • Own end-to-end: carry work from hypothesis through experiment design, execution, analysis, decision, rollout, and post-launch measurement. Read research papers deeply, get inspired ideas, and turn them into shipped outcomes
  • Make hard calls in subjective systems: decide when a regression is real, when a win is noise, when a benchmark is overfit, when to ship despite mixed signals, and when to kill a promising direction. Communicate the reasoning crisply

Who You Are:

  • 5 to 8 years of AI experience, with meaningful time spent either training and fine-tuning models, or designing rigorous evaluations and measurement systems for them. Both paths are equally valued for this role
  • Hands-on with the modern AI stack and fluent in Python (Go a plus) for research workflows: training pipelines, eval harnesses, data processing, and statistical analysis. Comfortable with transformers, RLHF/DPO/RL for agents, eval frameworks (Inspect, lm-eval-harness, or equivalent), prompt optimization, judge models, and agent frameworks. You pick up new tooling in days
  • Take pride in numbers that move. You measure first, opine second. You can defend why a benchmark is the right benchmark, why a metric isn't gameable, and why a result is statistically real
  • Comfortable in subjective, probabilistic systems. You reason about noise floors, confounds, distribution shift, judge bias, and selection effects without flinching. You know when to trust a number and when to suspect it
  • Enjoy going deep into the long tail. Sifting through large volumes of agent behavior to find the rare, hidden failure mode energizes you, not drains you
  • Understand models like friends. You have intuitions about how a model will behave on a new task before running it, and you update those intuitions when reality disagrees. You know what came out last week, why it matters, and which paper from two years ago is suddenly relevant again
  • Independent operator with leadership presence. You scope your own work, push back on weak ideas (including your manager's), and bring others along through clarity and conviction rather than consensus-seeking
  • Ship fast without compromising rigor. You know which corners are safe to cut and which are load-bearing. Bias toward velocity, but never at the cost of honest measurement

Benefits and Perks:

  1. Daily Meals: Lunch and Dinner provided
  2. Family Insurance: 5 Lakhs worth of coverage for you and your family
  3. Unlimited Paid Time Off: Take the time you need to recharge and come back refreshed
  4. Flexible Working Hours: Work arrangements that fit your life and commitments

Let's build the future of software together.

Skills

Python

Similar Jobs

30

Senior Specialist, AI Research

AIA Careers · CN-OCG International Center, Cheng Du, China

Yesterday

Senior AI/ML Research Engineer

Intuitive Surgical · Sunnyvale, CA, United States

3 days ago

Senior Equity Research Analyst, AI Platform

Versant · Englewood Cliffs, NEW JERSEY, United States · Hybrid

4 days ago

Senior Localization Engineer / Research Engineer, Robotics AI (Visual Navigation)

Grab · Singapore, Singapore · Onsite

4 days ago

Senior Applied AI Research Engineer

Accesa · Employees can work remotely, Romania · Remote

6 days ago

Senior AI Research Engineer

motorolasolutions · Bangalore, India (ZIN122)

6 days ago

Senior AI Research Engineer

Motorola Solutions · Bangalore, India (ZIN122)

6 days ago

Senior AI Research Scientist- Time-Series Foundational Models

Robert Bosch · Sunnyvale, CA, United States · Hybrid

6 days ago

Sr Director, Analyst - Tech CEO Business and Strategy Advisor on AI Orchestration (Tech CEO Research Organization)

Gartner · Remote - Texas, United States of America · Remote

1 week ago

Senior Solutions Architect, Higher Education and Research, Multimodal and Physical AI

Nvidia · Germany, Munich

1 week ago

Senior Solutions Architect, Higher Education and Research, AI and HPC

Nvidia · Munich, BY,DE, DE +1 · Remote

1 week ago

Senior Solutions Architect, Higher Education and Research, Multimodal and Physical AI

Nvidia · Munich, BY,DE, DE

1 week ago

Senior Solutions Architect, Higher Education and Research, Multimodal and Physical AI

Nvidia · Zürich, ZH,CH, CH

1 week ago

Senior Solutions Architect, Higher Education and Research, Multimodal and Physical AI

Nvidia · GB +1 · Remote

1 week ago

Senior Solutions Architect, Higher Education and Research, AI and HPC

Nvidia · Germany, Munich +1 · Remote

1 week ago

Senior Solutions Architect, Higher Education and Research, Multimodal and Physical AI

Nvidia · Switzerland, Zurich

1 week ago

Senior Engineer, Architecture & Performance Research Engineer for Data Center and Agentic AI CPU

Samsung Semiconductor · San Jose, California, United States · Onsite

2 weeks ago

Software Engineer AI Research (Mid-Senior)

Nexthink · Lausanne, VD, Switzerland · Hybrid

3 weeks ago

Senior AI Research Scientist (Model-based RL)

Phaidra · Remote - United Kingdom +1 · Remote

3 weeks ago

Senior Quantitative Equity Research Analyst, AI Platform

Versant · Englewood Cliffs, NEW JERSEY, United States · Hybrid

3 weeks ago

Senior Research Scientist- Robotics AI

Robert Bosch · Sunnyvale, CA, United States · Hybrid

3 weeks ago

Senior Expert for Generative and Multimodal AI to lead Advanced AI research and Innovation

Robert Bosch · bengaluru, India

1 month ago

Senior AI/ML Engineer - Research Data AI and Predictive Modeling (Vaccine R&D)

Pfizer · USA - New York - Pearl River, United States of America

1 month ago

Senior AI/ML Engineer - Research Data AI and Predictive Modeling (Vaccine R&D)

Pfizer · USA - New York - Pearl River, United States of America

1 month ago

Senior AI/ML Engineer - Research Data AI and Predictive Modeling (Vaccine R&D)

Pfizer · USA - New York - Pearl River, United States of America

1 month ago

Sr. Director, UX Research - AI

ServiceNow · Santa Clara, CALIFORNIA, United States · Hybrid

1 month ago

Sr. Data Scientist - AI Research

Lyrahealth · United States · Remote

1 month ago

Senior Manager/ Director, Technical Program Management - AI Research

Salesforce · California - San Francisco, United States of America · Onsite

1 month ago

Senior Member of Technical Staff - AI Research

Salesforce · California - Palo Alto, United States of America · Onsite

1 month ago

Senior AI Research Engineer, Algorithms

Samsungresearchamerica · 665 Clyde Avenue, Mountain View, CA, USA

1 month ago