Hiring.Camp

Director, Product Management AI Quality & Evaluation

Factset

·

Today

Location
India, Hyderabad, DVS, SEZ-1 – Orion B4; FL 7,8,9,11 (Hyderabad - Divyasree 3)
Type
Full-time
Department
IT
Seniority
Director
Source
Workday

Description

FactSet creates flexible, open data and software solutions for over 200,000 investment professionals worldwide, providing instant access to financial data and analytics that investors use to make crucial decisions.  

At FactSet, our values are the foundation of everything we do. They express how we act and operate, serve as a compass in our decision-making, and play a big role in how we treat each other, our clients, and our communities. We believe that the best ideas can come from anyone, anywhere, at any time, and that curiosity is the key to anticipating our clients’ needs and exceeding their expectations.  

We ship AI into workflows where being confidently wrong is expensive. A banker builds a pitchbook from generated tombstones and charts. A portfolio manager acts on AI-generated attribution commentary, and an analyst asks a conversational assistant a question and gets back an answer synthesized across filings, transcripts, estimates, news, and their own firm's internal data. In each case the output looks finished, and it carries FactSet's name into work a client will act on. 

When AI gets something wrong in these workflows, the cost is contractual and reputational, and in some cases regulatory. The error does not stay inside our product. It travels into a client's own work product and into the decisions they make from it. 

AI has moved from a feature inside a few products to a capability across the platform, and our quality practice needs to scale with it. Teams have built their own test sets and their own working definitions of good, which is how most organizations start. At platform scale, conversational assistance, banker workflows, portfolio commentary, research management, and the APIs clients build on top of each need their own measures of quality, evaluated on shared tooling so the results can be compared and audited in one place. 

This role exists to build that function. You will define the quality and evals playbook for AI at FactSet, own the platform that measures it, and make it straightforward for every team to prove the quality of what they build. You are the first product hire into this charter, and you will build the team behind it. 

What you will own 

Own the evaluation platform. You are the product leader for FactSet's shared AI evaluation platform, paired with an engineering counterpart who owns the technical build and operation. Together you stand it up and run it as enterprise infrastructure: tracing, offline and online eval harnesses, LLM-as-judge pipelines with calibrated human review, golden dataset management, and regression suites wired into CI. You own the roadmap, the adoption strategy, the integration surface into product teams' existing workflows, and the commercial and roadmap relationship with the platform vendor. Your engineering counterpart owns the technical relationship. You will know it worked when your colleagues across the enterprise run their own evals on it without your team in the loop. 

Define the golden pathways. With your engineering counterpart, develop and maintain the golden pathways for evaluation across the enterprise: the documented, opinionated way a team instruments a system, builds a dataset, runs an eval, and wires it into their release process. A product team should be able to follow the paved path without designing an evaluation approach from scratch, and you own keeping those pathways current as practice and tooling change. 

Define how quality is measured. Define the quality taxonomy for AI systems across the portfolio: what a failure is, how failures are classified, and which failure classes are non-negotiable, plus the further classes you define with product leadership and with the PMs who own each capability. This becomes the shared language the organization uses to argue about quality, and you are its steward. 

Own human review operations. Evaluation at scale depends on people producing labels on a schedule: judge calibration sets, ground-truth datasets, and adjudication of disagreements. You own that operation, including how reviewers are sourced, trained, and measured for consistency, what it costs, and how it scales as coverage grows. Your engineering counterpart builds the tooling those reviewers work in. 

Set and hold the floor. You define the minimum every AI capability must satisfy before it reaches a client. The floor is a coverage requirement: a capability ships with evals defined and running, tracing in place, a documented failure taxonomy, and a measured baseline it can be held against later. Product teams set their own quality targets above that line and are expected to aim well above it. 

Drive adoption. Enterprise platform mandates fail when the platform is slower than the spreadsheet it replaces. You will land this team by team: understand what each product group does for quality today, help them move onto shared tooling and pathways, and make the shared path the easier one. This is a sustained internal go-to-market effort and a core part of the job. 

Close the loop from production. Design the path from a client-reported failure to a permanent regression test. The program succeeds when the same failure cannot ship twice. That matters far more than the number of evals in existence. 

Unite the practice. Bring evaluation leaders at FactSet together into a small group of PMs and specialists operating as a federated center of excellence. You own the standard and the platform, and product teams own their own quality outcomes against it. 

What we are looking for 

  • 10+ years in product management, with 4+ years on AI/ML or data-intensive products, and experience managing product managers 

  • Demonstrated experience defining evaluation methodology for LLM or ML systems: metrics, failure taxonomies, golden datasets, and quality thresholds that engineering teams actually adopted 

  • Experience taking an internal platform from zero to broad adoption across product teams who were not required to use it, including the migration, enablement, and change-management work that entails 

  • Experience operating a paired product and engineering leadership model, where neither person has authority over the other's function 

  • Hands-on fluency with modern evaluation approaches, including offline and online eval, LLM-as-judge, human-in-the-loop review, RAG and retrieval metrics, and agent trajectory evaluation, enough to design a pipeline and critique someone else's without needing to implement it 

  • Familiarity with LLM observability and evaluation tooling such as LangSmith, LangFuse, Braintrust, or Weights & Biases. We care that you have opinions about this category rather than experience with any one product in it 

  • A track record of standing something up rather than inheriting it. You have built a function, a standard, or a platform from nothing inside a large organization 

  • Comfort holding a bar under commercial pressure, and the judgment to know when the bar is wrong 

  • Ability to influence senior stakeholders across time zones without formal authority over their teams 

  • Strong written communication. This role produces standards documents that people have to be able to follow six months later 

Helpful, not required: financial services or capital markets domain knowledge; experience with regulated or audited AI deployments; familiarity with model risk management practice. 

Company Overview: 

FactSet (NYSE:FDS | NASDAQ:FDS) helps the financial community to see more, think bigger, and work better. Our digital platform and enterprise solutions deliver financial data, analytics, and open technology to more than 8,200 global clients, including over 200,000 individual users. Clients across the buy-side and sell-side, as well as wealth managers, private equity firms, and corporations, achieve more every day with our comprehensive and connected content, flexible next-generation workflow solutions, and client-centric specialized support. As a member of the S&P 500, we are committed to sustainable growth and have been recognized among the Best Places to Work in 2023 by Glassdoor as a Glassdoor Employees’ Choice Award winner. Learn more at www.factset.com and follow us on X and LinkedIn. 

At FactSet, we celebrate difference of thought, experience, and perspective. Qualified applicants will be considered for employment without regard to characteristics protected by law. 

Skills

Risk Management

Similar Jobs

30

Director, Product Management

Connect wise·US-Remote·Remote

2d ago

Director, Product Management

Walmart·Change Building AR Bentonville Home Office, US

3d ago

Director, Product Management

Walmart·Change Building AR Bentonville Home Office, US

3d ago

Director, Product Management

Mastercard·London, England

3d ago

Director Product Management

Humana·Work at Home - Ohio, US·Remote, Hybrid

5d ago

Director, Product Management

Hdsupply·GA250 - Atlanta GA, US

5d ago

Director, Product Management

Curinos·New York, NY

5d ago

Director, Product Management

Verisk Careers | Verisk·US·Hybrid

6d ago

Director, Product Management

Mastercard·Toronto, Canada

6d ago

Director, Product Management

Q2Ebanking·Austin, Texas

1w ago

Director, Product Management

Cognition+

1w ago

Director, Product Management

Shell Recharge Solutions

1w ago

Director, Product Management

Mastercard·New York City, New York +2

1w ago

Director, Product Management

Fortive·Austin, TX

1w ago

Director, Product Management

Fluke·Austin, TX·Onsite

1w ago

Director, Product Management

Paypal·Austin, TX

1w ago

Director, Product Management

PayPal·USA - Texas - Austin - Corp - Alterra Pkwy, US

1w ago

Director, Product Management

Global Appliedsystems·US

1w ago

Director, Product Management

Mastercard·London, England

1w ago

Director, Product Management

Diligent Corporation·Budapest, Hungary +1

2w ago

Director, Product Management

Regeneron·SLEEPY HOLLOW, US

2w ago

Director, Product Management

HIRERIGHT·Nashville, TN

2w ago

Director, Product Management

Gnw·New York, US +1·Remote

2w ago

Director, Product Management

Commerceiq·Bengaluru, Karnataka

2w ago

Director, Product Management

Illumio·Sunnyvale, California - HQ·Onsite

2w ago

Director, Product Management

Understood·New York, NY +1·Hybrid

2w ago

Director, Product Management

Litmos·US +1

2w ago

Director, Product Management

Coupang Internal·Seoul, South Korea

3w ago

Director, Product Management

Mastercard·Lisbon, Portugal

3w ago

Director, Product Management

Mastercard·Kuala Lumpur, Malaysia

3w ago
Director, Product Management AI Quality & Evaluation at Factset | Hiring.Camp