
FactSet creates flexible, open data and software solutions for over 200,000 investment professionals worldwide, providing instant access to financial data and analytics that investors use to make crucial decisions.
At FactSet, our values are the foundation of everything we do. They express how we act and operate, serve as a compass in our decision-making, and play a big role in how we treat each other, our clients, and our communities. We believe that the best ideas can come from anyone, anywhere, at any time, and that curiosity is the key to anticipating our clients’ needs and exceeding their expectations.
We ship AI into workflows where being confidently wrong is expensive. A banker builds a pitchbook from generated tombstones and charts. A portfolio manager acts on AI-generated attribution commentary, and an analyst asks a conversational assistant a question and gets back an answer synthesized across filings, transcripts, estimates, news, and their own firm's internal data. In each case the output looks finished, and it carries FactSet's name into work a client will act on.
When AI gets something wrong in these workflows, the cost is contractual and reputational, and in some cases regulatory. The error does not stay inside our product. It travels into a client's own work product and into the decisions they make from it.
AI has moved from a feature inside a few products to a capability across the platform, and our quality practice needs to scale with it. Teams have built their own test sets and their own working definitions of good, which is how most organizations start. At platform scale, conversational assistance, banker workflows, portfolio commentary, research management, and the APIs clients build on top of each need their own measures of quality, evaluated on shared tooling so the results can be compared and audited in one place.
This role exists to build that function. You will define the quality and evals playbook for AI at FactSet, own the platform that measures it, and make it straightforward for every team to prove the quality of what they build. You are the first product hire into this charter, and you will build the team behind it.
Own the evaluation platform. You are the product leader for FactSet's shared AI evaluation platform, paired with an engineering counterpart who owns the technical build and operation. Together you stand it up and run it as enterprise infrastructure: tracing, offline and online eval harnesses, LLM-as-judge pipelines with calibrated human review, golden dataset management, and regression suites wired into CI. You own the roadmap, the adoption strategy, the integration surface into product teams' existing workflows, and the commercial and roadmap relationship with the platform vendor. Your engineering counterpart owns the technical relationship. You will know it worked when your colleagues across the enterprise run their own evals on it without your team in the loop.
Define the golden pathways. With your engineering counterpart, develop and maintain the golden pathways for evaluation across the enterprise: the documented, opinionated way a team instruments a system, builds a dataset, runs an eval, and wires it into their release process. A product team should be able to follow the paved path without designing an evaluation approach from scratch, and you own keeping those pathways current as practice and tooling change.
Define how quality is measured. Define the quality taxonomy for AI systems across the portfolio: what a failure is, how failures are classified, and which failure classes are non-negotiable, plus the further classes you define with product leadership and with the PMs who own each capability. This becomes the shared language the organization uses to argue about quality, and you are its steward.
Own human review operations. Evaluation at scale depends on people producing labels on a schedule: judge calibration sets, ground-truth datasets, and adjudication of disagreements. You own that operation, including how reviewers are sourced, trained, and measured for consistency, what it costs, and how it scales as coverage grows. Your engineering counterpart builds the tooling those reviewers work in.
Set and hold the floor. You define the minimum every AI capability must satisfy before it reaches a client. The floor is a coverage requirement: a capability ships with evals defined and running, tracing in place, a documented failure taxonomy, and a measured baseline it can be held against later. Product teams set their own quality targets above that line and are expected to aim well above it.
Drive adoption. Enterprise platform mandates fail when the platform is slower than the spreadsheet it replaces. You will land this team by team: understand what each product group does for quality today, help them move onto shared tooling and pathways, and make the shared path the easier one. This is a sustained internal go-to-market effort and a core part of the job.
Close the loop from production. Design the path from a client-reported failure to a permanent regression test. The program succeeds when the same failure cannot ship twice. That matters far more than the number of evals in existence.
Unite the practice. Bring evaluation leaders at FactSet together into a small group of PMs and specialists operating as a federated center of excellence. You own the standard and the platform, and product teams own their own quality outcomes against it.
Helpful, not required: financial services or capital markets domain knowledge; experience with regulated or audited AI deployments; familiarity with model risk management practice.
FactSet (NYSE:FDS | NASDAQ:FDS) helps the financial community to see more, think bigger, and work better. Our digital platform and enterprise solutions deliver financial data, analytics, and open technology to more than 8,200 global clients, including over 200,000 individual users. Clients across the buy-side and sell-side, as well as wealth managers, private equity firms, and corporations, achieve more every day with our comprehensive and connected content, flexible next-generation workflow solutions, and client-centric specialized support. As a member of the S&P 500, we are committed to sustainable growth and have been recognized among the Best Places to Work in 2023 by Glassdoor as a Glassdoor Employees’ Choice Award winner. Learn more at www.factset.com and follow us on X and LinkedIn.
At FactSet, we celebrate difference of thought, experience, and perspective. Qualified applicants will be considered for employment without regard to characteristics protected by law.

FactSet creates flexible, open data and software solutions for tens of thousands of investment professionals around the world, providing instant access to financial data and analytics that investors use to make crucial decisions.
For 40 years, through market changes and technological progress, our focus has always been to provide exceptional client service. From more than 60 offices in 23 countries, we’re all working together toward the goal of creating value for our clients, and we’re proud that 95% of asset managers who use FactSet continue to use FactSet, year after year.
As big as we grow, as far as we reach, and as successful as we become, we stay connected to our clients and to each other.