← Back to all opportunities

Evaluating Moral Competence in AI Systems

Society Ethics Tech / Data x Direction

Supervision
Violet Gordon, with Jesse Parent
Term
Mid-September through the first week of December 2026, with the opportunity to continue on the project afterward
Applications
First review September 7. Later proposals considered as capacity allows; there is no fixed closing date.
Compensation
Unpaid. See below.

About this position

This position continues recent work from Violet Gordon within our Data x Direction program: developing ways to actually examine AI morality, rather than take it on faith from how a system's outputs read.

What we are looking for is the question you would pursue and how you would go about it.

Background

Julia Haas and colleagues at Google DeepMind argue, in a 2026 Nature roadmap, that evaluating AI morality has to move past moral performance — whether a system's outputs look right — and toward moral competence: whether those outputs are produced because of morally relevant considerations.

The distinction matters because these systems are trained on enormous amounts of human writing already full of moral reasoning. Fluent, appropriate-sounding moral output is exactly what you would expect whether or not anything underneath is reasoning about it.

This is an open call for people interested in that question: what would count as evidence of moral competence in an AI system, and how would you go about finding it. We are not looking for a specific method or a predetermined answer. A proposal could be empirical, interpretive, or conceptual — drawing on interpretability and mechanistic analysis, instruments from moral psychology, evaluation and benchmark design, or philosophical work on what would count as evidence either way. Bring the approach that fits your training and interest, as long as it engages the distinction directly.

In scope: empirical, interpretive, or conceptual work that engages the performance and competence distinction directly.

Not in scope: general debate about whether AI is or could be conscious, sentient, or morally considerable in itself. That is a related question, but a different one from the one this roadmap asks.

Who this is for

Graduate students, postgraduate researchers, early career researchers, and undergraduates prepared to work at that level. Backgrounds in cognitive science, philosophy, machine learning or interpretability, or moral psychology are all relevant, and no single one is required.

What to submit

A proposal of 500 to 1000 words covering four things:

  1. The question and the claim. What you would work on, and what a claim about it would look like that you could actually be wrong about.
  2. Your approach. What you would do, and what you would try first to find out whether it is tractable at all.
  3. A timeline for the term, from mid-September to the first week of December, in three or four phases rather than week by week, with a midpoint you could be held to.
  4. What you want from the term, including whether you would want to work closely with Violet Gordon or run more independently.

Also submit a CV. A letter of recommendation is optional and there is a slot for one. Not including a letter will not count against your proposal.

Compensation and credit

These positions are unpaid. We do not offer stipends.

We will work with your university to arrange experiential learning or internship credit where your program allows it, and we welcome applicants who bring their own funding or fellowship support. If either applies to you, say so in your proposal and we will work out the arrangement.

How proposals are reviewed

We read for whether the claim is genuinely arguable, whether the scope is achievable within the term, whether the approach shows real familiarity with the material, and whether your own interest in the question is legible.

Every proposal is read in full by a person. Nothing in this process runs through keyword matching, automated ranking, or AI screening of any kind, and there are no recorded interviews without a person present. If we interview you, you will be talking with someone who can answer your questions.

Details

  • Duration 12–14 weeks, mid-September to early December 2026, with the opportunity to continue on the project afterward
  • Compensation Unpaid. University credit and outside funding supported.
  • Location Remote
  • Level Graduate, postgraduate, and early career researchers; prepared undergraduates considered

Apply

Cover the four items under “What to submit” above: the question and the claim, your approach, a timeline for the term, and what you want from it.

Optional. Not including one will not count against your proposal.

Or email start@jopro.org with your proposal and CV.