AI safety prizes

By OscarD🔸 @ 2026-07-31T20:53 (+6)

This is a linkpost to https://oscardelaney.substack.com/p/ai-safety-prizes

Rather than paying up front for AI safety research (push funding), perhaps we should pay after the fact for the work that made the most progress (pull funding). This way, you only pay for work that was actually valuable.[1] When we know what the target is, but not how to get there or who is best placed to solve the problem, a prize is a useful incentive structure.

Benefits of prizes

Prizes work well to incentivise innovation when:

Historically, prizes have worked well in e.g. DARPA’s autonomous vehicle challenges to source diverse talent into the field.

Costs of prizes

There are also some risks and downsides of prizes to be aware of:

AI-specific considerations

Prizes may be an especially good fit for the new wave of AI philanthropists to fund, because:

Practically, what prizes could we establish?[3]

If we were to launch one or more prizes, getting good publicity would be key. For that, the prizes should be about fun/interesting problems (as with the Jane Street/Dwarkesh one) and endorsed by fancy people (a $1B ‘Amodei-Altman Alignment Award’, anyone?).

I’d be keen to hear from people with ideas on what prizes (if any) we should be setting!


  1. ^

     Although the prize will not necessarily be counterfactual in causing the winning work to happen.

  2. ^

     One important step is to set up the prize from a reputable funding vehicle that has a very low risk of bankruptcy or equivalent trouble.

  3. ^

     Jason Hausenloy lists some ideas.