RSS

Elliott Thornley

Karma: 1,219

I work on AI alignment. Right now, I’m using ideas from decision theory to design and train safer artificial agents.

I also do work in ethics, focusing on the moral importance of future generations.

You can email me at thornley@mit.edu.

AI Philos­o­phy Com­pe­ti­tion: $11,000 in prizes.

Elliott Thornley1 Sep 2026 6:28 UTC
13 points
2 comments1 min readEA link

Challeng­ing the un­aware­ness ar­gu­ment with third ac­tions and mixed actions

Elliott Thornley17 Aug 2026 14:43 UTC
39 points
15 comments10 min readEA link

Tie train­ing can make DPO/​RLHF-trained AIs gen­er­al­ize better

Elliott Thornley6 Jul 2026 16:21 UTC
14 points
0 comments15 min readEA link

Risk-Averse AIs

Forethought24 Jun 2026 11:35 UTC
35 points
8 comments5 min readEA link
(www.forethought.org)