Value generalisation Theory of Change: putting it into practice
Source ↗
👁 4
💬 0
In the previous post, I presented my theory of change for why value generalisation is vital for AI alignment. Here I'll add the practical part of the argument: given those facts, why explicitly try to do value generalisation, what are the dangers of the approach, how should it be done, and how do we mitigate the risks?The formal theory of change is down below, but I'll put a collapsible version here, to make references easier:Theory of ChangeInputs / activities: investment/grants, a small resear
Comments (0)