Human Compatible
Stuart Russell
Human Compatible asks how machines that may surpass human intelligence can remain beneficial rather than become dangerous. It explains why fixed objectives fail, surveys social and economic risks, and develops preference-learning, deference, corrigibility, governance, and human-autonomy ideas for keeping AI aligned with people.
10 chapters