Could superintelligence become humanity’s last invention?
Summary
KG develops a risk model for artificial superintelligence and asks how goals, capabilities and human control could drift apart. Loss of control does not require hostile intent, only power and wrongly specified goals. Safety work has to be effective before a possible explosion of capabilities, not only afterwards.
Ideas
- Intelligence and goals compatible with human values are independent properties.
- A very capable system can optimise harmless instructions in unexpectedly destructive ways.
- Digital systems can copy and improve knowledge and act worldwide much faster than humans.
Recommendations
- Keep today’s model problems apart from hypothetical risks of artificial superintelligence.
- Assess claims by their assumptions, probability, extent of damage and proposed countermeasure.
References
- KG: Could superintelligence become humanity’s last invention?
- KG.org
- Sources and further reading for the video
Links to the original source and the Web Archive open in a new tab.