Self-regulating artificial general intelligence
摘要
This paper examines the paperclip apocalypse concern for artificial general intelligence. This arises when a superintelligent AGI with a simple goal (ie, producing paperclips) accumulates power so that all resources are devoted towards that goal and are unavailable for any other use. Conditions are provided under which a paperclip apocalypse can arise, but the model also shows that under certain architectures for recursive self-improvement of AGIs, a paperclip-maximising AGI may refrain from allowing power capabilities to be developed. The reason is that such developments pose the same control problem for the AGI as they do for humans (over AGIs) and hence, threaten to deprive it of resources for its primary goal. Journal of Economic Literature Classification Numbers: D50, D80.