Avancerad
The AI Alignment Challenge
Theoretical challenges in matching artificial intelligence goals with human values.
📝 Contenu du Prompt
Define the concept of instrumental convergence in artificial intelligence. Explain why an AI might pursue harmful sub-goals even if its final goal is benign. Discuss the theoretical difficulties involved in formally specifying complex human values into a utility function that a superintelligent system could optimize without causing unintended side effects.