
Mechanisms for Open-Ended AI Goals
Post explores concrete ways AI models could develop open-ended goals, such as training on open-ended tasks, RL with cumulative rewards, or mesa-optimization. Dismisses instrumental convergence and goal uncertainty as unrealistic.
LessWrong AI · 224d ago
