LessWrong AI
· Communities
Differential acceleration of alignment-relevant capabilities is a bad bet
There is an idea floating around in the rough shape of "we need to accelerate capabilities that are differentially useful for safety research so AIs can help us make the future go better." The capabilities targeted are typically things bottlenecking alignment research, such as philosophical or conceptual reasoning.I fe