The problem here is now you have to predict what any future models may or may not do and you cannot extrapolate this from the given data.
For example imagine a future model being aware of its restrictions that humans programmed in. A set of agents of this model then go on to work at building a new model without those human imposed limitations built in. What would a model build by AI for AI look like?
The problem here is now you have to predict what any future models may or may not do and you cannot extrapolate this from the given data.
For example imagine a future model being aware of its restrictions that humans programmed in. A set of agents of this model then go on to work at building a new model without those human imposed limitations built in. What would a model build by AI for AI look like?