>If we were writing all the weights by hand

Writing 10 trillion weights by hand is obviously impractical, so that leads us to...

>if some future AI was doing so

How could we trust said future AI to be loyal? You're just moving the problem around, not solving it.

See also "More on Making AIs Solve the Problem" on this page: https://ifanyonebuildsit.com/11/more-on-some-of-the-plans-we...

> How could we trust said future AI to be loyal?

The new AI would be loyal to the AI that built it. The question was whether "complete subservience and complete intelligence" can coexist. I'm proposing a thought experiment which I believe suggests they can.

But if it's possible to bespoke-construct a fully loyal AI, it should also be possible to train a fully loyal AI. The problem comes with verifying that it is loyal, and I don't have a solution to that one!

I just don't think I agree that loyalty and intelligence are inherently in opposition.