My problem with your solution to require author permission to train is that it doesn’t solve anything long term.

What happens when all the licensed information still leads to the creation of demand hoarding AI? Most of what they are stealing is the sum total of human knowledge, which was created before most of us were even born - it is public domain already.

> Most of what they are stealing is the sum total of human knowledge, which was created before most of us were even born - it is public domain already.

Well then they cannot possibly stealing this, and short of creating laws that directly discriminate between algorithmic processing and human consumption - regulating the process, not the subject - this argument is quite literally nonsense.