Maybe unpopular prediction:

I think vision models will come more into play for validating things. It’s the most like consciousness, and less like - as you put it an LLM validating its own assumptions.

It’s at least an independent way of analyzing the work (as glyphs and images).