Right, but we have no clue why, and how the emergent behavior they show works.

If we would know that, there would be no need for interpretability research.