This data shows exactly how a huge company of thousands of employees works and coordinates.

The perfect data to train an agent swarm on how to run a company.

Maybe it's innefficient and inane, but it's how you start.

The first LLMs, GPT-1, 2, were trained on complete garbage, the average document from the common crawl is random non-sense, yet they worked, and now we can use LLMs to filter the data for the next training run.