Thanks for answering. That's really interesting. Maybe agents flip this around. Instead of humans maintaining executable docs, the actual work generates the document. It could become something like a PR-style review layer for agent work. You don't necessarily need to understand the underlying code or tooling, but you can inspect what changed, why it changed, and approve or reject it. Do you think that would address any of the scaling problems you saw?
Hey in this case, you might find https://higherlevel.to/ relevant to you!
Very happy to hear what you think of it. I've built it specifically for reviewing outcomes instead of impementations.
Humans define what correct looks like, agents implement and attach evidence (image, videos, etc.) that humans can review.
I went to give it a try but it's a hosted SaaS and no privacy policy or anything?
Hey sorry for that I should have mentioned that it's in beta!
I've updated the landing page so it's clear that it's in beta and added a small "your data" page for now https://higherlevel.to/your-data