I wish you would provide more information. About time, costs, reliability (how many duds did you have? did you need to do any babysitting or etc) and so on But it is a very impressive looking demo, for sure.

this is a one-shot result but i have a really really lengthy prompt: https://github.com/PhiloLabs/fable51-worlds/blob/main/union-... with clear guidance in using subagents and self-QA loop. ~2 hour (extensive subagents usage), total ~8M tokens, ~$33 under API

Did you need to iterate on the prompt, or did you have a model help you author it? I frequently have problems with orchestration instructions in-prompt, and your is huge. Maybe this is just better with Fable? I honestly haven’t used it much.