I think this is what obstacles on the path to AGI look like now. It’s random things that would be obvious to a human but are unrepresentative in how an AI views the world and therefore it suddenly becomes seemingly incapable, despite having basically superpowers for proximal work.

I don’t mean that to say AGI is here or easy or necessarily that close but it’s likely going to feel like one thing after another until one day most of these things that make you think “how could something so capable be that dumb” are largely solved.

Yes, and right now it seems we get around this problem by spawning 10k agents (that's what OpenAI did for the stokes problem) and hoping that at least one of the 10k catches this and does it right - which it very likely will.