MCP apparently reduces probabilistic failures - aka the common fatal flaw of all of these robots (hallucinations, missing stuff in the API doc, etc).

This makes it a little more interesting to me, knowing those results.

It definitely underlines what we already know about the specific weaknesses of LLMs replies/results.

>MCP apparently reduces probabilistic failures

Does it really do that much difference? I mean in the end robots process MCP output the same way they would process API docs.

I think that was a problem 6 months ago, but GPT 5.6 Sol on xhigh doesn't have those sorts of issues. I don't think it'll last. Things are moving fast.

People say that with every single model release, and every single time they're wrong. I bet you anything that it's no different this time than the last dozen times.

i'm burning 500m tokens a day "writing code" for 83 days straight. one of those projects i'm building integrates netbox, stripe, quickbooks, mercury and deel all together via api. i wrote zero lines of code and read zero api docs. it does exactly what i want and gives me a 360 degree view of every aspect of my multi million dollar arr biz.

how about you?

360°? That's weak.

The real experts have a 720° FoV with multiverse time travel.

I’d be willing to believe that!

Corporate still runs lots of bullshit for compliance, though.

what happens when openai offers compliance as a service?

Why don’t you tell us instead of asking questions

I’m getting most of it, but fair critique lol

It’s not clear, but if you have a handle on how it might go, I’m interested in that take (saying this to anyone).