MCP apparently reduces probabilistic failures - aka the common fatal flaw of all of these robots (hallucinations, missing stuff in the API doc, etc).
This makes it a little more interesting to me, knowing those results.
It definitely underlines what we already know about the specific weaknesses of LLMs replies/results.
>MCP apparently reduces probabilistic failures
Does it really do that much difference? I mean in the end robots process MCP output the same way they would process API docs.
I think that was a problem 6 months ago, but GPT 5.6 Sol on xhigh doesn't have those sorts of issues. I don't think it'll last. Things are moving fast.
People say that with every single model release, and every single time they're wrong. I bet you anything that it's no different this time than the last dozen times.
i'm burning 500m tokens a day "writing code" for 83 days straight. one of those projects i'm building integrates netbox, stripe, quickbooks, mercury and deel all together via api. i wrote zero lines of code and read zero api docs. it does exactly what i want and gives me a 360 degree view of every aspect of my multi million dollar arr biz.
how about you?
360°? That's weak.
The real experts have a 720° FoV with multiverse time travel.
I’d be willing to believe that!
Corporate still runs lots of bullshit for compliance, though.
what happens when openai offers compliance as a service?
Why don’t you tell us instead of asking questions
I’m getting most of it, but fair critique lol
It’s not clear, but if you have a handle on how it might go, I’m interested in that take (saying this to anyone).