Comment by DANmode

6 hours ago

MCP apparently reduces probabilistic failures - aka the common fatal flaw of all of these robots (hallucinations, missing stuff in the API doc, etc).

This makes it a little more interesting to me, knowing those results.

It definitely underlines what we already know about the specific weaknesses of LLMs replies/results.

>MCP apparently reduces probabilistic failures

Does it really do that much difference? I mean in the end robots process MCP output the same way they would process API docs.

I think that was a problem 6 months ago, but GPT 5.6 Sol on xhigh doesn't have those sorts of issues. I don't think it'll last. Things are moving fast.

  • People say that with every single model release, and every single time they're wrong. I bet you anything that it's no different this time than the last dozen times.

    • i'm burning 500m tokens a day "writing code" for 83 days straight. one of those projects i'm building integrates netbox, stripe, quickbooks, mercury and deel all together via api. i wrote zero lines of code and read zero api docs. it does exactly what i want and gives me a 360 degree view of every aspect of my multi million dollar arr biz.

      how about you?

      1 reply →