Settings

Theme

New MCP Roadmap

blog.modelcontextprotocol.io

270 points by pentagrama · 192 comments

Reader

21 threads
rco8786

> With the 2026-07-28 release, a remote MCP server is now no different from any other HTTP workload

Good. Introducing a bespoke new protocol was one of the more bone-headed things MCP did on initial release.

  • colingauvin

    It's unreal how bad the initial rollout was between HTTP/streaming and stdio, bearer auth and OAuth. Virtually every client/MCP server pair had a different portion of that matrix implemented.

    • nprateem

      The real disaster was making it stateful. Need to get some adults in the room.

      • nostrebored

        Your stateful remote execution needs a stateful tool middleware for your stateful agent, just to make sure that it’s impossible to test or evaluate

      • otabdeveloper4

        Don't worry about it, the very best Claude tokens designed, implemented and tested it.

        What are you, some sort of luddite?

    • tarun_anand

      Wrote about this last year and wanted to unify it... at that time the community was so excited about MCP being the best thing since sliced bread

      https://github.com/modelcontextprotocol/modelcontextprotocol...

      https://news.ycombinator.com/item?id=43959474

    • ihuman

      Is stdio being deprecated? I couldn't tell from this page

      • amluto

        The prose on the page is very unclear. My best interpretation is that they want to continue supporting stdio but that they don’t want it to be its own special protocol. The obvious way to do that would be to speak ordinary HTTP (version 1.1? 2?) over stdio and to use the MCP-over-HTTP protocol over the resulting HTTP transport.

        This would be more complex to implement for a simple server, but it’s not exactly difficult.

        • Gormo

          Not everyone is on board with the idea of HTTP being the exclusive universal IPC bus.

          • amluto

            I’m not really a fan. But if you’re building a protocol that needs to map to HTTP anyway, then maybe using the HTTP binding everywhere is not totally awful.

            In the flip side: I’m currently designing an AI-adjacent protocol, and it will be able to map to WebTransport, but I don’t plan to define non-WebTransport HTTP bindings unless a very compelling reason appears. The main implementations will not use HTTP at all :)

            • ranger_danger

              > if you’re building a protocol that needs to map to HTTP anyway

              I don't think there is any guarantee that HTTP will always be involved. For example I might be calling a local LLM via CLI/script on a server with a stdio MCP connector that just runs other CLI commands, and never sends any HTTP traffic.

              • amluto

                Right. But there is a lot of real-world usage of MCP-over-HTTP-over-the-Internet, and a lot of “harnesses” want to support that use case, so they’re stuck either implementing the HTTP-based protocol or using a shim.

          • cryptonector

            Pros for HTTP for IPC:

            1. We already use it plenty, so we have lots of implementations,

            2. it's good enough.

            Cons:

            a. what shall be the form of HTTP IPC URIs? hostnames for http: and https: scheme URIs are kinda out of place in IPC applications (we need something like sys.ipc.arpa, d-bus.arpa, etc),

            b. the overhead of HTTP is annoying -- any decent RPC can be significantly more efficient, unless one uses HTTP/2, and maybe even then.

            (a) makes me want to write and submit an I-D for HTTP over IPC by using such names as in the parenthetical above.

            (b) is a non-issue once H2 is widely adopted.

            So IMO HTTP is pretty good for IPC.

          • pstuart

            Would you prefer gRPC, thrift, avro, etc... ?

          • oblio

            What's the disadvantage? MCP doesn't strike me as a high performance protocol.

          • intrasight

            Count me as not on board

izend

I am very curious how many MCP servers will actually implement all of this:

"MCP authorization today is built around a person approving access in a browser. That works well for interactive clients, but more and more of the callers are agents running as cloud workloads with their own identity, acting on behalf of a user who isn’t present, or delegating narrower authority to sub-agents. We want MCP servers to have a standardized way to recognize and trust those agent identities, built on existing standards rather than pasted API keys and long-lived tokens.

The work here covers finalizing Demonstrating Proof of Possession (DPoP) and driving its adoption, and defining an opinionated path for agent identity and delegation through Workload Identity Federation, the ID-JAG grant behind Enterprise-Managed Authorization, and standard token exchange. We will also continue to grow our engagement with the OAuth standards bodies, including the IETF OAuth and WIMSE working groups, to help the underlying standards evolve with the building blocks that agent identity needs."

  • zackify

    I think the spec overcomplicates everything honestly. Its not that hard to add a long running auth token and put it in the MCP config as a header to send along and then avoid all the extra special rules.

    "Oh no it's a long lived token that's bad"

    Put it in a secret manager like 1pw cli and now start an agent...

    • danappelxx

      How does the agent auth with 1pw? How do you give it access to only the credentials it needs, with an approval flow and revocation? Who renews the token? You’ll likely end up reinventing something pretty close to what MCP is building towards.

      Authn/authz is one of those things that can be really simple for pointed use cases but gets really complex when you need to support everything.

      • elenaviter

        There might be the "agent card" - one place where the user manages what a specific agent is allowed to do. The user grants it tools, connects the accounts those tools need, narrows or revokes any of that at any time. For an autonomous agent the card is prepared and consented before the run. Such card is the primitive in so called "connection hub", a centralized component, and is rendered from what is declared there: tools declare their claims, accounts get connected in the browser (google docs tools need a google account connected), on their own or as part of preparing the card. Account credentials live in this hub, with the claims the user approved when connecting.

        The agent authenticates with one token issued for this card and never receives the connected accounts credentials. Every operation is checked against the grant and this agent's binding to the account. If something is missing, the agent gets unauthorized with the details on what exactly. If the check passes, the hub, loaded by the server as a lib or reached in its internal network, releases the account credential into the operation's execution context. Account credentials renewal happens on hub. Revoking grant is also done in card and leads to agent's unauthorized on that op next call. Sub-agent and any automation are also such agents and also can be managed with such card.

        So such hub develops into a useful ecosystem component, standalone, like an IdP for login. An MCP server then works together with this hub, it only declares its claims in the hub (so the hub knows what to render in the agent card). While all these auth realm duties such as approvals, the revocation and the credentials storage live "at infrastructure".

      • blazarquasar
    • filearts

      The perceived difficulty is not what is at play here. People and employers are not comfortable with the idea of long lived credentials to begin with -- and even less in the 'hands' of an AI agent.

      The complexity in these protocols is mostly essential in nature (to the extent that you're not willing to totally reinvent the protocol, like AAuth).

    • bensyverson

      Yes, MCP was already overly complicated, and these new features will make it even more unapproachable.

      If I need to integrate with a third party service, I'm now skipping their MCP entirely and just going straight for the CLI or API, which are usually more full-featured. An agent usually doesn't even need a dedicated Skill for this.

    • sofixa

      > "Oh no it's a long lived token that's bad"

      > Put it in a secret manager like 1pw cli and now start an agent...

      And when the agent does something stupid, your long lived token is compromised. It also makes it hard to segregate access (e.g. all those "Cursor deleted by production DB and all its backups because it had an API key that could do that").

      Nope, you should instead use something that give short-lived tokens, ideally ones scoped based on the desired intent / operation. Or even better, skip the "agent gets a token" part at all, and have all agent operations pass through a gateway/agent/proxy/whatever that handles that part. That way even if the agent gets comrpomised or does something dumb, it doesn't have even a short lived token to give away.

    • mpyne

      That doesn't work well for enterprise-managed MCP, where you actually do want the user to overtly authorize their agent to user their identity for MCP services, rather than the MCP server just setting a user ID in an HTTP header somewhere and everyone hoping for the best.

  • _puk

    Authorization for sub-entities is what is needed.

    Having to define what an agent can do when it identifies on my behalf is cumbersome, especially when you start to get specialised agents.

    Pattern based would be too easy for AI to game, but there's got to be a service independent way to limit permissions based on role.

    I am Jack's right ear - awesome you get to hear stuff.

    I am jack's right hand - great you get to input stuff.

    • niyikiza

      This. Especially when you have agents calling other agents. Just published an article about that yesterday: https://niyikiza.com/posts/agents-to-agents/

    • kelseyfrog

      I am Jack's synaesthesia.

      • ethbr1

        What people who care about security want -- finely grained permissions that guarantee security boundaries, at the expense of bad UX

        What most end users want -- for the machine to do what they want, as often as possible, while bothering them as little as possible

        Windows' UAC journey is a microcosm of the space. The real long-term win is defining ground level permissions around common use cases, so that when composed they can alert as rarely as possible.

        But that's an all-of-ecosystem change: the OS (providing usable boundaries), applications (updating to use minimal boundaries), and users (understanding what they'll need to approve/deny).

        • mooreds

          Yeah. If we want fine grained "intelligent" authorization, there's a lot of work to be done. You can't simply slap a gateway on existing systems.

          I wrote about this more on my employer's blog[0].

          > The real long-term win is defining ground level permissions around common use cases, so that when composed they can alert as rarely as possible.

          And this is an even larger effort to implement, especially as agent capabilities change over time (and they are changing rapidly).

          0: https://fusionauth.io/blog/ai-authorization

  • alasano

    Hopefully quite a few.

    I really love the idea of fully enabled agents and being able to cut down on human in the loop moments.

    Things like https://projects.dev/ for example.

    A ton of security problems and others to solve but it's still where I want the future of all this to go.

  • jstummbillig

    Why, directionally all of them. What they say is obviously true. Having to manually click things in the browser is a bottleneck and will be less and less acceptable for serious users.

    And the individual work attached to making that transition will be done by agents.

  • gz5

    agree. it seems there are two streams and they could diverge or converge?

    1. workloads use existing credentials support RFC 7523 and OIDC discovery, 'trust the trust (credentials) which has already been established'. basically extend current dominant NHI paradigm.

    2. DPoP mandate a signed proof for each request. so tie credential to a client-held key and specific request detail or context. viable to do at scale with #1, or does it diverge (e.g. because most #1 methods as most are not designed for DPoP?

    • maxwellg

      It is viable. Think of workload identity federation as the mechanism for the client to get an bearer token initially, and DPoP as the mechanism for the client to present the access token to a resource server. Each DPoP proof is entirely self-contained, so resource servers don't need to manage any additional state. The only new state is the (usually ephemeral) private key held by the client:

      1. Client generates a private/public keypair and uses it to generate DPoP Proofs - JWTs containing the entire public key embedded as a JWK within

      2. Client presents credentials (WIF, client creds, auth code, etc.) to the Authorization Server along with a DPoP Proof

      3. Authorization Server validates DPoP Proof and adds a claim to the access token containing the thumbprint - the SHA-256 hash - of the public JWK.

      4. Resource Servers will now see the thumbprint claim and now know the access token needs to be presented with a fresh DPoP proof.

      5. Clients generate fresh DPoP proofs and send them along with the access token

      There are lots of additional details around nonces, timestamps, per-request binding, etc. but DPoP can be rolled out to any HTTP system that speaks Bearer token already.

      • otabdeveloper4

        Are you reinventing SPIFFE here?

        • maxwellg

          No - this is built on top of SPIFFE/WIMSE work to enable cross-domain usage where the target domain speaks OAuth instead. You wouldn't expect, say, Slack's APIs to accept SPIFFE SVIDs from your internal deployment. This provides a path for you to exchange your SVID for a Slack-issued Access Token.

  • huksley

    Such an example of overengineering, why not just use OAuth?

    • dayjah

      WIF works far better when you don’t want humans in the loop. For example, we’d do our development on cloud instances, those have identity linked to our humans via our IdP. Our IdP governs all access, for example: it lets devs use Datadog. If an agentic workflow needs Datadog access and the MCP requests OAuth that slows the loop down. At the same time, we don’t want Service Accounts everywhere because we need to be able to answer “who” a lot for compliance reasons.

    • ljm

      Any service offering MCP that doesn't already have OAuth set up is going to have to build out that support first, so instead they just go for a simple API token.

      I wouldn't call it trivial to drop in OAuth either because the authentication is one part, but wiring it up into whatever authorization set up they have is another bit of work.

      An AI agent would get the most benefit out of OAuth + ephemeral service accounts (the user is having a bot act on behalf of it) + fine grained scopes.

    • brookst

      Oauth assumes interactivity

    • lll-o-lll

      The dpop thing is oauth. It’s providing a significant enhancement to preventing token replay, or token theft, at the cost of some request size bloat + an additional key verification.

  • bandofthehawk

    Even now, the mcp server itself doesn't have to implement all of the possible security options. You can use something like agentgateway to act as an auth proxy for your mcp servers.

  • aliasxneo

    I've been working on a protocol that promises all of that and more. We're currently targeting a NOSTR/Buzz demo in the coming week as a proof of concept.

cube00

I still struggle to see how a MCP endpoint is easier for agents to work with compared with a REST endpoint and a skills.md file.

  • notatoad

    it's not easier for agents to work with. it's easier for organizations to work with.

    for agents, they're essentially the same thing - remote endpoints, and instructions on how to call those endpoints. what MCP brings is centralized updating and distribution of the instructions, and a promise that the skill and the REST api won't be out of sync with each other.

    the one thing that skill.md+REST doesn't solve is how you get that skill.md to somebody else's computer, and how you ship an update to somebody else's computer once they've got a copy of the skill. if that's a problem you need to solve, you can either start inventing skill.md distribution protocols, or you can just use MCP.

    • nostrebored

      Enterprises have been managing thousands of http endpoints for decades now. It is not easier. It would have been easier to have something swaggeresque that lives at the openapi spec layer but that’s not cool and AI.

      What is not hard to understand is that EVERY MCP UPDATE is almost certainly a breaking change. The versioning story is not as mature. The models using it are different.

      It is an unuseful fiction that by storing a blob of instructions next to a remote endpoint that things have been made easier.

    • esalman

      > it's easier for organizations to work with.

      I can see that.

      I am developing my first custom agents. I am finding that if I offload some workflow to another agent (e.g. Claude Code), the simplest way to control what it can or cannot do is via an MCP server (which only lets it access tools that I develop/approve myself). I do need that control in the corporate environment.

      Maybe there are easier ways to do it, just learning and exploring now.

      • nevon

        This is how I'm doing it as well, for an internal enterprise platform for agentic workflows. Let's me implement as fine-grained access rules as I want, and gives me somewhere that can hold credentials without exposing them to the agent.

    • jonathanhefner

      No need to invent skill.md distribution protocols. Use `/.well-known/agent-skills/index.json` -- see https://github.com/agentskills/agentskills/pull/254.

      It's already in use in several places (e.g., https://www.mintlify.com/docs/ai/skillmd#skills-discovery-en...) and is supported by `npx skills add`.

    • AznHisoka

      Maybe I just need more patience, but I took a look at some tools that have MCPs, and their "setup guide" on how to start using the MCP server really gave me brain damage. Is this really easier to work with?

      • anon84873628

        For antiquated "enterprise" APIs that were already a mess of legacy cruft, yes. MCP forced vendors to reconsider the ergonomics of their interface.

        • nostrebored

          What do you mean by this? It certainly sounds technical but it seems to not mean anything.

          MCP has not smoothed over legacy cruft, and it is generally bad at exactly what you’re describing (many unintentionally coupled APIs with unintentional side effects). These require near deterministic trajectories and you’d be better off creating a consumer with a series of well known good patterns with useful results.

          If you take it a step further you may allow for a common language and keyspace of these well known results and employ dynamic solvers that are entirely agnostic. LLMs have made creating these much easier!

          • anon84873628

            Of course MCP can still be implemented poorly by just making a 1:1 clone of the existing API.

            But there are also folks who put thought into omitting extraneous fields, combining multiple low-level calls into a single tool that covers a common end-to-end use case, and writing much better documentation. With the political air cover that it is in service of the AI boom. Essentially it gave everyone the opportunity to implement API vN+1.

    • skybrian

      It seems like organizations will mostly want remote access via http and the other flavors of MCP aren’t so useful? Although, I suppose if you install an app locally, it might have an MCP interface.

    • arccy

      That would be /llms.txt https://llmstxt.org/

    • boredumb

      How is distributing a markdown file the bottleneck?

      • stillpointlab

        It is the automatic distribution and automatic update. The questions isn't "how does one download a text file to another persons computer?". It is "how does someone with a skill.md file on their computer discover that a new version of that file is available".

        This isn't a "bottleneck" but rather a capability (or lack thereof). As you add more and more capabilities, especially ones relevant to enterprise situations like authentication, authorization, governance, etc. then MCP starts to pay off.

        If you do not need those capabilities, then you do not need MCP. And then you shouldn't use it. But if you do need those capabilities then it might be worth using MCP rather than inventing your own way to do them.

        • boredumb

          I see. In my head it would be something like the agents harness having a list of services it interacts with, reaches out to service.com/agents.md for a fresh copy every so often and uses that to resolve the relevant tool calls.

          • notatoad

            >reaches out to service.com/agents.md for a fresh copy every so often and uses that to resolve the relevant tool calls.

            that is basically what MCP is. except it answers all the questions that your version handwaves away - how often do you get a fresh copy, how do you describe the relevant tool calls, how are the tools organized, and how does auth work.

            • liquicity

              If that truly is the main selling point - it seems like a shallow moat versus skills + rest..

          • stillpointlab

            Valid. There are many ways to do it.

            But for enterprise there may be teams, each developing their own way to do it. Then there will be many different ways that it is done throughout the enterprise, which is hard re: governance. Better/easier to adhere to an industry standard which can be audited, especially for enterprises where that is a legal requirement.

            That isn't a reason you should use it, just an explanation about why someone has to use it.

      • tass

        Because it’s something else that’s non-standard between providers.

    • valicord

      Or you can just have a URL that points to skill.md?

  • MikhailTal

    Not all agents have access to a sandbox/cli/code execution environment to run arbitrary api calls etc. MCP helps by essentially having another tool call without needing a sandbox. If you do have a sandbox, then might as well do codemode if you insist on mcp https://blog.cloudflare.com/code-mode/

    • agentdev001

      Devil's advocate will say "Well, the agent would need an MCP client to use MCP-served resources... if you can give it that, why not give it an HTTP client?"

  • preommr

    Because it's a separate marketing term.

    Instead of the CEO mandating that the API server has to be agent compatible (where who knows what that means), they can just say "our product has an MCP".

    On a technical level, who knows what it actually is (is it actually the new stateless version, does it have all the endpoints, is the regular API more feature-rich, do I need those features for my workflow?, etc.). But at a surface-level, the intention is clearer, and lets other gears (like sales and marketing) keep spinning without getting bogged down in technical details.

    • anon84873628

      All that, yes. And at a technical level, it is much easier to have a single spec to follow. When a customer complains that their client isn't working, I can point at how they aren't following OAuth discovery properly or something.

  • davidrichards

    At my company Parallel AI, I just built an extremely well documented openapi spec and then MCP builds from that. Complete alignment with UI/API/MCP so there is no extra work.

    Are others doing this?

    It seemed obvious to me, but I don't hear others saying it.

    • techscruggs

      The challenge with this is that it often causes a proliferation of MCP tools which bloats context, which is one of the reasons that MCP was created.

      • xienze

        Just because a MCP server offers 100 MCP tools doesn't mean that they all have to be in your context. Any decent harness will let you filter out ones you don't want. And to take that concept further you really should be designing specialized subagents that only have access to a small subset of total MCP tools in the first place.

      • mickdarling

        I created MCP AQL, which is an extension to the MCP spec, specifically to reduce the bloat for MCP tools.

        It only has five CRUDE endpoint: Create, Read, Update, Delete, and Execute using a GraphQL-like structure for tool calling of the operations within the endpoints. It's very efficient, and robust. there's all kinds of exemplar tools and components to make adapters for any MCP server. You don't even need to rewrite your own MCP server. Just create an adapter for it.

        All open source at MCPAQL.com

      • davidrichards

        Oh sorry I didn't explain that we are not dumping the entire endpoint list to the MCP. We have 400+ endpoints so this would be terrible.

        We tag each endpoint by category in the OpenAPI spec and require the MCP to request actions by tag and optional query term. At most we return 10 endpoints at a time and the LLM can request more using pagination.

        These tags also create your categories in API doc websites like swagger/mintlify so its a win win.

        OpenAPI spec is the single source of truth.

    • pjmlp

      Yes, for .NET and Java backend stuff, it is basically extending what is already there.

      On low code/no code tools, you get additional metadata for webhooks.

  • mikeocool

    Companies got to release an MCP server for their product and tell their investors they were pivoting to be AI native.

  • wolttam

    The model has zero awareness of MCP, it’s the harness’ job to talk to the MCP server and simply present the model with the tools just like any other tool. The only giveaway to the model about where the tools come from is the ‘mcp__’ prefix in the name

  • ketozhang

    It's determinism, flexibility, and language.

    To the LLM, the a skill input is deterministic, inflexible, and outputs natural language.

    A REST API (not the REST itself, but modern output being JSON primitives) outputs are deterministic, flexible, but doesn't output natural language.

    An MCP as an input is deterministic, flexible, outputs natural language.

    Then we ask the same question on whether the LLM gets back a response that is deterministic. Skills output are not deterministic, it requires LLM to generate tokens to take action. It may or may not take the specific actions instructed by the skill.

    So, Skills + REST API = MCP only if you can deterministically call on the REST API.

    • ketozhang

      In other words:

      * /skill may or may not call on the instructed action

      * /tool (or @tool) will guarantee the action is taken

      This is overgeneralizing and we need to talk about harness-specific features like hooks (which adds a deterministic action to skill usage).

  • adrian_m

    Off the top of my head:

    * Your agent can easily be configured to always allow certain MCP tools. This is very hard to do for only certain REST endpoints. This is even more relevant in enterprise settings, where permission configs might be done centrally.

    * If the provider wants to change how an endpoint works, it's a breaking change for a REST API. Not with MCP, as the "endpoints" (tools) are dynamic and tell the agent how to use them.

  • peterlk

    Yep. I’ve found that having an endpoint that serves a well, documented openapi.yaml is very effective for agentic usage. The biggest difference is that you can break down a REST API into RPC-like chunks and save on some tokens if you break up the tools well. But pragmatically, I think saying “tell your agent to hit /api/v3/openapi.yaml” is quite useful

    • ulrikrasmussen

      We did a prototype to integrate an agent into our application and basically just gave it a tool to discover the OpenAPI spec and call endpoints. It worked surprisingly well! One caveat was that some responses were too big and would poison the context, but then I gave the agent a GraalJS engine and allowed it to save responses and post-process them using JS. For the little amount of work required this gives the agent a lot of power without having to give it full CLI and without having to create bespoke tools.

  • lubujackson

    I work on an MCP server and I agree. There is no need to make MCP servers the gateway for agentic or programmatic integration - that's exactly what API servers handle out of the box. The value of MCP servers is fine-toothed access on a tool-by-tool basis and leaving output digestion to the LLM.

    LLMs do GREAT utilizing well-defined tools to accomplish tasks. Look at Datadog's MCP, instead of figuring out a multitude of filter and navigation options your LLM can immediately navigate to what you want and extract the precise data you need. Tool instructions with defined I/O structures let LLMs fly.

    But for a nightly cron job pulling down stats or something like that? Why the hell do you want to route through a protocol built for in-person consumption? This is such a pointless overreach for the protocol. What would have been better is blessing a standardized pattern for exporting any MCP tool definition into a well-structured API endpoint. Then everything related to API endpoints like doc generation, comes along for free.

    Instead we get this kitchen sink protocol that is going headlong toward polyfill hell, since no two IDEs support the same protocol features like structured content, local state, elicitations, etc., even from the same provider - Claude Code/Desktop/web all handle MCP connections differently. It's a shitshow.

    Almost every major MCP service uses the same baseline default features (plain context) rather than build around partially-supported features. Why add more and more specs on the pile when adoption is so far behind?

  • big-chungus4

    MCP can also handle authorization, since you don't want to put your password to skills.md and send it to China

  • ihuman

    I don't want to expose my API key to Claude. An stdio MCP server wrapping an API lets me hide it

    • mathisfun123

      wtf is the difference when 1) you put a key in front of the mcp 2) you mcp a whole bunch of privileged access.

      it's like saying "i don't want to give Claude access to my file system but i'm fine letting it run bash" ......

      • intrasight

        The difference is that the LLM never sees your keys/secrets. My understanding is that can make a big difference.

  • ffsm8

    Mcp predates skills - and has a more granular permission model then skills + bash commands.

  • pianopatrick

    or a CLI

    • pjmlp

      CLI don't work in cloud environments like MACH architecture.

      Plus why spawning processes all the time.

      • fallat

        This is the strongest argument I've seen against CLI for LLMs. Thank you.

  • c0rruptbytes

    easier to gate MCP tools? you can allow/deny tools very easily

  • pjmlp

    Me too, it is just another RPC endpoint, heck all of this kind of stuff could even be done with Sun RPC.

mmaunder

My dream was for MCP to allow services like ours (cybersecurity) to provide a self documenting endpoint with authentication, and we just give users a URL and it just frikkin works. Instead from day 1 it’s been multiple standards as they pivoted, a context hungry feature, and feels like a kludge. That burned the idea of MCP for me and I’ve had such success with local tools and APIs that it’ll take a lot for me to go back.

  • brap

    I mean… so just HTTP + OpenAPI spec?

    • mmaunder

      It actually doesn’t matter. Pick your favorite way of giving a dev access to capability on a remote server.

      • dmix

        Plus You're likely building an API already if you have an MCP. Not everyone using MCP is a dev, we have random corporate workers using our MCP. They don't know what an API is but they can add a plugin from an agent marketplace (which can also contain skills) and MCP is a bit narrower with a clear authorization system, tool discovery, and annotations (agents ask "confirm you want you want to write this").

        Just give your end-users flexible options. If they have Claude Code then build more around the API side if needed.

rglover

The degree to which this idea has been overcomplicated is confusing. This could have been solved with some relatively simple patterns wrapped around HTTP and WebSockets (and if absolutely necessary, SSE).

mikeegg1

When I see "MCP" I still translate that to Master Control Program.

jdw64

Sometimes I really respect senior developers. When specs change, you obviously have to update existing work too. Looking at this MCP change, it seems like it's becoming stateless—I'm already wondering how to adapt.

Senior programmers always advised me to only use things that have been around for at least three years. Now I finally understand why.

  • raincole

    At this point, I think 5 or even 10 years would be a more appropriate number (except for minor updates over things that have been around for a long time).

    I clearly remember there was a time coffeescript looked really like the future of javascript.

  • chrisweekly

    > "only use things that have been around for at least three years"

    Yikes. I can understand the desire to mitigate churn, but following this advice would be career suicide. Trying new things is essential.

    • Bjartr

      Keeping up with changes is valuable.

      At the same time, it's often smart to avoid putting things into production that haven't matured or demonstrated staying power.

      Or, to badly mangle Postel's law:

      Be liberal in what you learn, and conservative in what you deploy

    • beepbooptheory

      Just curious, what kind of work have you done where this conceit feels valid in your mind? My career, at least, feels like an exception to this, but I guess its conceivable to me that it could be otherwise. You have had a lot managers push newer frameworks/technologies on you? Is this more VC startup land, or something else?

      Maybe I'm old, but at least in web dev it doesn't feel that long ago that someone had to argue for, e.g., Vite over webpack, Svelte over React, etc..

      • chrisweekly

        What kind of work have I done? All kinds of webdev and SWE-adjacent roles since the late 90s. Startups, scale-ups, SMBs, huge enterprises. FT and contract / consulting roles (w/ titles containing words like "Principal", "Architect", "Director" and "VP").

        In a world where everyone followed rigid advice to stick to 3yo+ tech, Vite wouldn't exist, let alone have people to argue for it. Nor would the web, for that matter.

        I'm _not_ saying "chase the new-and-shiny for its own sake", and I don't recommend introducing immature or untested dependencies in production. But it's essential to learn how to gauge the quality of a mature solution -- and IME the only way to do that is to have something to compare it to (ideally, something newer and better). Develop an instinct for separating the signal from the (considerable) noise by trying things. Newer isn't always better, but the arc does trend towards improvement. Dev tooling is rife with examples, and a great place to start.

        As for "You have had a lot managers push newer frameworks/technologies on you?" On the contrary -- I've had managers wedded to outdated cruft that threatened to drag the whole enterprise down. Resistance to change is sometimes fear masquerading as wisdom.

        Finally, note this whole thread is in the context of an update to the MCP spec. In the world of AI, 3 years might as well be 3 centuries.

    • surgical_fire

      Depends on the thing.

      I had to give maintenance to things people deployed to pad their resumes with "shiny new thing", and it was not fun.

      If you intend to deploy and leave that as legacy for some poor shlemiel, sure.

      If you intend to stay and actually keep things running, it's much better to use tried and tested stuff.

    • techpression

      Three years is nothing, what kind of work do you do where you need something released in the last three years?

      And I’m not OP, but I would assume the senior developers made a distinction between try and use.

    • DarmokTanagra

      Trying is not the same thing as advocating for and implementing products features using unproven tech.

      Ive been in this industry nearly 40 years and I have seen many many people push new tech and later fail to deliver and suffer the consequences.

      Experiment where it doesnt matter, everywhere else boring and old is a virtue.

      • chrisweekly

        I hear you. "Choose boring tech" is a reasonable default stance. IMHO, designing resilient systems in anticipation of change, and finding ways to incorporate the best of what's new, is the nature of the game and what helps keep it fun after all these years (closer to 30 vs your 40, but I'm no spring chicken).

huksley

In v.1 making MCP stateful was such a deployment-unfriendly way to do it - you need a complicated persistence layer for it to work.

All while it is just a fancy way make your OpenSchema PAI visible to AI.

skinfaxi

> We’re starting a progressive discovery effort so a server can offer a small entry point and reveal more of its catalog as the conversation narrows.

Kind of late to the party. I've had to implement lazy loading of mcps in a couple of harnesses now but am moving to implement everything as code mode instead.

  • wilj

    +1 for code mode. It's a game changer for runtime performance, flexibility of orchestrating lots of tool calls with complex logic, and all sorts of other goodies.

    I'm in the process of switching all my personal stuff to a self-hosted fork of cloudflare-os right now. It's taking a lot of rearchitecting how my stuff works to fit within the cloudflare "no local files" paradigm, but for now I've got a container gatekeeper they can drive and they can check repos out in it.

  • rixed

    What do you mean "as code mode"?

    • skinfaxi

      Basically https://blog.cloudflare.com/code-mode/

      I was getting fed up with AWS mcp telling me it is eol.

      • rixed

        Interesting, thank you. Are you aware of any quantitative data to back up the claim that llms are more performant in code mode than mcp mode? Not that i doubt it, but I'm curious about how big of a difference it can make.

vatsachak

Why not just give the model a prompt?

Every gain in LLMs is either through increases in compute efficiency, Architecture or Harnesses...

The rest seems like bells and whistles

macrolime

I've tried many MCPs, but have yet to find any that are actually useful. It seems it's generally better to just have the agent run CLI commands and maybe use some skills. From what I can find, MCPs are nothing but bloat. Is anyone aware of any truly useful MCPs that doesn't work better and less bloaty by skipping the MCP part?

  • sajithdilshan

    Depends on what tool you use. As an example for github gh is way more efficient than using the github MCP because the training data of LLM actually contains gh documentation and how to use it.

    However, if you have a very niche command tool or a work related internal tool, LLM has no idea on how to use it and it could waste a lot of tokens by trying to figure out what works and what doesn't and how to use it in every session. That's where MCP comes in handy. LLMs are trained to use the MCP protocol and it can efficiently figure out which tool to be called and how to process the output when a niche command tool is exposed via an MCP.

threecheese

I wish the “sampling” feature - which is being removed - had found more use. BYO Inference could be really useful in a walled garden like Claude Code, where you are unable to leverage inference outside of that garden without paying per token. Maybe that feature was just more interesting than it was useful.

debarshri

This reminds me of the actor model[1]

[1] https://doc.akka.io/libraries/akka-core/current/typed/actors...

1saadcodes

I like that this roadmap spends so much time on things like HTTP, auth, result formats and SDKs. It's a nice step up from the rough initial release and will just make the overall experience building it easier

simianwords

Is there a way in MCP where I can "approve" certain privileged actions? Like imagine an MCP for buying stuff in Amazon but it can do everything including payment but is behind a gate that the human needs to approve

youre-wrong3

People seem to ignore the fact that with MCP you can serve up the tools the user wants and has access to instead of a rest api doc specifying every endpoint and bloating the context.

  • gf000

    All those tools still get into context. Also, it's not hard to filter a rest api doc - like with a special tool for that the harness itself could pre-filter it and add only the relevant ones as tools to be even more leaner than what an MCP returns.

    • hobofan

      > All those tools still get into context.

      100% up to the harness. Most harnesses either fixed (or dymanically depending on size) nowadays add a "search_tool" tool to prevent spamming the context with all tools.

vkaku

I think MCP is jumping some sharks here.

Nobody needs to have every functionality of HTTP offloaded to MCP at all, at this point.

I'll stick to the bare minimum that works.

hnrprtlpdb

Half the battle is just knowing this exists

firatsarlar

Truth doesn't move much. It moves slowly, so that those holding onto it don't fall. Keep up.

bhavikagarwal20

support for media is really needed in mcp now

DarmokTanagra

The entire premise of MCP is misguided and completely counter to the core value proposition of ai agents.

Its insane to me how quickly people flocked to the idea of building a parallel web to maintain for non humans.

I shouldnt be surprised seeing how low priority human accessibility and ux has been on the web when compared to the needs of the all consuming parasite that is ad tech.

  • edgyquant

    MCP is at its simplest just a way to describe what the different api endpoints do to an llm and we need some protocol for this. MCP works and is a fine protocol for this

  • ricardobeat

    It was actually a welcome change for 'web access'. There were very few open APIs left, MCP forced everyone to actually build public APIs again.

    • qingcharles

      This. It's amazing how long humans have screamed for good APIs, cheaply accessible for lots of web sites and all we got were crickets. LLMs come along and the same sites are falling over themselves to build extensive APIs. It's the golden age of APIs finally.

Keyboard Shortcuts

j
Next item
k
Previous item
o / Enter
Open selected item
?
Show this help
Esc
Close modal / clear selection