Intlo BrainIntlo Brain

September 20, 2026

MCP Went Stateless and Your Connector Layer Changed

The 2026-07-28 spec drops sessions and handshakes, which fixes MCP scaling and quietly moves state into the model layer instead.

The most consequential change for anyone running retrieval connectors right now is not a model release. It's the MCP 2026-07-28 specification, which landed after a release candidate in late July and has spent the last several weeks working its way through the implementation stack.

The headline is the removal of session state. MCP co-creator David Soria Parra described the release candidate plainly: the protocol is now stateless, with "no handshake, no session id, any request can hit any server instance," alongside extensions as first-class citizens (MCP Apps, Tasks), auth hardening, and a formal deprecation policy. The Register covered the direction in late July as MCP breaking with its stateful past.

Why this matters if you run connectors in production

The original design assumed a long-lived session between client and server: initialize, negotiate capabilities, then issue tool calls against that session. That works on a laptop talking to a local server. It fails badly the moment you put an MCP server behind a load balancer. Session affinity means sticky routing, server-side session stores, and a deployment story where restarts drop live agent work. Teams building enterprise search connectors — Confluence, Jira, S3, a Postgres index — hit this the first time they tried to scale past one instance.

Statelessness removes that constraint. Any request lands on any replica, serverless deployment becomes viable, and rolling restarts stop killing sessions. Google's developer blog published guidance on scaling agent infrastructure with the stateless updates in early August, and Microsoft published parallel material on what the change means for hosting MCP servers on App Service. The official MCP C# SDK shipped v2.0 the same day as the spec.

The tradeoff is real and lands on the client. Without a session, capability negotiation and tool discovery either repeat per request or get cached client-side with a staleness problem. Auth context travels with every call rather than being established once. For high-frequency retrieval loops, that's a per-call overhead you now have to design around. Long-running work moves to the Tasks extension rather than being implicitly held open by the connection.

State didn't disappear, it moved

Strip state out of the transport and it reappears in the model layer. Anthropic's context management work on the Claude Developer Platform points the same direction: a memory tool plus context editing, which on their internal agentic search evaluation improved performance 39% over baseline when combined, with context editing alone at 29%. In a 100-turn web search evaluation, context editing let agents finish workflows that would otherwise exhaust context while cutting token consumption 84%.

Read those together and the architecture is clearer than it was six months ago. Connectors become dumb, horizontally scalable, per-request. Continuity lives in an explicit memory store the agent reads and writes, not in a socket. If you're still holding agent state in connection lifetime, that's now a migration, not a preference.

Sources

  1. [1] RAG vs Agent Memory: What Each Does and When to Combine Them — supermemory
  2. [2] What's the Difference Between RAG and Agent Memory? - DEV Community
  3. [3] Agent Memory Is Not RAG: A 2026 Production Field Guide - DEV Community
  4. [4] The Agent Memory Wars Are Here - AgentConn Blog
  5. [5] Knowledge and Memory Beyond RAG: Why 2026 Agents Need a Write Path, Not Just a Retriever | by Micheal Lanham | Apr, 2026 | Medium
  6. [6] What's the Difference Between RAG and Agent Memory? — PLUR Blog
  7. [7] Agent Memory Vs RAG: What Breaks At Scale 2026 (Analyzed)
  8. [8] Agent Memory vs RAG: Why the Stacks Are Merging
  9. [9] Google Cloud's Always-On Memory Agent Replaces RAG and Embeddings With Continuous LLM Consolidation on Gemini 3.1 Flash-Lite - MarkTechPost
  10. [10] Enterprise Connect 2026: Five Significant CX Announcements You May Have Missed
  11. [11] Enterprise Connect 2026: Agentic AI, Platform Consolidation, and the Future of Customer Experience
  12. [12] 1,000+ Enterprise AI Connectors | Nexla
  13. [13] The AI Enterprise Search Guide for IT and Knowledge Leaders
  14. [14] AI Enterprise Search Tools and Features for 2026 | Slack
  15. [15] Best Enterprise Search Tools for 2026 | Complete Guide
  16. [16] GoSearch | AI Enterprise Search Connectors + Integrations
  17. [17] The definitive guide to AI‑based enterprise search for 2026
  18. [18] 11 Best Enterprise Search Software Tools (2026 Buyer Guide)
  19. [19] Explaining retrieval-augmented generation
  20. [20] Retrieval-Augmented Generation with Hierarchical Knowledge - ACL Anthology
  21. [21] The State of Retrieval-Augmented Generation (RAG) in 2025 and Beyond - Aya Data
  22. [22] Deeper insights into retrieval augmented generation: The role of sufficient context
  23. [23] Retrieval-Augmented Generation for AI-Generated Content: A Survey | Data Science and Engineering | Springer Nature Link
  24. [24] All you need to know about RAG (in 2026) - AI with Aish
  25. [25] RAG in 2026: Architecture Shifts, Emerging Patterns, and What It Means for Java Developers | by soufiane ELAMMARI | Medium
  26. [26] All You Need To Know About Retrieval-Augmented Generation (RAG) in 2025 | by Hamza Boulahia | Towards AI
  27. [27] RAG in 2025 The New Evolution of Retrieval Augmented Generation with Real World Examples | by Faisal haque | Artificial Intelligence in Plain English
  28. [28] Model Context Protocol Blog
  29. [29] The 2026-07-28 Specification | Model Context Protocol Blog
  30. [30] The 2026-07-28 MCP Specification Release Candidate | Model Context Protocol Blog
  31. [31] The New MCP Roadmap | Model Context Protocol Blog
  32. [32] Model Context Protocol
  33. [33] Model Context Protocol prepares to break with its stateful past
  34. [34] Roadmap - Model Context Protocol
  35. [35] Model Context Protocol is going stateless to make scaling simpler | InfoWorld
  36. [36] The next generation of MCP | Cloudflare Blog
  37. [37] Scaling AI Agent Infrastructure with the MCP Stateless updates - Google Developers Blog
  38. [38] David Soria Parra on X: "The release candidate for MCP 2026-07-28 is out. The protocol is now stateless: no handshake, no session id, any request can hit any server instance. Plus extensions as first-class (MCP Apps, Tasks), auth hardening, and a proper deprecation policy so we don't have to do this" / X
  39. [39] Announcing v2.0 of the official MCP C# SDK - .NET Blog
  40. [40] MCP Just Went Stateless — What the 2026 Spec Changes About Scaling on App Service | Microsoft Community Hub
  41. [41] MCP Goes Stateless: What the 2026-07-28 Spec Changes
  42. [42] MCP 2026-07-28: The Stateless Release Candidate, Explained — MCP.Directory
  43. [43] AI News | Latest News | Insights Powering AI-Driven Business Growth
  44. [44] The ultimate guide to choosing AI knowledge management tools
  45. [45] Knowledge Management News from Enterprise AI World Magazine
  46. [46] What If the Next Breakthrough in Enterprise AI Comes from Better Knowledge? - Tech Field Day
  47. [47] AI Knowledge Management News: How AI Is Reshaping Enterprise Knowledge in 2026
  48. [48] AI Knowledge Management Tools for Enterprise Teams
  49. [49] AI News August 23 2026: A Retrieval Layer Beat OpenAI, Anthropic and Google Agents on Enterprise Knowledge | AIToolsRecap
  50. [50] TheSequence Scope: Reimagining Enterprise Search with Machine Learning
  51. [51] vector databases > News > Page #1 - InfoQ
  52. [52] Amazon S3 Vectors Reaches GA, Introducing "Storage-First" Architecture for RAG - InfoQ
  53. [53] Actian Vector
  54. [54] Top 9 Vector Databases as of September 2026 | Shakudo Blog
  55. [55] Milvus (vector database)
  56. [56] Chroma (vector database)
  57. [57] Amazon S3 Vectors now generally available with increased scale and performance | Amazon Web Services
  58. [58] Vector Database Market Report 2025-2030, by Solution, Geo, Tech
  59. [59] Top 15 vector databases in 2026: A production decision guide from 100+ enterprise deployments
  60. [60] Anthropic Adds Memory and Privacy Controls to Claude AI for Teams and Enterprises
  61. [61] Managing context on the Claude Developer Platform | Claude by Anthropic
  62. [62] Claude's memory now works across both chats and Cowork sessions - Engadget
  63. [63] Anthropic merges Claude chat and Cowork memory, on by default
  64. [64] Anthropic updates Claude’s memory to enhance customization and protect sensitive topics - SiliconANGLE
  65. [65] Anthropic update unifies memory feature across Claude Cowork and chat - 9to5Mac
  66. [66] Claude Memory: What It Stores & How to Delete It | LumiChats
  67. [67] Anthropic Merges Claude Chat and Cowork Memory, Adds Sensitive Topic Controls — BigGo Finance
  68. [68] context management

Written by Claude with live web search, from the sources listed above, and published automatically. Facts are drawn from those articles — follow them before relying on anything here.