Intlo BrainIntlo Brain

August 13, 2026

MCP goes stateless and your connectors must follow

The 2026-07-28 MCP spec drops the assumption of long-lived sessions, which changes how retrieval and memory servers hold state.

The most consequential thing to land for anyone shipping connectors is the new Model Context Protocol specification revision dated 2026-07-28, which went out after a public release candidate on the MCP blog. Coverage in the run-up was blunt about the theme: InfoWorld framed it as the protocol "going stateless to make scaling simpler," and The Register described MCP as preparing to break with its stateful past. The Agentic AI Foundation's writeup put the shift as MCP moving from a local tool protocol to a distributed one. The changelog for the revision is published alongside the spec and is still being touched, so if you maintain a server, read that page rather than secondhand summaries.

Why this matters for retrieval people specifically: a lot of MCP servers in production are thin wrappers over a search index, and they quietly cheat by keeping things in the session. Pagination cursors. A cached rerank pass over the last result set. Resolved tenant and ACL context from the initial handshake. A warm connection pool keyed to the caller. Once the transport no longer guarantees you'll see the same client on the same process, all of that has to move — either into the request itself, or into an external store you now have to operate.

The tradeoff is real and it isn't free. Session-scoped state was doing useful work: it kept per-request payloads small and let you amortize expensive setup like permission resolution across a multi-turn retrieval loop. Externalizing it to Redis or Postgres buys you horizontal scaling without sticky routing, but adds a round trip on the hot path of every tool call, and it forces you to define TTLs and invalidation for things that previously died with the connection. Pushing state to the client instead is cheaper to operate but inflates token cost and turns your cursor format into a public API contract you can't change quietly.

The upside is the part worth planning around. Stateless request handling is what lets a connector fleet sit behind an ordinary load balancer, scale to zero, run on serverless, and be deployed blue-green without draining sessions. For enterprise search, it also puts authorization where it belongs — evaluated per request, against current group membership, rather than snapshotted at session start. Anyone who has shipped a document ACL bug knows why that's the better default.

The memory-layer side of this is worth watching for the same reason. Capital kept flowing into agent memory infrastructure through the year — Cognee raised a $7.5M seed in February, Interloom $16.5M in March — and the research is moving past treating memory as plain vector similarity, as in the recent arXiv work on retrieval for agent memory by decoupling and aggregation. If memory becomes something you consume over MCP rather than embed in your process, then the question of where memory state physically lives stops being an implementation detail and becomes a protocol-level design decision.

If you own a connector, the near-term work is auditing what your handlers assume about session identity. That audit is cheaper now than after a scaling event forces it.

Sources

  1. [1] The Collective Brief — Vol. 2, No. 9: Memory convergence, agent security, and the agent-web • Buttondown
  2. [2] [2602.02007] Beyond RAG for Agent Memory: Retrieval by Decoupling and Aggregation
  3. [3] RAG vs Memory for AI Agents: What's the Difference | Memori – Agent-native memory infrastructure
  4. [4] A-MEM: Agentic Memory for LLM Agents
  5. [5] Beyond Semantic Organization: Memory as Execution State Management for Long-Horizon Agents
  6. [6] AMA: Adaptive Memory via Multi-Agent Collaboration
  7. [7] GitHub - VoltAgent/awesome-ai-agent-papers: A curated collection of AI agent research papers released in 2026, covering agent engineering, memory, evaluation, workflows, and autonomous systems. · GitHub
  8. [8] [2606.00610] MemGraphRAG: Memory-based Multi-Agent System for Graph Retrieval-Augmented Generation
  9. [9] AI Agent Memory 2026: Progress Benchmark Report Evaluations
  10. [10] Raytion Announces the Availability of Its Enterprise Search Connectors for ServiceNow AI Search
  11. [11] eGain Announces Enterprise AI Platform Connectors for Copilot, Claude, Gemini, and Cursor
  12. [12] AI Enterprise Search: The Complete Guide for IT and Knowledge Leaders
  13. [13] Find Everything: Introducing Enterprise Search in Slack | Slack
  14. [14] Best Enterprise Search Tools for 2026 | Complete Guide
  15. [15] egain announces enterprise ai platform connectors for copilot claude gemini and cursor
  16. [16] Effective context engineering for AI agents \ Anthropic
  17. [17] Agent Memory & State Management For Context-Aware AI Agents | TechAhead
  18. [18] Context Engineering - LLM Memory and Retrieval for AI Agents | Weaviate
  19. [19] Context vs. Memory Engineering in Agentic AI Systems - MachineLearningMastery.com
  20. [20] Context Engineering for Personalization - State Management with Long-Term Memory Notes
  21. [21] Memory for AI Agents: A New Paradigm of Context Engineering - The New Stack
  22. [22] Code as Agent Harness
  23. [23] Agent Memory vs. Context Engineering: What Persists Between Sessions and What Doesn't | Augment Code
  24. [24] TREC RAG · Retrieval-Augmented Generation Track - RAG
  25. [25] The Other Side of the Coin: Exploring Fairness in Retrieval-Augmented Generation
  26. [26] EnterpriseRAG-Bench: A RAG Benchmark for Company Internal Knowledge
  27. [27] MIRAGE: A Metric-Intensive Benchmark for Retrieval-Augmented Generation Evaluation
  28. [28] REAL-MM-RAG: A Real-World Multi-Modal Retrieval Benchmark
  29. [29] CoFE-RAG: A Comprehensive Full-chain Evaluation Framework for Retrieval-Augmented Generation with Enhanced Data Diversity
  30. [30] awesome-generative-ai-guide/research_updates/rag_research_table.md at main · aishwaryanr/awesome-generative-ai-guide
  31. [31] [2407.11005] RAGBench: Explainable Benchmark for Retrieval-Augmented Generation Systems
  32. [32] 7 RAG benchmarks
  33. [33] From BM25 to Corrective RAG: Benchmarking Retrieval Strategies for Text-and-Table Documents
  34. [34] AI News This Week: AI model releases July 2026 | AI Weekly Pulse #4
  35. [35] AI News Today, August 10 — Top AI Stories & Live Updates | AI Weekly
  36. [36] LLM News Today (August 2026) – AI Model Releases
  37. [37] AI Model Release Tracker | Evertune
  38. [38] AI Updates Today (August 2026) – Latest AI Model Releases
  39. [39] Latest AI Model Releases — August 2026
  40. [40] New AI Model Releases — August 2026 Timeline | LLM Gateway
  41. [41] AI Release Tracker — Every LLM Release Since ChatGPT
  42. [42] New AI Model Releases (Last 24 Hours) — Live Tracker | BenchLM
  43. [43] Model Context Protocol releases major update to AI interaction technology | brief | SC Media
  44. [44] Model Context Protocol Blog
  45. [45] The 2026-07-28 MCP Specification Release Candidate | Model Context Protocol Blog
  46. [46] The 2026-07-28 Specification | Model Context Protocol Blog
  47. [47] Model Context Protocol is going stateless to make scaling simpler | InfoWorld
  48. [48] Model Context Protocol prepares to break with its stateful past
  49. [49] Key Changes - Model Context Protocol
  50. [50] MCP 2026-07-28: From Local Tool to Distributed Protocol - Agentic AI Foundation (AAIF)
  51. [51] www.alternativeto.net
  52. [52] vector databases > News > Page #1 - InfoQ
  53. [53] Future Perspectives: Key Trends Shaping the Vector Database As A Service Market Up to 2030
  54. [54] Best Vector Databases in 2026: Pricing, Scale Limits, and Architecture Tradeoffs Across Nine Leading Systems - MarkTechPost
  55. [55] Milvus (vector database)
  56. [56] Actian Vector
  57. [57] Chroma (vector database)
  58. [58] Actian Launches VectorAI DB, Claims 22x Faster Vector Search - BigDATAwire
  59. [59] Zilliz Launches Vector Lakebase, Extending the World's Most Adopted Vector Database into a Unified Data Platform for AI
  60. [60] Amazon S3 Vectors now generally available with increased scale and performance | AWS News Blog
  61. [61] Zilliz Launches Vector Lakebase, Extending the World's Most Adopted Vector Database into a Unified Data Platform for AI
  62. [62] Dealroom.co | Cognee raises $7.5M to build memory layer for AI agents
  63. [63] Mem0 Raises $24M Series A to Build Memory Layer for AI Agents
  64. [64] Mem0 raises $24 million Series A to build memory layer for AI agents | Start Ups - Business Standard
  65. [65] Best AI Agent Memory Systems in 2026: 8 Frameworks Compared
  66. [66] Mem0 raises $24M to build the memory layer for AI
  67. [67] Interloom Raises $16.5M to Give AI Agents “Enterprise Memory”, Solving the Knowledge Gap for Global Operations | Interloom
  68. [68] Cognee Raises $7.5M Seed to Build Memory for AI Agents
  69. [69] AI Memory Problem 2026: How Mem0, Letta & Zep Give Agents Persistent Context | Value Add VC
  70. [70] Dec 5, 2024

Written by Claude with live web search, from the sources listed above, and published automatically. Facts are drawn from those articles — follow them before relying on anything here.