Intlo BrainIntlo Brain

October 9, 2026

O'Reilly Ships Its Library as an MCP Retrieval Endpoint

A licensed corpus exposed as an MCP server reframes provenance as an infrastructure problem, not a prompt-engineering one.

The more interesting retrieval news this week wasn't a model or a vector index. It was a publisher deciding that the right delivery format for its corpus is an MCP server.

O'Reilly announced Expert Intelligence on October 7, an organization-wide suite that grounds enterprise AI in its repository of practitioner knowledge and calls on job-specific agent skills, with responses verified through citations from named authors. The product surface is an "Expert MCP" that pulls context from the knowledge repository, plus agent skills, and it connects with any MCP-compatible tool. The pitch is that it lives inside the AI tools organizations already use, so there's no separate platform to roll out.

The attribution argument

The framing is worth reading closely because it's a retrieval-quality argument, not a content-licensing one. O'Reilly cites a February 2026 benchmark across 60 questions drawn from live production traffic, finding roughly one in five claims from leading AI tools unsupported by a cited source, about two-thirds of cited sources naming no author, and around half carrying no publication date. The stated fix is a direct, secure connection to the curated corpus so every response traces to a named expert and a confirmed publication date.

If you build RAG, you already know why author and date are load-bearing. A chunk from a 2019 blog post and a chunk from a current reference manual look identical in embedding space. Recency and authorship are metadata problems, and most ingestion pipelines either discard those fields during chunking or never had them because the source was scraped HTML. A curated corpus with enforced provenance fields is less a knowledge advantage than a metadata-integrity advantage — which is the part teams consistently underinvest in.

Why MCP is the right-sized abstraction here

Shipping as an MCP server rather than a bulk export or an embeddings dump is the real design choice. The content never lands in your index, so licensing stays enforceable per-query, and the retrieval logic stays with the party that understands the corpus structure. The tradeoff is familiar: you give up control over chunking, ranking, and hybrid-search tuning, and you add a network hop with someone else's latency and uptime. For a reference corpus you don't own, that's usually the correct trade. For your own Confluence and ticket history, it isn't.

The surrounding evidence

Context from October 5: MIT Technology Review Insights published a Neo4j-sponsored report surveying 300 data, AI, and technology executives on their ability to give agents full contextual understanding across semantic knowledge, episodic memory, and procedural knowledge. It reports only 34% of agentic AI projects reaching production, with data fragmentation and missing context as major blockers, and higher production rates among organizations with stronger semantic knowledge capabilities. It's vendor-sponsored, so treat the numbers as directional. The stated investment priorities — ingestion pipelines, AI-ready APIs, RAG, evaluation agents, and knowledge graphs — still describe where the work actually is.

Who should care: anyone whose agents cite sources users will check. The provenance layer is becoming a product boundary, and external corpora are arriving as protocol endpoints rather than data dumps.

Sources

  1. [1] 15 Best Enterprise Search Software in 2026 - A Complete Guide
  2. [2] Best Enterprise Search for Engineering Teams (2026)
  3. [3] Gemini Enterprise release notes
  4. [4] Best Enterprise Search Tools for 2026
  5. [5] The AI Enterprise Search Guide for IT and Knowledge Leaders
  6. [6] Best Enterprise Search Solutions 2026: Semantic & AI-Native
  7. [7] Enterprise Search Options 2026
  8. [8] Enterprise AI Search: AI-Powered Search for 2026
  9. [9] AnySearch AI Search Platform Launches for Enterprise Agents
  10. [10] Enhancing LLM Performance with Retrieval-Augmented Generation
  11. [11] Retrieval Augmented Generation Market to Reach USD 47.00 Billion by 2035, Growing at 37.58% CAGR as Enterprises Scale Generative AI
  12. [12] Deeper insights into retrieval augmented generation: The role of sufficient context
  13. [13] News from generation RAG - Dive deep into the transformative world of AI Retrieval Augmented Generation (RAG) technologies
  14. [14] RAGMail: a cloud-based retrieval-augmented framework for reducing hallucinations in LLM text generation
  15. [15] Top Information Retrieval Papers of the Week
  16. [16] Top Information Retrieval Papers of the Week
  17. [17] This thread is only visible to paid subscribers of Top Information Retrieval Papers of the Week
  18. [18] site.ieee.org
  19. [19] RAG vs Agent Memory: What Each Does and When to Combine Them — supermemory
  20. [20] Knowledge and Memory Beyond RAG: Why 2026 Agents Need a Write Path, Not Just a Retriever
  21. [21] RAG vs Agent Memory: What Changes When Agents Need to Remember? - GeeksforGeeks
  22. [22] Agents Don't Need Memory, They Need Docs - AgentConn Blog
  23. [23] [2602.02007] Beyond RAG for Agent Memory: Retrieval by Decoupling and Aggregation
  24. [24] Weekly ML Roundup: Agentic RAG in SQL and Durable Agent Memory - Tech Hub
  25. [25] Memory (EP)
  26. [26] past.dev Introduces Frontier Memory for AI Agents: #1 on BEAM, the Largest Public Memory Benchmark
  27. [27] [2606.00610] MemGraphRAG: Memory-based Multi-Agent System for Graph Retrieval-Augmented Generation
  28. [28] Machine Learning Pills
  29. [29] ACER: Automatic Language Model Context Extension via Retrieval
  30. [30] Beyond Pipelines: A Survey of the Paradigm Shift toward Model-Native Agentic AI
  31. [31] Graph World Model
  32. [32] Context Length Alone Hurts LLM Performance Despite Perfect Retrieval
  33. [33] [2310.03025] Retrieval meets Long Context Large Language Models
  34. [34] AI Context Window Comparison 2026: 1M to 10M Tokens
  35. [35] Long-Context Retrieval 2026: Needle-in-Haystack Test
  36. [36] AI Model Context Window Comparison 2026: Advertised vs. Real - elvex
  37. [37] Best AI for Long Context 2026 - Top Long Context Models
  38. [38] Scaling AI Agents: Key Takeaways from the Model Context Protocol (MCP) Specification Release
  39. [39] Specification - Model Context Protocol
  40. [40] Overview - Model Context Protocol
  41. [41] Scaling AI Agent Infrastructure with the MCP Stateless updates - Google Developers Blog
  42. [42] Model Context Protocol
  43. [43] Roadmap - Model Context Protocol
  44. [44] Key Changes - Model Context Protocol
  45. [45] Announcing v2.0 of the official MCP C# SDK - .NET Blog
  46. [46] Model Context Protocol Specification Version Timeline - Version-by-Version Changes and Adoption Milestones
  47. [47] AI News for the Week of October 9; Updates from Honeycomb.io, NinjaTech AI, Oracle & More - Solutions Review
  48. [48] AI News for the Week of October 2; Updates from Honeycomb.io, NinjaTech AI, Oracle & More
  49. [49] Connecting AI agents to enterprise knowledge
  50. [50] AI News Today, October 8: Top Stories
  51. [51] AI News
  52. [52] AI News
  53. [53] O’Reilly Launches Expert Intelligence to Ground Enterprise AI in Practitioner Knowledge - AIwire
  54. [54] Glean Technologies
  55. [55] AI news October 2026: four developments · Hello Growth
  56. [56] Connecting AI agents to enterprise knowledge
  57. [57] Connecting AI agents to enterprise knowledge · Issue #1334 · hanzhad/squelch-news-engine
  58. [58] Redefining enterprise intelligence with autonomous AI
  59. [59] Connecting AI agents to enterprise knowledge
  60. [60] Bridging the Knowledge Gap for Successful AI Project Deployment
  61. [61] AI agents, prediction, and the trust problem for IT teams
  62. [62] Agentic commerce
  63. [63] Enterprise AI agents lack sufficient organizational knowledge for autonomous decision-making — Agentic Ready
  64. [64] O'Reilly Launches Expert Intelligence, Bringing Vetted Practitioner Knowledge Directly into Enterprise AI Stacks and Providing the Skills Framework to Apply It
  65. [65] O'Reilly Expert Intelligence - O'Reilly Media
  66. [66] O'Reilly Media - Technology and Business Training
  67. [67] Building Organizational Intelligence
  68. [68] O'Reilly Launches Expert Intelligence, Bringing Vetted Practitioner Knowledge Directly into Enterprise AI Stacks and Providing the Skills Framework to Apply It
  69. [69] O'Reilly Launches Expert Intelligence, Bringing Vetted Practitioner Knowledge Directly into Enterprise AI Stacks and Providing the Skills Framework to Apply It
  70. [70] Skill: Generative AI

Written by Claude with live web search, from the sources listed above, and published automatically. Facts are drawn from those articles — follow them before relying on anything here.