You Pointed RAG at SharePoint. Nobody Owns What Goes Stale.




The RAG pilot looked sharp. SharePoint was wired as the source of truth. Retrieval demos pulled the right policies. Leadership signed off. Six months later the assistant still answers fluently, and half the answers lean on pages nobody has touched since the restructure.
That is the part most RAG programmes underfund. Connecting SharePoint is an integration project. Keeping the corpus true is an operating model project. If you only finish the first one, you did not ship retrieval. You shipped a polite way to recycle stale intranet content.
Pointing a crawler at site collections feels like progress. Folders appear. Chunks land in the vector index. Citations look official. The missing question is quieter: who decides when a document is still fit to answer a customer, a staff member, or a regulator?
SharePoint is usually a collaboration surface, not a governed knowledge product. Drafts sit next to finals. Old versions hang around because deleting them feels risky. Project sites outlive the project. Regional teams keep parallel copies of the same procedure. RAG does not know the politics. It knows similarity.
In Australian enterprises this shows up as a familiar split. IT owns the connector and the index. The business owns the documents in theory. Knowledge management owns the taxonomy on a slide. Nobody owns the weekly decision that a page should be unpublished from the retrieval path.
If no named owner can take a document out of the answer path this week, you do not have a knowledge base. You have a museum with a search bar.
Stale RAG rarely fails with a red error. It fails with confident answers built from the wrong generation of truth:
Your evaluation set from go-live will not catch this for long. The questions stay the same. The documents move. Without freshness checks, the demos keep passing while production quietly drifts.
Picture a national retailer that connected RAG to the policy SharePoint for store managers. Launch week was clean. Then a returns rule changed in one state. The intranet article was updated by the legal team. The PDF attachment in an old site library was not. The assistant kept citing the attachment because it chunked cleanly and ranked high. Store staff trusted the citation. The escalation queue filled with "but the AI said".
Or take a bank ops assistant pointed at procedures libraries. The connector ran nightly. Ownership of each library stayed with whoever created the folder years ago. When a payments workflow changed, the new SOP landed under a different site. The old SOP stayed live in the index for weeks because nobody owned a de-index request.
Neither case is a model failure. Both are corpus governance failures wearing a retrieval UI.
Give RAG the same operational seriousness you give a customer channel.
If your RAG runbook only covers chunk size, embedding model, and top-k, add a second chapter for corpus ownership. Make it as boring and mandatory as access reviews.
Before you expand the SharePoint footprint:
RAG against SharePoint can be excellent. It stays excellent only while someone is paid to keep the corpus honest. The connector was the easy part. Ownership of what goes stale is the product.
