ArchitectureLegal Services

AI search over a firm's precedents and know-how: an architecture that lawyers can trust

AI knowledge search for a law firm is retrieval-augmented generation over the firm's own documents: it finds the right precedent clause, prior advice or practice note and answers with pinpoint citations to it. The hard parts are not the model. They are deciding which documents count as authoritative, carrying matter-level permissions and ethical walls into every query, keeping superseded templates out of answers and testing for invented authority before lawyers rely on it.

Reviewed 6 min read

On this page
  1. What lawyers actually ask a knowledge search tool
  2. Five layers of a permission-aware precedent search system
  3. Choosing the corpus: curated know-how, all matter documents or both
  4. Ingesting from the document management system without losing context
  5. One precedent query, from question to cited answer
  6. Keeping answers current and testing for invented authority
  7. A hypothetical search for negotiated liability caps in one sector
  8. Design risks specific to legal knowledge search
  9. Questions and answers
  10. Sources

What lawyers actually ask a knowledge search tool

The questions are narrower than a general chat assistant suggests. Lawyers want the firm's standard clause for a particular point and the fallback positions it has agreed before; the last piece of advice the firm gave on a regulatory question; the practice note explaining how a procedure works; or the closing checklist used on a similar deal. They want the document itself, not a fluent paraphrase of it.

That shapes the architecture. Answers must point to a specific document and passage the lawyer can open. The tool must know which documents are the firm's approved positions and which are one-off negotiated outcomes. And it must never show a lawyer a document they could not open in the document management system, because an ethical wall breached through a search box is still a breach. Playbook-based review of an incoming contract is a different job, covered by the AI contract review use case.

Five layers of a permission-aware precedent search system

Lawyer interface01Grounded answering02Permission-filtered search03Curated index04DMS and KM sources05
  1. Lawyer interface

    Search and drafting panel inside the tools lawyers already use, showing answers with linked citations.

  2. Grounded answering

    A language model drafts only from retrieved passages, cites each one and declines when support is weak.

  3. Permission-filtered search

    Hybrid keyword and semantic search, filtered by the user's DMS access and ethical walls before ranking.

  4. Curated index

    Chunks with metadata: practice area, document type, status, approval date, matter and access list.

  5. DMS and KM sources

    The document management system, the know-how library and the matter system that hold the originals.

Conceptual layering of a legal knowledge search system. Real deployments may merge layers; it does not describe any particular firm or product.

Choosing the corpus: curated know-how, all matter documents or both

Corpus scope decides answer quality and risk more than model choice does.

ConsiderationCurated KM corpus onlyAll matter documentsTiered: curated first, matter documents labelled
Reliability of answersHigh: approved precedents and notesMixed: drafts, superseded and negotiated versionsHigh for curated hits; matter hits marked as examples
CoverageLimited to what KM has written upVery broad, including rare pointsBroad, with the trust level shown
Permission complexityLow: mostly firm-wide accessHigh: matter access lists and ethical walls throughoutHigh for the matter tier
Effort to keep currentOwned by KM lawyersHard: nobody curates matter filesKM owns the top tier; status metadata drives the rest
Best fitFirst release, or firms with strong KM librariesRarely the right starting pointSecond phase, once permissions are proven

Most firms will start with the curated corpus and add a labelled matter tier later.

Ingesting from the document management system without losing context

A firm's document management system, for example iManage or NetDocuments, holds versions, document types, authors, matter numbers and access controls. Ingestion should carry all of that into the index, not just the text. Version history tells the system which draft was final; document type separates an executed agreement from a markup; the matter number links each passage back to its access list.

Chunk documents along their legal structure, by clause, schedule or heading, rather than by fixed length, so a citation lands on a whole clause. Keep defined terms and cross-references with the chunk where possible, because a limitation clause without its definition of losses can mislead.

Permissions deserve one firm rule: the index stores each document's access list and the retrieval layer filters by the requesting user's identity and ethical-wall membership before anything is ranked or shown. The SRA's guidance treats systems that prevent access to confidential information, including separate access controls, as part of effective measures for information barriers1. The implementation pattern for filtering by identity at query time is covered in detail on our permission-aware retrieval page.

One precedent query, from question to cited answer

questioncheck access and wallsfiltered searchpermitted passagespassages and questionanswer with citationsverify citations existrecord sourcesanswer and links01Lawyer02Search service03Index04Language model05Audit log
  1. Lawyer

    Asks for the firm's position on a clause.

  2. Search service

    Resolves the user's identity and wall memberships.

  3. Index

    Returns only passages the user may see.

  4. Language model

    Drafts an answer from the passages supplied.

  5. Audit log

    Records the query, sources shown and user.

  1. Lawyer to Search servicequestion
  2. Search service to Search servicecheck access and walls
  3. Search service to Indexfiltered search
  4. Index to Search servicepermitted passages
  5. Search service to Language modelpassages and question
  6. Language model to Search serviceanswer with citations
  7. Search service to Search serviceverify citations exist
  8. Search service to Audit logrecord sources
  9. Search service to Lawyeranswer and links
Conceptual sequence for a single query. Filtering happens before the model sees any text, and citations are checked against retrieved passages before display.

Keeping answers current and testing for invented authority

Build these into the release gate and into monthly monitoring.

0 of 7 checked

A hypothetical search for negotiated liability caps in one sector

Questions and answers

Should a firm build knowledge search on its document management system or on a separate index?

Usually a separate search index fed from the document management system. The DMS remains the system of record for documents and permissions, while the index adds chunking, embeddings and metadata such as document status. The key requirement is that the index keeps each document's access list synchronised with the DMS, so permission changes take effect in search quickly.

How is AI knowledge search different from AI legal research tools?

Research tools search published law: legislation, case law and commentary from legal publishers. Knowledge search covers the firm's own work product: precedents, prior advice, practice notes and executed documents. The two complement each other, but they carry different risks. Knowledge search must enforce client confidentiality and ethical walls; research output must be verified against authoritative sources before it is relied on.

Can the system draft from precedents as well as find them?

Yes, if drafting stays grounded. The safer pattern is to retrieve the firm's approved clause and adapt it to the facts the lawyer supplies, showing the source clause alongside the draft so the lawyer can see every change. Drafts produced without a retrieved source should be clearly marked as such or blocked for practice areas where the firm requires precedent-based drafting.

Sources

  1. Confidentiality of client information (guidance) — Solicitors Regulation Authority · checked 10 October 2026

More in Legal Services

Back to Legal Services

Next step

Scope a first precedent search corpus with your KM team

Tell us which document management system you use, how your know-how library is organised and the questions lawyers ask most often. We will propose a starting corpus, the permission model it needs and a test set your knowledge lawyers can own.

Discuss your knowledge corpus