Modern knowledge management · museums, libraries & collections

Your collection knows more than any exhibition — or any catalogue — has ever shown.

A museum's greatest asset was never the objects, and a library's was never the shelves. It's what the institution knows about them — and almost all of that knowledge is currently unreachable, both to the public and to your own staff.

Neo Gens Co., Ltd. For museum and library directors, and collection leadership 16 min read

A visitor stands in front of a ceramic bowl. The label gives them forty words: what it is, roughly when it was made, and the name of the donor. They read it in eight seconds and move on.

Behind that bowl sits an enormous amount of knowledge. A curator spent three years establishing where it actually came from. There is a conservation report describing a repair that tells you something about how the object was used. There is correspondence in the archive from the collector who acquired it, mentioning the village he bought it in. A researcher published a paper connecting its glaze to a kiln site two hundred kilometres away. Somewhere in the textiles department is a garment collected in the same expedition, by the same person, in the same month.

None of that reaches the visitor. Most of it doesn't reach your own staff either. It exists — in a collection management system, in a filing cabinet, in a PDF, in a departmental spreadsheet, and above all in the memory of people who will eventually retire.

This is not a technology failure. Museums are, historically, the most disciplined knowledge-management institutions humanity has built. Cataloguing, provenance research, accession records, the practice of citing your evidence and distinguishing an attribution from a certainty — museums invented this rigour centuries before anyone used the phrase "data governance." The discipline is already there. What's missing is a layer that lets the discipline connect.

The museum's problem isn't that it doesn't know. It's that what it knows can't reach across itself.

The model, borrowed from a much smaller collection

We built our first knowledge system for a personal library — one person's books, catalogued properly. It sounds trivial next to a national collection. It wasn't, and the reason is instructive.

The rule we imposed was simple: nothing is stored as a bare fact. Every statement about every object is held as a claim, carrying who asserted it, on what evidence, and at what status. A publication date isn't "1962." It's "1962, asserted by two independent references, corroborated." An attribution that rests on a single dealer's catalogue stays visibly marked as a single-source claim, forever, until something better arrives.

Books were chosen as the first domain because a barcode on a physical object gives you ground truth you can hold in your hand. It was the hardest possible test bed for a claim-and-evidence model — and everything the model needed to survive there, a museum needs more.

Because a museum object is exactly this structure, only richer. Look at what your accession record already contains and you'll find the same four properties finance demands of any asset:

  • A register — you know what you hold, and where it physically is.
  • Provenance — where it came from, through whose hands, on what documentation.
  • Valuation — not only insurance value, but scholarly and interpretive significance, which changes over time.
  • An audit trail — every claim about the object traceable to the research that produced it.

Museums already do all four, better than most industries. The difference is that they do it in prose — in reports, cards, files and expert memory — and prose cannot be traversed. You cannot ask a filing cabinet a question that crosses three departments.

What changes: the object stops being a record and becomes a node

The shift is not from paper to digital. Most museums made that shift long ago and were disappointed by it, because digitising a catalogue card produces a digital catalogue card. The shift is from storing descriptions of objects to modelling the relationships between them.

Today · the catalogue record
Glazed stoneware bowl
Acc. 1974.221 · Ceramics · store 4, bay 12

Stoneware with celadon glaze. Probably 14th century. Acquired 1974, gift of a private collector. Condition: repaired rim. See conservation file.

Connected · the same object as a node
Glazed stoneware bowl
Acc. 1974.221 · 9 relationships · 4 departments

The same record, plus every relationship the institution has ever established around it — each one carrying its source.

Illustrative example. Nothing new was researched to produce the right-hand version — every relationship shown already existed somewhere in the institution. It had simply never been expressed in a form a question could travel through.

This is the whole of it. You are not being asked to generate new knowledge. You are being asked to stop losing the knowledge you already produced. The connections above were made by real curators, conservators and researchers over decades. They exist in documents that no longer speak to each other.

Once the objects are nodes and the relationships are explicit, something becomes possible that was not possible before: the collection can be asked a question — by a curator, by an exhibition designer, or by a fourteen-year-old standing in gallery three.

SAME EXPEDITION GLAZE ANALYSIS DONOR PAPERS USE EVIDENCE Stoneware bowl ACC. 1974.221 · CERAMICS Woven textile TEXTILES DEPARTMENT Kiln site research paper LIBRARY · PUBLISHED 1998 Collector correspondence INSTITUTIONAL ARCHIVE Conservation report CONSERVATION · 2011
Four departments that rarely share a system, connected through one object. The horizontal links are the interesting ones — a textile and a research paper that have a relationship neither department knew about.

The visitor experience this makes possible

Here is where it stops being an infrastructure conversation and becomes a visitor one.

When the collection is connected, a visitor can put a question to an object and receive an answer drawn from your institution's own scholarship — with the curator's research attached to it. Not a plausible paragraph generated by a language model that has never seen your catalogue. An answer assembled from records that each carry a source, in the voice of an institution that knows the difference between what it has established and what it merely suspects.

Three visitors, three very different questions, same afternoon:

Illustrative exchanges. The third is the most important one — and the one no conventional visitor-facing chatbot will ever produce.

Notice what happened in the third exchange. The system did not invent an answer. It said what the museum knows, said what it doesn't, and explained why the gap exists. That is not a failure of the experience — for a curious visitor it is the single most memorable thing a museum can say, because it is an invitation into the actual practice of scholarship.

This behaviour is not a matter of tone or prompt wording. It's structural. If every fact in the system carries its evidentiary status, then "we don't know" is simply what the system returns when it reaches a node with no corroboration. An institution built on provenance gets honesty for free.

Which is why this is a back-of-house project

It is tempting to treat a conversational visitor experience as an interpretation project — a front-of-house commission, scoped like an audio guide. It isn't, and the distinction matters more in your sector than almost anywhere else.

A visitor-facing AI is not an exhibit. Exhibitions close. Answers don't. Every answer it produces is an institutional statement, issued in your name, hundreds of times a day, without a curator reading a word of it first. A model attached to a pile of PDFs will answer fluently and will eventually invent a provenance — and the correction will be public, and it will be yours to make.

An institution that exists to hold society's knowledge cannot let statements without provenance, without evidence, without a rule that governs them, go out under its name — not to a visitor, not to a school group, not to a journalist. That obligation doesn't disappear because the statement was generated rather than written.

So the sequence is not negotiable: the experience is a consequence of the layer, never a substitute for building it. Prepared properly, the constraint becomes the most distinctive feature of the whole thing — an institution whose AI is structurally incapable of overclaiming is offering visitors something no general-purpose assistant can.

A museum that can say "we don't know, and here is why" is more trustworthy than one that always has an answer. Visitors can feel the difference.

Exhibitions that assemble themselves — and a curator who edits rather than searches

Exhibition development currently begins with a curator holding an idea and then spending months finding out what the museum actually has that speaks to it. That search is limited by one thing: how much of the collection any single person can hold in their head. A curator who has worked in ceramics for twenty years will assemble a superb ceramics exhibition and will never know that the perfect supporting object has been sitting in the textiles store the whole time.

On a connected collection, that inverts. The curator poses the theme — everything in this institution that passed through a particular trade route, every object acquired during a single decade of colonial expansion, everything showing evidence of repair and reuse — and the collection proposes the set, across every department, with the evidence for each inclusion attached.

The curator's judgement is not replaced. It's moved to where it's actually valuable: deciding what the exhibition means, rather than spending months discovering what exists. And every proposed connection can be walked back to the research that supports it, so nothing enters a gallery on the strength of a machine's suggestion alone.

What this does to repeat visitation

The economics of a museum turn on a hard question: why would someone come back? The permanent collection is, by definition, permanent. Most institutions answer with temporary exhibitions, which are expensive and infrequent.

A connected collection offers a different answer. The objects don't change — but the paths through them do, and they can be recomposed continuously from what the institution already owns:

  • A different thread each visit. The same forty galleries, entered through the story of trade, or of repair, or of women collectors, or of a single decade — each thread drawn from existing scholarship, not newly written.
  • A route shaped by what you asked last time. A returning visitor who spent twenty minutes on Japanese lacquer is offered the objects that connect to it, most of which are not in the lacquer gallery.
  • Depth on demand. The same object answers a nine-year-old and a specialist differently, because the knowledge behind it is layered rather than flattened into a single label.
  • Storage brought into play. The majority of most collections is never on display. Connected, those objects can participate in a visit even when they can't physically be in the room.

The part that compounds: every question is an asset

Now the loop closes, and this is the part that changes a museum's trajectory rather than just its visitor experience.

Every day, visitors ask your institution hundreds of questions. Today those questions evaporate. Front-of-house staff field them, curators occasionally hear a good one, and none of it accumulates.

On a knowledge layer, each question is recorded against the object it was asked about — and questions the collection could not answer are the most valuable output of the entire system. They are a research agenda, generated continuously by the public, ranked by genuine curiosity rather than by academic convention.

DAY 1

A visitor asks

"Who actually made this, not who collected it?" — asked of an object whose maker was never recorded.

SAME DAY

The gap is logged

The system answers honestly and records an unanswered question against that accession number.

OVER TIME

A pattern appears

The same gap is hit two hundred times across a category of objects. That's now a visible, ranked research priority.

NEXT SEASON

The answer enters

A curator researches it, the finding enters as a sourced claim, and every future visitor gets it.

The compounding loop. Knowledge enters through the front door as well as the back — but it only ever becomes fact through curatorial judgement, never automatically.

This is the sense in which a museum's knowledge can grow rather than merely accumulate. Not because a machine generates it — because the public tells you, every single day, where the interesting holes are, and the institution finally has somewhere to write that down.

There's a failure mode we watch for and design against: a system that only surfaces what people already engage with becomes an echo chamber, quietly burying the difficult and unfashionable parts of a collection. So disagreement is treated as a first-class citizen — competing attributions are held side by side rather than resolved into a false consensus, and the system is built to ask periodically what in the collection contradicts a given interpretation.

What this is worth to the institution

Stewardship

Expertise stops leaving with people

The reasoning behind an attribution — not just its conclusion — is captured while the person who made it is still here to explain it.

Programming

More exhibitions from the same collection

Thematic routes recomposed from existing scholarship cost a fraction of a loan exhibition and can change several times a year.

Research

A visible, evidence-ranked research agenda

Gaps become countable. That is unusually persuasive material for a grant application or a research partnership.

Development

Donor stories that hold up

Provenance chains you can show, object by object, with sources — increasingly a condition of serious philanthropy, not a nicety.

Access

The stored collection participates

Objects that will never fit in a gallery can still be part of a visit, a school programme or an online route.

Governance

Answers that survive scrutiny

For restitution enquiries, insurance, or a board asking how a claim was established, the trail is already there.

A catalogue tells you what an institution holds. It has never been able to tell you what the institution knows.

The same argument, one building over: libraries

A reader comes to the enquiry desk with a question that sounds simple. They are researching the textile trade in this town, and they want to know what the library has. The librarian does three things: runs a catalogue search, mentions that local studies is on the second floor, and suggests emailing the archivist. Three answers to one question — and no way to tell the reader whether those three bodies of material overlap, contradict each other, or describe the same family twice under two spellings of a name.

This is not a failure of the librarian, and it is emphatically not a failure of discipline. If any institution should have been spared this problem, it is the library. Libraries formalised the organisation of knowledge before the phrase existed — authority control, controlled vocabularies, cataloguing rules maintained across generations and shared between institutions that had never met. Libraries solved interoperability while everyone else was still arguing about file formats.

Which is precisely why the gap is easy to miss. The discipline is exhaustive about the record and nearly silent between records. A catalogue entry describes an item with remarkable precision and tells you almost nothing about how that item stands in relation to the pamphlet in the archive, the oral history in the digital repository, or the annotated copy in special collections. The record is the unit, and the record stops at the item.

Most libraries have already bought something meant to close this. A discovery layer searches every store at once and returns a single ranked list. It is genuinely useful, and it is not the same thing: a longer list of places to look is not an answer. Federating search across four silos leaves you with four silos and one search box.

General collections

Described precisely, related to nothing

Everything catalogued to standard, with subject headings that group titles by topic — and no way to express that this monograph rests on the archive holdings two floors below it.

Special collections

The material readers most want, described least

Rare books, annotated copies, association copies. The scholarship establishing why a particular copy matters usually lives in a curator's head, a finding aid, or a decade-old accession note.

Archives & manuscripts

Described at the box, found at the box

Arrangement by provenance and hierarchy is the right model for archives. It is also a different model from the catalogue, which is why a reader generally has to know the material exists before they can find it.

Digital repository

Preserved, cited, and stranded

Digitised runs, institutional research output, born-digital deposits. Preservation is solved. Connection to the physical holdings that share their subject, period or provenance is not.

The change is the one the museum makes, in different furniture. The record stops being a description and becomes a node: a person is one person whether they appear as an author, a correspondent in a manuscript collection, a donor in an accession file, or a subject in a local history pamphlet. A place is one place across four spellings and two centuries. Once those identities are asserted once, with evidence, the reader's question crosses the stores by itself — and so does the next reader's.

Libraries also hold one asset in this that museums do not, and they currently discard nearly all of it. The reference interview is the richest available record of what a collection is being asked for, and it survives nowhere. A question arrives at the desk, a librarian answers it out of expertise and memory, and the reasoning — which sources they tried, which failed, why they went to the archive instead of the catalogue — evaporates the moment the reader leaves. Captured as claims and gaps rather than as door counts, that traffic becomes a standing account of what readers want and what the collection cannot yet answer. Which is also, precisely, the evidence a collection development case has always struggled to produce.

  • Subject expertise stops retiring. The specialist who knows why this collection is shaped the way it is can be recorded reasoning, not merely recommending.
  • Special collections begin to earn. Material currently found only by readers who already knew it was there becomes reachable from an ordinary subject question.
  • Collection development gains evidence. Gaps become countable, and countable gaps are what funding cases are built from.
  • Teaching and holdings meet. What a department is teaching this term can be related to what the library actually holds — including the parts of it that are not books.

Everything from here applies to both kinds of institution. The honesty below is the same honesty, and so is the starting point: one question your institution already wishes it could answer.

Where we're honest with you

Before anyone signs anything
  • This does not replace curators or librarians, and cannot. The system holds claims and their evidence. Only a specialist promotes a claim to an established fact. Any vendor telling you a machine can do that scholarship is selling you a liability.
  • Your specialists' time is the real cost. Not fees — calendars. The knowledge worth modelling is inside your people, and it cannot be extracted from documents. If that time can't be committed, the work won't hold, and we'd rather say so now than in month four.
  • Automatic classification is not solved. It gets you part of the way and then needs human judgement. We won't claim otherwise.
  • Start narrow or don't start. One collection, one real question, end to end. Institution-wide programmes that model everything at once are how these projects die.
  • Your knowledge stays yours. The layer runs inside your institution's boundary. Sensitive material — donor terms, valuations, restitution correspondence — is never handed to a public model, and your team owns the result outright.

Where an institution starts

Not with a technology selection. With a question your institution already wishes it could answer.

01

Pick one question you couldn't answer last year

Something a curator or a subject librarian genuinely tried to research and abandoned because the collection couldn't be searched that way.

02

Model the objects it touches — with your curators in the room

Not the whole collection. The slice that question crosses, usually spanning two or three departments.

03

Connect what already exists

Conservation reports, archive correspondence, published research, past exhibition texts. Almost nothing here is new work.

04

Answer the question, then show the board the trail

The answer matters less than the fact that every step of it can be walked back to a source.

05

Let your team do the next collection

That's the actual deliverable. The graph is only the evidence that the capability transferred.

The bowl, revisited

Return to the visitor in front of the ceramic bowl. Nothing about the object has changed. It is the same stoneware, the same repaired rim, the same eight seconds of attention available.

But now the label is a doorway. The visitor can ask why the rim was repaired and learn that somebody, centuries ago, valued this bowl enough to mend it rather than discard it. They can follow that thread to a textile collected on the same expedition and see two objects that left the same village together and have sat in different parts of this building ever since. They can ask who made it and be told, honestly, that the museum doesn't know — and that this is precisely the question the ceramics department is working on, partly because so many visitors have asked it.

Every one of those facts already belonged to the museum. The institution simply had no way to hand them over.

That's what modern knowledge management is for. Not a new system to buy. A way of holding what you already know so that it can finally be given away — to your visitors, to your researchers, and to the colleagues who will run this institution after everyone currently in the building has gone.

Bring us the question you couldn't answer.

90 minutes. Bring your curators or subject librarians, and one question your institution couldn’t answer.