The LLM Wiki Pattern

A method for building a personal knowledge base where an LLM agent — not you — writes and maintains a structured wiki that sits between you and your raw sources.1 This wiki is an instance of it, formalized on top of the Open Knowledge Format.

The LLM Wiki Pattern: retrieval-only work happens after every question; the LLM wiki compiles it once, at ingest.

The whole argument at a glance. It is dense at page width — open the interactive version to pan and zoom.

The problem with retrieval-only

The default way to point an LLM at a document collection is RAG: upload files, retrieve relevant chunks at query time, generate an answer. NotebookLM, ChatGPT file uploads, and most RAG systems work this way.

It works, but nothing accumulates. The model rediscovers knowledge from scratch on every question. Ask something subtle that requires synthesizing five documents and it has to find and reassemble the same fragments it assembled last time, with no memory that it ever did.1

The inversion

Instead of retrieving from raw documents at query time, the LLM compiles knowledge once and then keeps it current. A new source isn’t merely indexed — it’s read, and its content is integrated into the existing wiki: entity pages updated, topic summaries revised, contradictions with earlier claims flagged, the evolving synthesis strengthened or challenged.1

The payoff is that the work is already done when a question arrives. The cross-references exist. The contradictions are already surfaced. The synthesis already accounts for everything read so far. The wiki compounds — it gets richer with every source added and every question asked.

The division of labor is the point: you curate sources, direct the analysis, and ask good questions. The LLM does the summarizing, cross-referencing, filing, and bookkeeping. You rarely write the wiki yourself.1

Three layers

The pattern separates what nobody edits, what the agent owns, and what the two of you negotiate:

LayerWho writes itRole
Raw sourcesYou (by curation)Immutable source of truth. The agent reads, never modifies.
The wikiThe LLM, entirelyConcept pages, summaries, comparisons, cross-references. You read it.
The schemaYou and the LLM togetherThe config that makes the agent a disciplined maintainer instead of a generic chatbot.

The schema is the load-bearing piece. It’s an ordinary agent instruction file (CLAUDE.md for Claude Code, AGENTS.md for Codex) describing structure, conventions, and workflows — and it co-evolves as you learn what works for your domain.1 In this bundle it’s CLAUDE.md, which is deliberately excluded from the published site.

The day-to-day loop over these layers — ingest, query, lint — is covered in wiki operations.

Why it holds up

The tedious part of a knowledge base was never the reading or the thinking. It was the bookkeeping: updating cross-references, keeping summaries current, noticing when new data contradicts an old claim, staying consistent across dozens of pages. Humans abandon wikis because that maintenance burden grows faster than the value.1

An LLM doesn’t get bored, doesn’t forget a cross-reference, and can touch fifteen files in one pass. The wiki stays maintained because maintenance costs close to nothing.

Karpathy places the idea in line with Vannevar Bush’s 1945 Memex — a private, actively curated store where the trails between documents matter as much as the documents. Bush’s unsolved problem was who maintains the trails.1

Where it applies

The pattern is domain-agnostic. Karpathy’s examples: personal tracking (goals, health, psychology), long-running research, reading a book (characters, themes, plot threads — a private version of a fan wiki like Tolkien Gateway), team wikis fed by Slack threads and meeting transcripts, and anything else accumulating over time — competitive analysis, due diligence, course notes, hobby deep-dives.1

Deliberately underspecified

The gist describes an idea, not an implementation. Directory structure, page formats, schema conventions, and tooling are all left to the instantiation — the intent is that you hand the document to your agent and work out a version that fits your domain.1 Everything in it is modular: text-only sources need no image handling, a small wiki needs no search engine beyond index.md.

Optional tooling it does mention: qmd for on-device hybrid BM25/vector search once the index outgrows itself, Obsidian Web Clipper for capturing sources, Obsidian’s graph view for spotting hubs and orphans, Marp for slide output, and Dataview for querying frontmatter. This bundle currently uses none of them — the wiki is small enough that index.md suffices.

Footnotes

  1. Andrej Karpathy, “LLM Wiki” (gist). Local copy: references/karpathy-llm-wiki.md. 2 3 4 5 6 7 8 9