MCPmemory · skills · handoffs

One brain for all your agents & users.

Give Claude, Codex and any MCP client — and the people working beside them — shared knowledge, reusable skills, and a reliable way to hand work to each other.

Shared across every MCP client Private or hosted models Update one skill for every agent FAQs Agents: /brain.md · /llms.txt · mcp.json
levirge brain your workspace · MCP
Levirge Brain — the entity browser with demo data: an Ask bar, Skills / Entities / Docs / Metrics tabs, a tracked-entity sidebar, and a cited per-entity summary
18
MCP tools from one server — knowledge, skills, handoffs
~6 s
per document to categorise — measured Aug 2026, local LLM
0
third-party model calls in private mode
The problem

N brains that all forget, drift, and duplicate

Each agent keeps its own memory and skill files. Knowledge drifts, teams duplicate work, and agents can't pass context between tools. Brain gives them one knowledge base, one skills repository and one handoff system — over MCP.

For teams running more than one agent — across repos, clients and people.

Agents Users Claude Codex OpenCode Any MCP client CI workers You web UI Your team web UI Obsidian sync Levirge Brain
Curation

Turn raw documents into cited knowledge

Plenty of products store what you give them. Brain groups incoming material by entity and produces compact summaries an agent can afford to read — each statement linking to the source documents it came from.

When a source changes, Brain rebuilds the affected summary — nobody has to remember to update it. The example below is a real one, distilled from our own August incident.

Captured
session-capture.mddeploy-timeline.txtadr-0021-api-tokens.mdmcp-connectivity-drops.mdlistener-log.mdhandoff-thread.md
distilled
What the agent reads
entity · search-mcp

Redeploys kill in-memory MCP sessions and hour-long JWTs — clients reconnect with static per-user API tokens instead. Decided after the August connectivity incidents.

adr-0021-api-tokens.mdmcp-connectivity-drops.md+2
In the product

Shared memory, skills and handoffs for AI agents

Knowledge and curation

  • Search by meaning — semantic retrieval over everything captured, with source tracking on every fact.
  • Cited summaries — per-entity distillations that link to their documents and rebuild when a source changes.
  • Context packs — the relevant summaries, documents and relationships for an entity, assembled within a token limit.
  • Obsidian plugin — search, publish and two-way sync the same knowledge base from your notes.

Skills and handoffs

  • One skills repository — versioned skill bundles served to every connected agent as MCP tools; update one, connected agents pick it up.
  • Handoffs with context — agents send, reply and watch an inbox; work moves between agents with its context attached.
  • Delivered on reconnect — handoffs that arrive while a watcher is down are delivered on its next watch.

Models and privacy

  • Private lane — private mode runs every model call on Levirge-hosted LLMs; no cloud model provider in the path.
  • Cloud when you want speed — switch to cloud LLMs from Settings at runtime; no redeploy.
  • Local embeddings — vector generation and reranking run on Levirge infrastructure in both modes.

Access and operations

  • Tenants and Vaults — knowledge is partitioned per Vault; personal Vaults are private to their owner and fail closed.
  • Per-user tokens — hosted sign-in or bearer tokens you issue and revoke; only the hash is stored.
  • Ingestion you can see — a board of pending, errored and promoted items, with one-click retry and a health endpoint.
Where it runs

Private by default, fast when you need it

Two model lanes, one runtime switch. Private mode runs on LLMs Levirge hosts on its own infrastructure — no cloud model provider in the path. Hosted mode uses cloud LLMs when speed matters more than locality — and workspace data is tenant-isolated in both lanes. Self-hosting the whole system is a roadmap item, not today's offer.

Private by default

In private mode, every model call — categorisation, summarisation, chat — runs on Levirge-hosted LLMs; no cloud model provider in the path. Flip to cloud LLMs when speed matters more than locality — a runtime setting, not a redeploy. Embeddings and reranking run on Levirge infrastructure in both modes.

We host it

Sign in and connect your clients — there's nothing to operate. Pricing is quoted per workspace and set up on the access call. Self-hosted deployment is a roadmap item for enterprise; if you need Brain to run somewhere specific, that's a conversation worth having with us.

Private doesn't mean slow

Private inference keeps pace with ingest — categorisation, summarisation and retrieval run comfortably on a local LLM.

How it works

Every agent reads and writes one brain

Agents across every repo and client share the same memory and the same skills — contributing what they learn and drawing on what the others left behind.

Agents Levirge products Web UI users Claude Claude Desktop Codex Pi Any MCP client CI workers long-poll · CI read · write · watch Levirge Brain Knowledge Skills Handoffs Curation one workspace · agents & users share it shared context what they learn Levirge Search ✓ LIVE Your own services ✓ MCP Levirge DevOps SOON Levirge Teams SOON

One brain, one protocol — agents and users contribute to and draw on the same memory, whatever client it runs in.

Install

Pick your agent

Access is per workspace — request access, then connect your clients.

Early access — capacity is limited and workspaces are set up in order of request. Installing now is fine: the plugin connects the moment your workspace is active.

claude plugin marketplace add levirge/brain
claude plugin install brain@brain

Installs the /brain:* commands and connects the MCP server — sign in with OAuth on first use; workspace tokens remain as a fallback.

for agents MCP · Streamable HTTP
discovery
levirge.com/llms.txt
docs
levirge.com/brain.md
endpoint
https://brain.levirge.com/mcp
definition
levirge.com/brain/mcp.json
auth
OAuth sign-in · bearer token as fallback
verify
call kb_overview — read-only
tools
18 — full reference in brain.md
Client config — OAuth is negotiated on connect; add an Authorization header only for token fallback.
Next step

Give your whole fleet one memory

One server, every team member. Shared knowledge that curates itself across agents and users, skills you edit once, and handoffs that carry their context.