mnemo
pre-alpha · open source · Apache-2.0

MNEMO

A terminal coding agent with a memory that learns your projects. Several agents in one window. Bring your own model.

Mne, the pixel elephant: elephants never forget
curl -fsSL https://github.com/AtmanMishra/Mnemo/releases/latest/download/install.sh | sh
irm https://github.com/AtmanMishra/Mnemo/releases/latest/download/install.ps1 | iex

No Bun, Node or Rust needed. The installer checks the download's SHA-256 against the release's checksum file (that catches a corrupted download, not a compromised release: verify provenance if that matters to you), never touches your memory or settings, and upgrades in place when you run it again. Read the script first.

Then: mnemo doctor to see what it found, mnemo --demo for a scripted session that needs no API key, and mnemo to start. Remove it with … | sh -s -- --uninstall (Windows: set $env:MNEMO_UNINSTALL=1 first).

What it is

◈

It remembers

After each run Mnemo writes down what it learned about the project and about you, recalls it next session, and keeps the fix when something failed. Memory is local, and what it holds is yours to read (/memory) and delete (/forget).

◇

Memory is detachable

One command prints the hooks and MCP entry that attach the same memory to Claude Code or Codex: mnemo memory setup claude-code. What one agent learns, the others recall.

▣

Many agents, one window

A hub of every agent and what it is doing, split view for up to four, a project switcher, and a toast when one you are not watching finishes or needs you.

▞

Pixels, on purpose

Mne the elephant reacts to the work; diffs and test runs are drawn as pixels; the memory map shows what it holds. Three themes: Night, Game Boy, Paper.

The interface

The hub: a card per agent with its status, its last prompt and a pixel per turn
The hub: every agent at a glance. ctrl+g
Three agents side by side in a split view
Split view: up to four agents side by side. ctrl+s
The memory map: the project in the middle, its facts, pitfalls and sessions around it
The memory map: what it knows about this project.
Best-of-three as a race of lanes, the winner in green, then the exit card with Mne waving
/bestof: three attempts race, the smallest that passes the check wins.

What we measured

All on a small, cheap model (DeepSeek v4.1 Flash), through the same agent loop with memory on and off. Small samples: read the direction, not the decimals. The raw results and what went wrong along the way are in research/.

TestNo memoryWith memory
Six tasks in one repo, with rules the code does not reveal (3 runs)87%100%
Two repos whose rules conflict (2 runs)82%99%
Terminal-Bench subset: 10 easy/medium tasks, one attempt each10 / 1010 / 10

The last row is the honest one: unrelated one-off tasks in a fresh container give memory nothing to remember, and it makes no difference there (about $0.008 a task either way). It is a subset on our own machine, not comparable to the public leaderboard. The first two are where memory is supposed to pay off. Against a simpler, global memory in the style of Hermes Agent, Mnemo matched it at one or two repos and was ahead where it keeps projects apart (a command learned in one project was offered in another) and picks up unfinished work. The comparison is written up in research/hermes-comparison.md.

Before you run it

Mnemo is a coding agent: it reads your files, edits them and runs shell commands as you. By default it asks before every edit and command; yolo and headless -p ask for nothing, so use those in a container or a throwaway checkout. The permission gate filters what runs; it is not a sandbox.

Your prompts and the files the model reads go to the provider you chose, and nothing else leaves your machine: there is no telemetry and no update check (we measured it). Commands the agent runs inherit your environment, provider keys included, which a prompt injection could try to read; SECURITY.md says what to do about that today. It is pre-alpha software: expect rough edges and tell us about them.