It arrives on its own
The memory is in context before reasoning starts. There is no tool for the agent to call, no moment where it must notice something is missing, and nothing for it to forget.
GRM reads the conversations and code already on your machine and puts the relevant piece into your agent's context before it starts thinking — no tool call, nothing for it to remember to do. It runs entirely on your computer and makes no network calls of any kind.
7 days free, counted from your first run — not from the day you downloaded it. Then $99 per machine per year, paid once.
Real output, captured from a live install — a decision from an earlier session
and the source that implements it, delivered without the agent asking for anything.
(The [2026-08-14] timestamp is recorded in UTC, which is why the conversation text inside reads 15 Aug 2026 after midnight in Indian Standard Time.)
Two things decide whether a memory is worth having: whether it turns up unasked, and what it costs to keep running.
Nothing to call, nothing to configure per project, and no bill that grows with how much you use it.
The memory is in context before reasoning starts. There is no tool for the agent to call, no moment where it must notice something is missing, and nothing for it to forget.
ok do it gives a search nothing to search on. At the end of the previous turn the agent already said what it was about to do, and the answer was fetched then and set aside.
The precise filename, function name, error string or compiler flag you asked about. No near-misses, and no invented detail presented as fact.
Across 715 recorded decisions it said nothing 297 times rather than crowd your context with a guess. Silence is what makes the rest of it trustworthy.
The ranking is arithmetic over words you already wrote. That is why there is nothing to meter, nothing to rate-limit, and no per-query bill — ever.
One document on your disk. Read it, copy it, back it up, delete it. Uninstalling takes all of it with you and leaves nothing behind.
One recording, one screenshot, both off this machine. Nothing staged, nothing sped up, no edits.
The memory reaches an agent two ways. On the left it is asked and answers. On the right nobody asked — it was already there.
Memory [recall]
fires and dated blocks come back. Then the agent ends its own answer with
[[next: …]], writing down what it expects to need. On the
following turn the answer arrives marked
“pulled from the pre-fetched context, no recall call needed.”
46 seconds.
Neither of these was re-shot for this page. Both were typed at working speed, typos and all — the model reads intent, not spelling.
Every decision the engine made was logged as it happened, then scored afterwards against what was actually written next.
Not a benchmark. Two months of one developer's real history, searched by the automatic half of the memory — the part that speaks without being asked.
| Words of conversation and source | 2,005,699 |
| Characters, as 24,625 memories | 13,694,306 |
| Distinct words | 32,323 |
| Blocks of source code | 2,404 |
| Days without a gap | 66 |
| Recorded decisions | 715 |
| Times it spoke | 418 |
| Times it stayed silent | 297 |
How the 56.7% is counted. An engine cannot be asked whether it helped — hand any model an answer that agrees with what it already thinks and it will say yes. So it is not asked. Every volunteered memory is compared against what was written afterwards, and counts only if material from it appears there. Anything the memory quoted back to itself is excluded.
It averages two very different days. One scored 74.6%, the other 34.3%. The high day moved between unrelated pieces of work; the low day was one long unbroken session on a single subject, where the answer to almost everything was already a few screens up. On that day the share of memory genuinely new to the session fell with it, 27.6% to 20.0% — there was simply less left to tell.
One way the measure is unfair to itself. A memory only scores when its words turn up in what gets written next, so hours spent writing code and shell commands score low however useful it was. The real figure is better than the one shown. It is reported this way on purpose.
None of it comes from the search tool. This scores only what arrived unasked. Memory the agent deliberately went looking for is not counted.
No account, no sign-up, no API key. The trial licence is inside the download.
One binary, three commands, then quit your agent and open it again. That is the whole procedure on every platform.
The binary, your 7-day key and a README — about half a megabyte. macOS is built for Apple Silicon; Linux is statically linked and runs on any x86-64 distribution. curl and tar already ship with Windows 10 and 11, macOS and every Linux, so there is nothing to install first — and nothing to click past.
It finds the agents already on this machine, wires itself into each of them, reads the history you already have, and starts the background half.
Your conversations are picked up automatically — every Claude Code, Cursor and Antigravity session already on the machine is read during the install. Your code is not, until you say which folders. The installer leaves an empty code_roots.txt beside the memory file; put one project path on each line. Keep it short — a handful of real projects beats every folder on the disk.
Closing the window is not enough — an agent reads its configuration at launch, so a session already open has not seen any of this. After that it is on, and there is nothing else to run. pm doctor checks every part of it and names the fix for anything that is down.
One command checks every part and names the exact fix for anything it finds. It is also the first thing to send us if you get stuck. Real output, captured from a live install.
You pay for the computers you run it on, once a year. Nothing renews on its own.
No tiers, no seats to administer, no per-token billing.
India: ₹2,999 per machine per year, payable in rupees. Same product, same key — ask for the rupee link.
pm id and mail what it prints to support@getgrm.tech — a 12-character fingerprint of the computer, derived locally and never transmitted by the product.pm_license.key and restart your agent. Your trial keeps running while you wait, so nothing stops.If something here is not answered, mail us — the answer comes from the person who wrote the code.
No. There is no cloud service, no embedding API, no vector database and no telemetry. The binary makes no network calls of any kind. Your memory is one file on your disk.
Sentences you and your agent already wrote, and blocks of code from folders you point it at, each with a date. Nothing is summarised by a model and nothing is sent anywhere to be processed.
The memory goes quiet and your agent carries on. The hook gives up within a quarter of a second rather than making you wait, restarts the background half, and tells the agent plainly that memory was unavailable — so a fault is never mistaken for you having no history.
About 1 MB, one file. No Python, no Docker, no database, no runtime to install — it imports nothing but the operating system's own libraries. Windows and Linux on x86-64, macOS on Apple Silicon.
Immediately. It reads the history already on your machine during install, so it knows your past work from the first question — you do not have to use it for a month first.
Delete the folder and the .pm directory in your home folder. That is all of it. Nothing else is written anywhere that matters.