.md file to compare - side-by-side diff against memory-consolidation
memory-consolidation
description: "Triggers on prompt mention of 'memory-consolidation'."
What it does for you
Tidies your assistant's memory by merging and pruning what it knows.
What it produces
A recent result, so you can see the kind of work it returns.
loading…
How to get it
These run inside the Snappy workspace. Want this working in your business? I set skills like this up with you, in one focused week.
For developers how this skill is built, graded, and how it runs
at a glance- the short version
what's inside - the parts that make up a skill 2/4 present
A skill is just a few plain-text files. Only the main one is required. The rest are optional, added as the work needs them. This is what the skill is made of; how it runs is just below.
state/skills/memory-consolidation/SKILL.md
present
state/lib/memory-consolidation.ts
not present
state/bin/memory-consolidation/
not present
state/skills/memory-consolidation/AGENTS.md
present
how it's graded - what counts as a good run 4 criteria · 1 deterministic · 3 judge
Each row is one thing a good run has to get right. deterministic means a quick check decides, pass or fail. judge means the AI reads the result and rates it. Grading each piece on its own (instead of one overall score) shows exactly where a run fell short, so the fix is obvious.
how it runs - the shared frame every skill uses 2/5 present
Every skill runs the same way. One part does the work, a separate part checks it, and a short loader hands the AI exactly what it needs for the job. Anything this skill doesn't use shows a one-line note saying why, on purpose, not by accident.
No separate check found. Without one, the part that makes the work could end up approving its own work, worth a closer look.
This skill doesn't fix its own gaps yet.
state/log/evals.ndjson - ALWAYS audit memories before claiming pass - score is 1.0 only if stale entries were actually removed
- A clean audit with no changes needed is 0.5, not 1.0 - the score asks "did the loop produce delta" not "did it run"
- Cron-scheduled (daily 6 AM); dispatcher auto-runs state/bin/<loader_slug>/run.ts when present
- Manual drift check: run npx tsx state/bin/memory-consolidation/run.ts from repo root
what it has learned - fixes written back in over time sample
When a run hits something this skill didn't handle, the fix gets written back into the skill so it doesn't happen again. FIXED means it was corrected on the spot. LOGGED means it's queued for a bigger rewrite. Either way, the skill gets a little better and never makes the same mistake twice.
- Loading feedback rows…
how the work flows- step by step
SKILL.md- the skill, written out in plain English
memory-consolidation
Backed by: state/bin/memory-consolidation/run.ts + state/lib/log.ts.
Cron job (daily 6 AM). The scheduled-agent dispatcher auto-runs state/bin/<loader_slug>/run.ts when that sidecar exists, so the memory-consolidation agent executes state/bin/memory-consolidation/run.ts. The sidecar audits memory markdown stores and appends state/log/memory-consolidation.ndjson on every tick so the heartbeat artery is concrete.
Steps
- Run
npx tsx state/bin/memory-consolidation/run.tsfrom the snappy-os repo root. - The sidecar reads
.mdfiles understate/log/memory,state/memory,
memory, and state/observations.
- It counts exact duplicate candidates without rewriting files.
- It appends one row to
state/log/memory-consolidation.ndjsonwith
{ ok, files_audited, entries_processed, deduped, pruned, merged }.
- Only perform destructive pruning/merging in a separate reviewed change after
proving the stale or duplicate entry is not still load-bearing.
Eval
Score 1.0 if memories were audited and a reviewed prune/merge changed files. Score 0.5 if the sidecar audited and appended the artery row with no changes. Score 0.0 if the sidecar could not read the memory roots or append the row.
Rubric
criteria:
- name: log_entry_exists
kind: deterministic
check: "A log entry for 'memory-consolidation' exists in 'state/lib/log.ts' with fields {ok, deduped, pruned, merged}."
- name: stale_entries_removed
kind: judge
check: "Auditor verifies that stale memory entries (referencing non-existent topics or files) have been removed from the '.md' files in the memory directory."
- name: duplicates_removed
kind: judge
check: "Auditor verifies that duplicate memory entries have been removed, resulting in unique content."
- name: related_memories_merged
kind: judge
check: "Auditor verifies that related memory entries covering the same topic have been merged into a single, consolidated entry."AGENTS.md- what the AI loads when this skill comes up
memory-consolidation - loader
Per-turn rules for the memory-consolidation skill. Full reference: state/skills/memory-consolidation/SKILL.md. Do not skip these.
Critical Rules
- ALWAYS audit memories before claiming pass - score is 1.0 only if stale entries were actually removed
- A clean audit with no changes needed is 0.5, not 1.0 - the score asks "did the loop produce delta" not "did it run"
- Cron-scheduled (daily 6 AM); dispatcher auto-runs
state/bin/<loader_slug>/run.tswhen present - Manual drift check: run
npx tsx state/bin/memory-consolidation/run.tsfrom repo root
Commands
| ui model | live composition via compose_inline, persisted as artifact lang_body, reopened with OpenArtifact | |invoke: run npx tsx state/bin/memory-consolidation/run.ts |backed by: Node fs primitives + state/lib/log.ts append helpers |eval log: state/log/evals.ndjson (skill: "memory-consolidation") |artery log: state/log/memory-consolidation.ndjson
Known Pitfalls
- "Stale" means references to things that no longer exist - verify before pruning, deletion is destructive
- Score 0.0 only if the sidecar could not read memory roots or append the artery row; a successful read with no changes is 0.5
Self-Test
An agent reading this should correctly:
- [ ] Score 0.5 (not 1.0) when audit completes with zero changes needed?
- [ ] Verify a "stale" reference no longer exists before pruning?
- [ ] Append the run to
state/log/memory-consolidation.ndjsonviastate/bin/memory-consolidation/run.ts?
Found a gap? Edit this file. <!-- footer-injection-point -->
api.ts- the code it can call
⚠ no api.ts - this skill has no typed action surface
scripts- helper scripts it can run
prose-only skill - 1 inline code block live in SKILL.md above (no state/bin/ sidecar yet).
how we check it- the checks, plus the last 10 runs
no recent runs logged - the eval contract is declared but nothing has been graded yet