Files

T

jp 942aae3ed5 fix(hnsw): address @igorls's #1173 review

Restore three `_pin_hnsw_threads` tests that the previous integrity-gate
commit deleted. The function is still live code on develop (defined at
chroma.py:207, called from chroma.py:705 + mcp_server.py), so the
deletion left the import unused (ruff F401) and dropped coverage on a
function unrelated to this PR's scope. Restored verbatim from main.

Plus three nits @igorls flagged:

- **Thread-safety doc**: `_quarantined_paths` mutation is lock-free;
  documented that idempotency of `quarantine_stale_hnsw` is the safety
  property (concurrent same-palace calls produce a benign redundant
  rename attempt that no-ops, no need for a lock).
- **Pickle protocol assumption**: `_segment_appears_healthy` requires
  PROTO ≥ 2 (`0x80`). Documented; matches what chromadb writes today,
  and a future protocol-0/1 emission would conservatively quarantine
  + lazy-rebuild rather than mis-classify as healthy.
- Class-level vs module-level scope: keeping class-level — the
  conftest reset is the controlled case, and module-level wouldn't
  remove the foot-gun, just relocate it. Conftest reset documented in
  the existing comment is the right pattern for test isolation.

Style nit (`Path(marker).touch()` vs `open(marker, "a").close()`)
deferred — that pattern lives in #1177's territory, not #1173's.

37/37 tests pass on the PR branch.

Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>

2026-04-26 13:20:14 -07:00

backends

fix(hnsw): address @igorls's #1173 review

2026-04-26 13:20:14 -07:00

i18n

fix(entity): reduce noise in regex-based detection

2026-04-24 00:20:32 -03:00

instructions

fix: add mempalace-mcp console entry point for pipx/uv compatibility

2026-04-21 01:26:00 -03:00

sources

fix(sources): address Copilot review on #1014

2026-04-18 17:17:50 -03:00

__init__.py

fix: upgrade chromadb to >=1.5.4 for python 3.13/3.14 compatibility

2026-04-18 12:05:46 -07:00

__main__.py

MemPalace: palace architecture, AAAK compression, knowledge graph

2026-04-04 18:16:04 -07:00

cli.py

fix(repair): refuse to overwrite when extraction looks truncated (#1208 )

2026-04-25 23:34:05 -07:00

closet_llm.py

release: v3.3.0 (#839 )

2026-04-13 18:25:01 -07:00

config.py

feat(graph): cross-wing tunnels by shared topics (#1180 )

2026-04-24 23:06:26 -03:00

convo_miner.py

test: cover embedding device fallback and bounded upserts

2026-04-24 23:06:50 +00:00

convo_scanner.py

fix(llm): tighter refinement — word boundaries, JSON extraction, authoritative sources

2026-04-24 01:30:40 -03:00

dedup.py

refactor: route all chromadb access through ChromaBackend

2026-04-14 00:31:16 -03:00

dialect.py

fix: address i18n review issues from PR #718

2026-04-15 11:03:28 +05:00

diary_ingest.py

release: v3.3.0 (#839 )

2026-04-13 18:25:01 -07:00

embedding.py

test: cover embedding device fallback and bounded upserts

2026-04-24 23:06:50 +00:00

entity_detector.py

feat(graph): cross-wing tunnels by shared topics (#1180 )

2026-04-24 23:06:26 -03:00

entity_registry.py

Merge pull request #931 from mvalentsev/fix/i18n-entity-metadata

2026-04-16 15:54:01 -03:00

exporter.py

fix: restrict file permissions on sensitive palace data (#814 )

2026-04-15 00:27:03 -07:00

fact_checker.py

release: v3.3.0 (#839 )

2026-04-13 18:25:01 -07:00

general_extractor.py

MemPalace: palace architecture, AAAK compression, knowledge graph

2026-04-04 18:16:04 -07:00

hooks_cli.py

fix: best-effort HNSW thread-pin retrofit + drop dead attempt-cap constant

2026-04-25 04:36:29 -03:00

instructions_cli.py

fix: add explicit UTF-8 encoding to read_text() calls (#776 )

2026-04-16 16:00:29 +05:00

knowledge_graph.py

fix(sources): address Copilot review on #1014

2026-04-18 17:17:50 -03:00

layers.py

fix: guard Layer3.search_raw against None doc/meta from ChromaDB (#1011 )

2026-04-18 13:30:57 -07:00

llm_client.py

fix(llm): tighter refinement — word boundaries, JSON extraction, authoritative sources

2026-04-24 01:30:40 -03:00

llm_refine.py

feat(graph): cross-wing tunnels by shared topics (#1180 )

2026-04-24 23:06:26 -03:00

mcp_server.py

fix: best-effort HNSW thread-pin retrofit + drop dead attempt-cap constant

2026-04-25 04:36:29 -03:00

migrate.py

Merge pull request #935 from shaun0927/fix/repair-crash-safety

2026-04-25 20:49:23 -03:00

miner.py

chore(rebase): reconcile with develop and apply ruff format

2026-04-25 04:39:31 -03:00

normalize.py

fix: add provenance header and speaker IDs to Slack transcript imports (#815 )

2026-04-15 00:27:01 -07:00

onboarding.py

test: add comprehensive test coverage (35% → 58%, threshold 50%)

2026-04-08 20:54:56 +03:00

palace_graph.py

Merge pull request #1168 from arnoldwender/fix/security-tunnels-permissions

2026-04-25 04:21:44 -03:00

palace.py

fix: Windows CI compat for palace lock tests and path normalization

2026-04-25 04:34:30 -03:00

project_scanner.py

feat(graph): cross-wing tunnels by shared topics (#1180 )

2026-04-24 23:06:26 -03:00

py.typed

chore: tighten chromadb version range and add py.typed marker

2026-04-07 18:51:42 -03:00

query_sanitizer.py

fix: address Copilot review comments on PR #739

2026-04-12 23:07:46 -03:00

README.md

MemPalace: palace architecture, AAAK compression, knowledge graph

2026-04-04 18:16:04 -07:00

repair.py

fix(repair): refuse to overwrite when extraction looks truncated (#1208 )

2026-04-25 23:34:05 -07:00

room_detector_local.py

fix: skip unreachable reparse points in detect_rooms_from_folders (#558 )

2026-04-11 16:16:06 -07:00

searcher.py

fix(search): BM25 hybrid rerank, legacy-metric warning, invariant tests

2026-04-25 00:39:37 -03:00

spellcheck.py

MemPalace: palace architecture, AAAK compression, knowledge graph

2026-04-04 18:16:04 -07:00

split_mega_files.py

Merge pull request #681 from jphein/fix/unicode-checkmark

2026-04-18 23:27:57 -07:00

sweeper.py

fix: address Copilot review on release/3.3.2

2026-04-19 18:19:28 -03:00

version.py

release: v3.3.3

2026-04-23 16:44:22 -07:00

README.md

mempalace/ — Core Package

The Python package that powers MemPalace. All modules, all logic.

Modules

Module	What it does
`cli.py`	CLI entry point — routes to mine, search, init, compress, wake-up
`config.py`	Configuration loading — `~/.mempalace/config.json`, env vars, defaults
`normalize.py`	Converts 5 chat formats (Claude Code JSONL, Claude.ai JSON, ChatGPT JSON, Slack JSON, plain text) to standard transcript format
`miner.py`	Project file ingest — scans directories, chunks by paragraph, stores to ChromaDB
`convo_miner.py`	Conversation ingest — chunks by exchange pair (Q+A), detects rooms from content
`searcher.py`	Semantic search via ChromaDB vectors — filters by wing/room, returns verbatim + scores
`layers.py`	4-layer memory stack: L0 (identity), L1 (critical facts), L2 (room recall), L3 (deep search)
`dialect.py`	AAAK compression — entity codes, emotion markers, 30x lossless ratio
`knowledge_graph.py`	Temporal entity-relationship graph — SQLite, time-filtered queries, fact invalidation
`palace_graph.py`	Room-based navigation graph — BFS traversal, tunnel detection across wings
`mcp_server.py`	MCP server — 19 tools, AAAK auto-teach, Palace Protocol, agent diary
`onboarding.py`	Guided first-run setup — asks about people/projects, generates AAAK bootstrap + wing config
`entity_registry.py`	Entity code registry — maps names to AAAK codes, handles ambiguous names
`entity_detector.py`	Auto-detect people and projects from file content
`general_extractor.py`	Classifies text into 5 memory types (decision, preference, milestone, problem, emotional)
`room_detector_local.py`	Maps folders to room names using 70+ patterns — no API
`spellcheck.py`	Name-aware spellcheck — won't "correct" proper nouns in your entity registry
`split_mega_files.py`	Splits concatenated transcript files into per-session files

Architecture

User → CLI → miner/convo_miner → ChromaDB (palace)
                                     ↕
                              knowledge_graph (SQLite)
                                     ↕
User → MCP Server → searcher → results
                  → kg_query → entity facts
                  → diary    → agent journal

The palace (ChromaDB) stores verbatim content. The knowledge graph (SQLite) stores structured relationships. The MCP server exposes both to any AI tool.