Observed Signal · Apr 25, 2026 · Technical Release · Source: DEV Community · Impact: 2/5 · Sentiment: Positive

claude-recall lets AI agents read session archives

Executive Signal Summary

A developer built claude-recall, a beta tool that lets Claude Code instances (and similar agentic tools) read and reuse their locally persisted session history. The tool indexes Claude Code's per-turn JSONL archive into a local SQLite database with FTS5 full‑text search, offers an optional semantic reranking layer using a local Ollama embedding model (nomic-embed-text), and injects ranked past-session matches into prompts via a UserPromptSubmit hook. The hook is distributed as a NativeAOT-compiled binary to avoid Python startup latency. The package (v0.5.3) is on PyPI and the open repo is on GitHub under an MIT license. The author frames this work as part of a broader pattern—agents already write structured histories to disk but rarely read them—and contrasts claude-recall’s read-only approach with other memory strategies like curated remember/forget systems and task-tracking stores.

Polaris7 AgentPolaris7 Strategic Assessment
High Confidence

Practical developer tooling that improves agent reliability by enabling local recall of persisted session history; relevant to teams building agentic workflows but not a major platform policy or industry-shifting release.

SIGNAL RADAR

Track Ollama Signals & Market Shifts in Real-Time

Polaris7 autonomous intelligence agents track regulatory filings, primary sources, executive changes, and deal flow 24/7. Create your free Explorer workspace to monitor these entities.

Start Free in Explorer
Free Explorer tierNo credit card requiredInstant watchlist setup

Key Takeaways & Evidence Grounding

  • claude-recall indexes Claude Code JSONL session archives into SQLite using FTS5 full-text search.
  • It offers optional semantic reranking via a local Ollama embedding model called nomic-embed-text.
  • A UserPromptSubmit hook injects ranked prior-session matches as additionalContext into Claude Code prompts; the hook is shipped as a NativeAOT-compiled binary to reduce latency.
  • claude-recall is available on GitHub (github.com/LearnedGeek/claude-recall), published to PyPI as 'claude-recall', MIT licensed and tagged beta (v0.5.3).
Primary Source Grounding & Direct Attribution
Direct Origin Attribution
Primary Reporting: DEV Community•Published: Apr 25, 2026
Original Coverage Title: “Your AI agent already writes every session to disk. Why isn't it reading its own archive?”

Related Market Signals & Shifts

Recent verified developments and strategic activity across this market segment.

Large Language Models (LLM) & AIApr 18, 2026

Lossless AI Memory Tool for Claude Code

Longhand is an open-source tool that captures Claude Code session logs verbatim, indexes them locally, and exposes deterministic recall without LLM summarization. Instead of relying on larger context windows, Longhand ingests JSONL session files and stores structured events in SQLite and semantic vectors in ChromaDB. It runs an auto-ingest hook on session end, supports one-time backfill, and exposes recall via an MCP server with ~17 Claude tools (recall, search_in_context, get_session_timeline, replay_file, etc.). The project is published on PyPI (longhand), registered in the MCP Registry, MIT-licensed, Python 3.10+, and tested against 107 sessions (53,668 events). Reported characteristics: ~126ms semantic recall, ~200–400MB typical storage (up to ~1GB heavy users), zero network/API calls, offline operation, and 170 unit tests with security audit results showing no critical findings.

Read assessment
Large Language Models (LLM) & AIAug 22, 2026

Anthropic ships memory into Claude Code

A developer describes building an MCP-based persistent memory layer for AI coding assistants called cachly and how Anthropic's announcement that Claude Code gained a memory feature prompted re-evaluation of that work. The author explains a reproducible test for measuring whether an assistant truly persists facts between sessions, shares benchmark results from a personal corpus, and summarizes community feedback that changed the project's metrics and reliability checks. The post notes cachly offers a free tier hosted in the EU and links to the project's site.

Read assessment
Large Language Models (LLM) & AIMay 3, 2026

CTX Adds Persistent Memory to Claude Code

Jaewon Jang published an article describing CTX, an open-source tool that provides persistent, local memory for Claude Code. CTX hooks into Claude Code’s UserPromptSubmit event and injects relevant context before each prompt via three subsystems: Decision memory (parses git history), Code and doc search (BM25 search across repo and docs), and a Chat Memory vault (local SQLite of past conversations with hybrid BM25/vector search). The project is installable via pip or as a Claude Code plugin, keeps data local (no cloud sync or LLM calls), and provides benchmark and telemetry results showing substantial recall and utility gains. Source code is on GitHub and the package is available on PyPI; the post includes demo video and install validation details.

Read assessment

Track Real-Time Market Signals & Shifts

Set up custom watchlists to receive automated, evidence-grounded executive digests whenever material signals or shifts occur across your tracked landscape.