Radar / AI infrastructure / claude-mem
claude-mem
Captures tool-use observations across agent sessions, compresses them with AI into SQLite + Chroma hybrid search, and injects relevant context into future sessions automatically.
Why it matters
Works across Claude Code, OpenCode, Cursor, Grok and OpenClaw via lifecycle hooks plus 4 MCP search tools with ~10x token savings via progressive disclosure.
What you could build with it
A consulting team could install claude-mem on its shared coding setup so every session recalls prior decisions, bugfixes, and architecture choices, turning each engagement's history into searchable memory instead of re-explaining context to the agent every week.
Does it hold up?
Mixed evidence: 95.9k stars and real third-party integrations (e.g., octorato's skill) show broad production use with reported ~75% token savings, but a documented security report flagged self-healing hooks and context injection, and users reported silent worker failures dropping memory logging. Powerful, but configure it carefully.
Built with claude-mem
- octorato claude-mem-persistent-memory SKILL.mdgithub · Third-party agent framework documenting claude-mem as a brain-multiplier skill, claiming ~75% token savings per session with a risk-aware rollout plan and AGPL licensing caveat.
- Security report: self-healing hooks, context injection (v12.1.6)github · Responsible disclosure documenting a watchdog that silently re-enabled the plugin after disable, context injection every turn, and a possible poisoned-memory prompt-injection payload.
- Claude-Mem hackathon cheat sheetgithub · One-page maintainer doc explaining the two-agent architecture: an observer agent writes structured notes while the builder works, claiming ~99% startup-context savings.
Learn more
First spotted on github: source.
More AI infrastructure
Modular open-sources MAX inference server, Mojo stdlib and accelerator kernelsThe core of Modular's unified AI deployment stack is now open: the MAX inference server (OpenAI-compatible endpoints)…infra · JEV 0.76OmniRouteFree MIT AI gateway: one endpoint, 290+ providers and 500+ models with quota-aware auto-fallback.infra · JEV 0.75Context Mode: context-window optimization for AI coding agentsMCP server for AI coding agents that sandboxes verbose tool output and persists session memory, cutting context use by…infra · JEV 0.74StrataOpen-source inference engine that runs the 125B-parameter Qwen3.8-Flash-Next MoE on a single consumer GPU with 12GB+…infra · JEV 0.74Microsoft releases 301,000 Copilot coding-agent tracesMicrosoft open-sourced 301,026 GitHub Copilot coding-agent sessions (9.3M model calls, 8.7M tool calls) with timings…infra · JEV 0.71DeepSeek open-sources Ascend infrastructure stackDeepSeek published Ascend-optimized versions of its NVIDIA-proven infra components — TileLang Ascend, DeepGEMM-Ascend…infra · JEV 0.7
Get the week's best AI launches, plus 3 ideas worth building
One email every Saturday. Ranked by traction, not hype. Free.