embeddings
91 servers · 8,367★ total
Fast and Accurate Code Search for Agents. Uses 99% fewer tokens than grep+read
The memory your AI should have had from the start. Automatic capture, automatic recall, 100% local. One SQLite file, zero cloud. Works with Claude Code, Claude CLI, Cursor, Codex CLI, Gemini CLI.
Local-first 3-in-1 AI memory layer & MCP server for Claude Code, Codex, Grok, Gemini, VS Code and Cursor. Fuses session history, codebase indexing & concept graphs in SQLite. Enables zero-cloud, privacy-first context & instant recall, supports multi-agent swarms.
Pharos — local-first agentic RAG for your team's document library: multi-format ingest, hybrid retrieval, enterprise ACL, dual HTTP + MCP exits.
On-device memory layer for AI agents. Claude Code, Hermes and OpenClaw. Hooks + MCP server + hybrid RAG search.
AI semantic search for Zotero, with a built-in MCP server for AI agents (Claude Code, Codex). Find papers by meaning. 100% local and private.
🧠 High-performance persistent memory system for Model Context Protocol (MCP) powered by libSQL. Features vector search, semantic knowledge storage, and efficient relationship management - perfect for AI agents and knowledge graph applications.
CLI + Local MCP - A shared structured memory store across Claude Code, Cursor, Windsurf, Antigravity, and every MCP client. Semantically queryable.
Local memory infrastructure for AI agents. Store knowledge and skills in isolated vaults you compose, control and query.
Give Claude semantic memory of your Obsidian vault — local semantic search over Smart Connections embeddings via MCP. Multi-vault, block-level, 100% private.
Persistent memory graph for AI agents. Facts, decisions, entities, and relationships that survive across sessions, tools, and providers. MCP server — works with Claude, Cursor, ChatGPT, and any MCP client.
Local-first context ingestion and retrieval for AI tools. SQLite + embeddings + MCP server for Cursor & Claude.
MCP server for semantic search using local Qdrant vector database and OpenAI embeddings
Self-hosted AI agent memory server with MCP, evidence provenance, typed claims, conflict detection, embeddings, recall, PostgreSQL, and pgvector
Generic markdown collection MCP server with FTS5 + semantic search, frontmatter-aware indexing, and incremental reindexing
Persistent memory system for AI coding assistants. Captures decisions, learnings, and context from coding sessions. Features hybrid search (semantic + BM25), MCP server integration, SQLite persistence with knowledge graph, and proactive memory surfacing. Written in Rust.
Self-hosted AI knowledge base with hybrid semantic search (pgvector + FTS + RRF), MCP server, multi-provider LLM inference (Ollama, OpenAI, OpenRouter, llama.cpp), multimodal ingestion (vision, audio transcription, speaker diarization), and knowledge graph. Rust + PostgreSQL.
A local-first, LLM-agnostic memory layer for AI assistants
Durable, local-first memory for AI coding agents over MCP — zero-dependency, curated & semantically de-duped, you own the data (SQLite + Markdown). Works with Claude Code, Codex & any MCP host.
Local-first MCP server giving AI coding agents fast, structured, and semantic context over any codebase. Zero config, zero cloud, full context.
AI-first, zero-dependency JavaScript database. Vector search, agent memory, MCP server, and encryption built in. Node.js, Bun, Deno, browsers, and edge runtimes.
Enhanced MCP server for semantic code search with call-graph proximity, recency ranking, and find-similar-code. Built for AI coding assistants.
Personal knowledge graph MCP server on Cloudflare D1 + Vectorize. Deploy your own sovereign second brain in one click — Claude captures atomic concepts, finds cross-domain analogies, and links them with substantive justifications. Latticework thinking as a service, running entirely in your Cloudflare account.
MoFlo — an opinionated, local-first AI agent orchestration toolkit for Claude Code: semantic memory, learned routing, gates, and spells. No API keys, no cloud, works out of the box.
🔒 Privacy-first MCP server for Claude using PostgreSQL + Ollama. Local alternative to cloud-based code context with full data sovereignty. No API keys, no external calls, 100% local.
Nautobot Model Context Protocol (MCP) Server - Contains STDIO and HTTP Deployments with Embedding Search and RAG.
Structured local memory storage and retrieval for LLM agents
🧠 Personal knowledge MCP server with vector database for Opencode. Store and retrieve knowledge using semantic search, powered by local embeddings.
Open-source, self-hosted knowledge backend for AI agents — hybrid search (vector + keyword), MCP server, 5 connectors, Docker-ready
Persistent semantic memory for AI agents on Raspberry Pi 5 — local Qdrant + MCP, no cloud, ~3s per query
Your coding agent starts every session with amnesia — memo fixes that, 100% on your machine. Persistent memory for Claude Code, Codex, Cursor & any MCP client: Markdown source of truth, hybrid search (MLX/CPU + sqlite-vec), time-machine, contradiction radar, nightly self-optimization. No cloud, no keys.
MCP server in Rust for AI agent persistent memory: branch-aware session handoffs, local ONNX embeddings, SQLite-backed semantic search.
Local-first, eval-first memory for long-horizon AI agents — no LLM at ingest. Python SDK + MCP server with source-traceable recall, belief revision, selective forgetting, and reproducible benchmarks.
Cross-session memory and recall for AI agents — git-synced knowledge base, hybrid semantic+TF-IDF search, auto-distillation with secrets scrubbing
Lightning-fast RAG for AI agents. ONNX-powered, 4-layer fusion, MCP server. No PyTorch.
Persistent semantic memory server for MCP - Give your AI long-term memory that survives across conversations. Lightweight Python server with SQLite storage and semantic search.
Persistent visual memory for AI agents — capture screenshots, embed with CLIP ViT-B/32, compare, recall. MCP server + Rust core library.
MCP server for BookStack — 56 tools covering the full API + semantic vector search. Rust/tokio/axum, dual transport (SSE + Streamable HTTP), OAuth 2.1, pluggable DB (SQLite/PostgreSQL+pgvector).
Lightweight RAG server for the Model Context Protocol: ingest source code, docs, build a vector index, and expose search/citations to LLMs via MCP tools.
Self-hosted semantic memory over MCP for an engineering team's tacit knowledge. Rust, Postgres + pgvector.
PeopleMesh is the AI-powered matching layer for modern organizations. It helps people discover the right colleagues, internal opportunities, communities, and projects through semantic search that understands context, not just keywords.
Standalone Node MCP server: semantic search + knowledge graph + vault editing for Obsidian, no plugin required
Local-first crystalline intelligence for AI agents: the knowledge that endures across sessions, taught as Markdown Domains and captured as Engrams. One Rust binary with MCP server, CLI and hybrid search.
A memory graph designed for the agent that uses it, not the human who feeds it. Engrama reconstructs context from associations on demand, replacing the "stuff everything into the prompt" reflex with targeted graph traversal. SQLite default, Neo4j optional.
Persistent cross-session memory for AI coding assistants. Works with Claude Code, Cursor, Windsurf, Cline, OpenClaude, and MCP-compatible editors.
A local, privacy-preserving knowledge-base MCP server with semantic search (RAG) for Claude and other AI assistants.
The continuity layer for everything you do with AI. A self-hosted, open source server any tool plugs into over MCP or REST: hybrid vector + lexical recall, an auto-built knowledge graph, sleep-style consolidation, procedural and persona tiers, OAuth 2.0, multi-tenant. SQLite or Postgres. MIT.
Semantic routing for MCP tools - .NET library that indexes MCP tool definitions and returns the most relevant tools via vector search
RAG for researchers: page-level citations from your personal library, LLM access via MCP. Ask your entire archive, get answers with the intelligence of leading AI.
Local-first semantic knowledge graph with magnetic-pull retrieval
A Claude Code plugin that gives Claude fully automatic, per-project cognitive memory with hybrid search, session lifecycle hooks, and local embeddings.
MCP server for AI memory -- hybrid search (BM25 + semantic + knowledge graph), temporal decay, local-first
🧠 Persistent memory for AI agents. SQLite for agent state. Zero cloud dependencies. Local embeddings. MCP-native integration with Claude Desktop/Code, Cursor, Windsurf & more.
Official TypeScript/JavaScript SDK for Cognipeer Console — OpenAI-compatible chat, batch, realtime voice, embeddings, RAG, MCP, agent tracing, and guardrails for multi-tenant AI products.
Persistent memory for Claude Code. Automatically indexes every conversation and provides production-grade hybrid search (BM25 + vectors + reranker) via MCP tools. 100% local, zero config, zero API keys.
"From Prompt to Cognitive Engineering". — AI: Designed, not Dreamed.
Brain-like persistent memory for AI agents. Hybrid search, knowledge graph, sleep-consolidation lifecycle, expert routing (HMoE), and a graph-Laplacian diffusion subsystem driving decay, consolidation, auto-link, and spectral retrieval. 65 MCP tools. Local, no cloud.
MCP server providing long-term memory for AI agents — knowledge graph with Qwen3 embeddings, continuous learning, and knowledge evolution
Local-first semantic search for your ChatGPT, Claude, and AI conversation history. Python CLI + MCP server for Claude Desktop. Uses sentence-transformers + ChromaDB. No API keys, nothing leaves your machine. Part of the Resonant ecosystem.
Knowledge engine for AI agents — persistent memory, vault, and brain that learns
Fast hybrid (BM25 + semantic) local code search for AI agents - pure Rust, persistent index, MCP/gRPC servers, tree-sitter symbols
Local AI memory system for Node.js — verbatim storage, semantic search, knowledge graph, MCP server. Port of the Python MemPalace.
MCP server for agent memory over HexxlaDB—ring retrieval, embeddings + lexical search, seams, facets, YAML persistence policy, localhost HTTP transport.
International Education Standards MCP — CCSS, Singapore MOE, IB MYP and more. Grounding layer for LLM-powered education tools.
Long-term memory your AI coding agents actually share. One brain, many agents — standalone embedded MCP server: SQLite + FTS5 + vector hybrid search, local ONNX embeddings, self-expiring diary. No server, container, or daemon.
MCP-native embedded memory database for AI agents built in Rust. REMEMBER/RECALL/FORGET/SHARE primitives with hybrid vector search, AES-256-GCM encryption, DuckDB/PostgreSQL backends & SDKs for Python, TypeScript and Go.
🧠 내 AI에게 '나'를 기억시키는 로컬 second-brain — 결정의 선택·이유·전제를 .md 노트로 남기고, MCP로 어떤 AI 앱(Claude Code·Desktop·Cursor)에서든 꺼내 쓰는 개인 기억 계층. 전부 로컬(Ollama 임베딩)·오픈소스.
Self-hosted RAG stack for crawl, scrape, search, ingest, query, and ask workflows with Qdrant, TEI embeddings, Chrome rendering, MCP/CLI/REST, and Gemini synthesis.
Docker container for centralized Ruflo MCP server with PostgreSQL (RuVector). Architecture: Ruflo (stdio) → Express proxy (Streamable HTTP /mcp).
Searchable MCP and AITool gateway for .NET, built on Microsoft.Extensions.AI with embedding-based tool discovery, lexical fallback, and unified execution
Local-first RAG engine with cited answers from your own docs. GraphRAG, sealed sharing, MCP-native, VS Code Copilot integration. Nothing leaves your machine.
Open-source context layer for LLM agents — a single Rust binary served over MCP that feeds Claude, ChatGPT, Codex or any MCP client the facts, decisions & history they need. Local-first, on-device embeddings, hybrid BM25 + vector + graph retrieval. Apache-2.0.
MCP tool to enable semantic code search in your project.
Self-hosted documentation RAG pipeline that serves semantic search to LLM agents over the Model Context Protocol.
Fast, local semantic search over web content for AI agents. Hybrid BM25 + potion-retrieval-32M embeddings, cross-page dedup, token-budget mode, MCP server, SearXNG bridge. ~90% fewer tokens than raw web_fetch.
Persistent, local-first memory for AI coding agents. Give Claude Code and Cursor memory across sessions over MCP — and see and edit exactly what your AI remembers. Open source, no API key, no cloud.
RE-call — Retrieval-Augmented Self-Recall: RAG over an AI agent's own memory that knows when it doesn't know (gap detection, freshness, anti-re-litigation). PostgreSQL + pgvector, hybrid retrieval + RRF.
Portable hybrid (BM25 + dense + RRF) retrieval engine and a label-free evaluation harness — extracted from a personal AI-assistant memory index and decoupled to run on any source tree.
Pre-release PostgreSQL-backed MCP memory server with hybrid search and hierarchical summaries
Local-first temporal knowledge graph memory for AI agents. Runs with Ollama — no API keys required. Single Go binary + SQLite. Hybrid search, MCP server, REST API, interactive TUI.
High-performance, pure Go memory middleware for AI Agents (Claude/Cursor). Features active file watching, auto-distillation via LLM, and 'Single Source of Truth' versioning. Native MCP support.
Local-first multilingual memory for Codex, Claude Code & MCP — auditable Qdrant recall, project isolation, and a single-instance desktop tray.
Agent memory system with retain-recall-reflect loop. Hybrid search, entity resolution, observation synthesis. TypeScript library + MCP server.
Persistent layered memory MCP server for AI agents — SQLite + FTS5 + vector hybrid search (RRF), multilingual, zero-cloud
Self-hosted personal AI assistant on Telegram. Powered by OpenCode + Claude Sonnet (free big-pickle). OpenClaw skills compatible. SQLite + MCP live memory with optional Ollama vector search.
MCP server for multilingual dictionary lookups with word relations (synonyms, antonyms, hypernyms, translations, etc.) covering all languages via ConceptNet, Wiktionary, and Datamuse
Perfect recall for imperfect machines. A shared memory layer for LLMs via MCP
Semantic code search for your repo, as a CLI and an MCP server. Bring any OpenAI-compatible embedding model. Zero dependencies.
MCP-native long-term memory engine — identity / event / ephemeral layers on Postgres + pgvector.