hybrid-search

44 servers · 936★ total

Pharos — local-first agentic RAG for your team's document library: multi-format ingest, hybrid retrieval, enterprise ACL, dual HTTP + MCP exits.

On-device memory layer for AI agents. Claude Code, Hermes and OpenClaw. Hooks + MCP server + hybrid RAG search.

Codebase Context gives AI agents understanding of your codebase through semantic code search, team conventions, patterns, and memory, so they use fewer tokens, spend less time, and produce better, more familiar output.

Persistent long-term memory for Claude Code via MCP — captures coding decisions, bugfixes, and context across sessions. Hybrid FTS5 + TF-IDF search with episode batching. Single SQLite DB, no external services. Alternative to claude-mem with 600x lower cost.

The #1 Obsidian MCP for AI memory — freshness-aware, cited, local-first and read-only by default. Dataview, Bases, PDFs, every agent.

Robot Memory - Persistent memory system for AI robots. MCP Server + hybrid search + spatial retrieval.

Federated, local-first search for an AI — one query across transcripts, files, knowledge graph, vector store, and the web, fused by trust-weighted RRF. Apache-2.0.

Persistent memory system for AI coding assistants. Captures decisions, learnings, and context from coding sessions. Features hybrid search (semantic + BM25), MCP server integration, SQLite persistence with knowledge graph, and proactive memory surfacing. Written in Rust.

Self-hosted AI knowledge base with hybrid semantic search (pgvector + FTS + RRF), MCP server, multi-provider LLM inference (Ollama, OpenAI, OpenRouter, llama.cpp), multimodal ingestion (vision, audio transcription, speaker diarization), and knowledge graph. Rust + PostgreSQL.

Give any MCP-capable agent persistent memory: remember/recall over a tiered store with hybrid vector + keyword retrieval. Single Go binary, SQLite or Postgres, embedded admin UI.

Local code search for AI coding agents: a CLI and MCP server with hybrid keyword + semantic search and SQL relevance-ranked aggregation over an index in plain files. No accounts, no keys, no server.

Standalone MCP server for Obsidian vaults — hybrid search, notes & files, memory, tasks, OAuth 2.1.

Persistent AI memory for Claude Code, OpenClaw, and any MCP-compatible agent. BM25F + vector hybrid, governance-aware, local-first, zero-infrastructure.

Open-source, self-hosted knowledge backend for AI agents — hybrid search (vector + keyword), MCP server, 5 connectors, Docker-ready

The most capable Calibre MCP server — full read/write tools plus multilingual semantic search over your whole ebook library or a single book. For Claude & any MCP client.

Persistent AI memory with hybrid search and embedded sync. Open, free, unlimited.

MCP server for German & EU law. Verified, citable legal context for any LLM. Daily updates from official sources, hosted in Germany

ChatGPT-like AI that runs 100% locally on your hardware. No subscriptions, no cloud, complete privacy. Multi-agent swarm + 10 MCP tools + hybrid RAG vector DB + . Runs on one GPU (RTX 5090 recommended)

A memory graph designed for the agent that uses it, not the human who feeds it. Engrama reconstructs context from associations on demand, replacing the "stuff everything into the prompt" reflex with targeted graph traversal. SQLite default, Neo4j optional.

The continuity layer for everything you do with AI. A self-hosted, open source server any tool plugs into over MCP or REST: hybrid vector + lexical recall, an auto-built knowledge graph, sleep-style consolidation, procedural and persona tiers, OAuth 2.0, multi-tenant. SQLite or Postgres. MIT.

MCP server for AI memory -- hybrid search (BM25 + semantic + knowledge graph), temporal decay, local-first

Persistent memory for Claude Code. Automatically indexes every conversation and provides production-grade hybrid search (BM25 + vectors + reranker) via MCP tools. 100% local, zero config, zero API keys.

Open-source team memory layer for AI coding agents — markdown files in git, user→team→org hierarchy, cross-vendor MCP server. Apache-2.0.

Musubi (結び) — Ai Agent shared memory and thought layer. The braiding of threads between presences.

Experimental MCP tool-list filtering proxy for large agent tool catalogs.

Memory module for LLMs and Agents with MCP

Fast hybrid (BM25 + semantic) local code search for AI agents - pure Rust, persistent index, MCP/gRPC servers, tree-sitter symbols

A team's knowledge, captured where the work happens, filed by an agent, and answered with citations you can check.

Self-hosted MCP server for hybrid semantic code search and repository intelligence.

Long-term memory your AI coding agents actually share. One brain, many agents — standalone embedded MCP server: SQLite + FTS5 + vector hybrid search, local ONNX embeddings, self-expiring diary. No server, container, or daemon.

Compression Runtime for Universal eXecution — a 60–95% token reducer for AI coding agents (Claude Code, Cursor, Cline, …). Single Rust binary, SQLite, 11 layers, local-first.

Local-first RAG engine with cited answers from your own docs. GraphRAG, sealed sharing, MCP-native, VS Code Copilot integration. Nothing leaves your machine.

MCP server for querying ARCA/AFIP web services documentation with natural-language semantic search across 49 official services

RE-call — Retrieval-Augmented Self-Recall: RAG over an AI agent's own memory that knows when it doesn't know (gap detection, freshness, anti-re-litigation). PostgreSQL + pgvector, hybrid retrieval + RRF.

Portable hybrid (BM25 + dense + RRF) retrieval engine and a label-free evaluation harness — extracted from a personal AI-assistant memory index and decoupled to run on any source tree.

Production-ready context database and MCP server for AI assistants. Provides code intelligence with hybrid search, call graphs, impact analysis, and semantic search for Cursor, VS Code, and Claude Desktop. Docker-first, 248 languages, auto-indexing.

Persistent identity and memory across AI tools - mcp-native, local-first, framework-agnostic, production-ready.

Pre-release PostgreSQL-backed MCP memory server with hybrid search and hierarchical summaries

Self-hosted MCP memory server with hybrid semantic + keyword search on Cloudflare Workers

Project-aware collection management based on Qdrant, including a Rust MCP, daemon and CLI: hybrid semantic, pattern and full-text (FTS5) search into single or cross-concerns collection. Dedicated collections for knowledge library, LLM behavioral rules, and an LLM scratchpad

Store and recall robotic experiment data to improve task success by learning from past parameters, trajectories, and outcomes.

AI Architect Codebase: cross-platform code intelligence MCP for Claude Code, Codex, Gemini, Cursor, VS Code, and Zed. Tree-sitter AST to LadybugDB graph, hybrid search, and impact analysis.

Local-first multilingual memory for Codex, Claude Code & MCP — auditable Qdrant recall, project isolation, and a single-instance desktop tray.

Persistent layered memory MCP server for AI agents — SQLite + FTS5 + vector hybrid search (RRF), multilingual, zero-cloud