Tagged “rag”
26 listings
Servers

headroom
Updated todayCompress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.

code-graph-rag
Updated todayby vitali87
The ultimate RAG for your monorepo. Query, understand, and edit multi-language codebases with the power of AI and knowledge graphs

neo
Updated todayby neomjs
Neo.mjs is a self-evolving software organism: a professional end-to-end AI engineering team whose cross-model swarm inhabits live apps via Neural Link, Active Hybrid GraphRAG, DreamService, and self-healing loops.

codanna
Updated 18d agoby bartolli
Local code intelligence MCP server and CLI for AI coding agents

Compartment
Updated todayEncrypted, fully offline agentic memory. One click install, GUI w/ memory map, all OS and agents. Logically and mathematically superior storage and retrieval.

caura-memclaw
Updated todayby caura-ai
Governed shared memory for AI agent fleets — multi-agent, multi-tenant, MCP-native. Trust tiers, keystone policies, audit trails, knowledge graph, self-improving retrieval. Apache 2.0.

lilbee
Updated 2d agoby tobocop2
The whole local AI stack in one executable: it runs and manages local AI models across every GPU, and it's a search engine you can talk to, with cited answers from your files, code, and the web. MCP server for coding agents, web crawler, TUI, CLI, REST API, Python library. No Ollama or LM Studio needed, works with both.

Vault-Agent-Memory
Updated 21d agoby zycaskevin
Local-first memory governance for AI agents: shared, reviewable, auditable memory via SQLite and MCP.

vybe-intelligence-vault
Updated todayby sairaman436
An auto-updating open-source vault for AI agents, RAG systems, MCP servers, prompts, tools, templates, and next-generation web development.

memo
Updated todayby jagoff
Your coding agent starts every session with amnesia — memo fixes that, 100% on your machine. Persistent memory for Claude Code, Codex, Cursor & any MCP client: Markdown source of truth, hybrid search (MLX/CPU + sqlite-vec), time-machine, contradiction radar, nightly self-optimization. No cloud, no keys.

MailFathom
Updated todayby Krzysztof318
A brain for your mail: MailFathom turns IMAP mailboxes into a self-hosted, AI-native service. Mail synchronizes into your own PostgreSQL, is indexed for search and retrieval, and is served to AI agents over the Model Context Protocol. Read-only today; semantic retrieval, answering, and gated write tools next. .NET 10, Apache-2.0.

memtomem
Updated todayby memtomem
Markdown-first, long-term memory infrastructure for AI agents. Hybrid BM25 + semantic search across markdown/code files via MCP.

murmur
Updated 2d agoby murmur-io
🎙️🧠 Local-first macOS meeting + document notebook — on-device Whisper transcription, source-linked notes, @brain and history Q&A, a read-only local MCP server, E2EE sharing, and plain Markdown exports. Run AI locally or explicitly opt into redacted cloud AI · AGPL-3.0

exomem
Updated todayby Artexis10
Self-hosted MCP server that makes your Obsidian/markdown vault searchable — text, PDFs, Office docs, images, audio — from any MCP client. Hybrid retrieval over a typed, governed corpus, sub-second at 50k notes; your files stay plain markdown.

inspeximus
Updated todayby DanceNitra
Zero-dependency agent memory + MCP server. Value-ranked recall, consolidation, and a first-class correction & erasure channel (revert, lineage-aware retraction, tamper-evident receipts). Measured integrity vs mem0/Graphiti.

Diariz
Updated 3d agoby kenhayward
Self-hosted, multi-user transcription platform: record or upload audio. Speaker-labeled, timestamped transcripts, Recognize speakers across recordings, Summarize, extract action items and chat over your transcripts with your own OpenAI-compatible LLM. Your Audio, your Server, your Model. Tested on Laptop RTX4070, Desktop RTX3090 and RTX5090

axon
Updated 2d agoSelf-hosted RAG stack for crawl, scrape, search, ingest, query, and ask workflows with Qdrant, TEI embeddings, Chrome rendering, MCP/CLI/REST, and Gemini synthesis.

captain-memo
Updated 5d agoLocal, cross-AI memory for coding agents — Claude Code, Codex, Gemini, Antigravity, Cursor, Kimi & more share one private corpus. Hybrid search, auto-injected context, no-API-key summarizers, runs fully local via Ollama.

search-mcp
Updated 6d agoCloudflare AI Search toolkit: MCP server, streaming /ask Worker, docs widget, and git-to-R2 corpus sync. AGPL.

ExternalBrain
Updated 14d agoby bejranonda
Self-improving, self-hosted memory across every AI coding tool, project, and team (Claude Code, Cursor, Copilot, any MCP client). Autoskill proposes new skills from your sessions, so each project improves automatically. Inspectable, grounded, and yours. Built for teams and enterprise. Open source, MIT.

agent-infrastructure-landscape
Updated todayby MrPeppersDev
AI agent memory & infrastructure landscape — comparative catalog of 912 systems × 68 columns covering memory layers, agent frameworks, runtimes, vector stores, knowledge graphs, MCP servers, benchmarks. Searchable with typed edges, lineages, citations.

mnem-o-matic
Updated 13d agoPerfect recall for imperfect machines. A shared memory layer for LLMs via MCP

nanohype
Updated todayby nanohype
Template catalog and SDK for AI subsystems — agents, RAG, MCP servers, eval harnesses, infrastructure. Part of the nanohype ecosystem.

opedd-mcp
Updated 2d agoby Opedd
Licensed, rights-cleared content for AI agents — verifiable license keys, on-chain proof, EU AI Act attestation. The MCP alternative to unlicensed scraping for RAG.