Enterprise DNA Enterprise DNA
M MCP Servers Developer low

kskurtveit/echocache

by Various

MCP server that caches expensive LLM results — HTTP-style freshness plus semantic recall of related answers

MCP

kskurtveit/echocache

Added 16 Sept 2026

#ai-agents #cache #claude #llm #mcp #mcp-server #model-context-protocol #typescript

Overview

MCP server that caches expensive LLM results. Uses HTTP-style freshness to expire cached entries and semantic recall to retrieve related answers. Written in TypeScript.

Best for

Best for
Developers building MCP servers that need to cut LLM costs.

Use cases

  • Cache repeated LLM calls with TTL-based freshness
  • Retrieve semantically similar past answers for new queries
  • Reduce cost and latency in MCP-based LLM workflows

How to use

Install

claude mcp add echocache -- npx -y echocache

Tools exposed

  • ECHOCACHE_DB_PATH
  • ECHOCACHE_MAX_ENTRIES
  • ECHOCACHE_MAX_BYTES
  • ECHOCACHE_DEFAULT_TTL_SECONDS
  • ECHOCACHE_SIMILARITY_THRESHOLD
  • ECHOCACHE_LINK_CANDIDATE_POOL
  • ECHOCACHE_ENCRYPTION_KEY
  • cache_get
  • cache_set
  • cache_query
  • cache_related
  • cache_invalidate
  • cache_stats

Tested with

Claude Desktop, Claude Code, Cursor, VS Code

Notes

MCP server that caches expensive LLM results. Uses HTTP-style freshness to expire cached entries and semantic recall to retrieve related answers. Written in TypeScript.

0 stars on GitHub. Last updated 2026-09-10. Licensed MIT.

Use cases

  • Cache repeated LLM calls with TTL-based freshness
  • Retrieve semantically similar past answers for new queries
  • Reduce cost and latency in MCP-based LLM workflows

Pros

  • Simple HTTP-style caching model
  • Semantic recall improves hit rate for paraphrased queries
  • Lightweight TypeScript implementation

Cons

  • Requires running and maintaining a separate MCP server
  • Semantic recall may need embedding infrastructure
  • No stars or community adoption yet

Indexed from awesome-mcp-servers-punkpeye and enriched against its public facts.

Pros

  • Simple HTTP-style caching model
  • Semantic recall improves hit rate for paraphrased queries
  • Lightweight TypeScript implementation

Cons

  • Requires running and maintaining a separate MCP server
  • Semantic recall may need embedding infrastructure
  • No stars or community adoption yet

Pairs with

Other entries in the index that connect to this one. Click through to see the chain.

Free 27-page guide

Get the free Developer’s Field Guide

A 27-page field guide to the AI coding workflow with Claude. Claude Code, MCP servers, the prompt patterns that work, and what to delegate. Free.

Enter your work email. We send it straight over, plus a few short notes worth knowing. Unsubscribe any time.

No spam. Unsubscribe any time.

Running a business, not writing the code? See the MCP servers picked for operators, and get your first one wired up with us.

Operator picks