Skip to content
Built here

Semantic cache for LLM calls

/ai--semantic-cache

/ai--semantic-cache is a Claude Code skill in the AI & Agents section. Adds a cache in front of LLM calls so repeated or near-duplicate requests get a stored answer, and measures whether it helps without serving wrong ones.

Author's description

Add semantic caching to LLM calls - embedding keys, thresholds, invalidation

How to ask for it

  • /ai--semantic-cache cache answers for similar questions, not just identical ones
  • /ai--semantic-cache tune the similarity threshold for the semantic cache
  • /ai--semantic-cache invalidate the cache when source documents change

Install

curl -fsSL https://raw.githubusercontent.com/sgomez-dev/claude-skills/main/install.sh | bash

After installing with the script, type /ai--semantic-cache. Using Cursor, Windsurf or Codex? Platform guides

Demo coming soon

↑↓ move · Enter open · Esc close