anthropics ★ 51k Claude Skill experimental Cache long system prompts and context blocks for up to 90% cost reduction.
Best match because its name, category, capabilities, or owner matches “caching”. Adoption and trust break close ties.
Stars 51k
Downloads —
Updated 2026-07-23
Skill installs — Trust pending Momentum not matched
$ pip install anthropic-cookbookCopy Prompt Caching
sgl-project ★ 31k Repository experimental Fast serving framework for LLMs with radix attention based prefix caching for lower cost.
Best match because its name, category, capabilities, or owner matches “caching”. Adoption and trust break close ties.
Stars 31k
Downloads 100357k
Updated 2026-08-03
Skill installs — Trust pending Momentum not matched
$ pip install sglangCopy Prompt Caching
KV cache management layer that speeds up LLM serving by reusing cached context.
Best match because its name, category, capabilities, or owner matches “caching”. Adoption and trust break close ties.
Stars 11k
Downloads 68k
Updated 2026-08-02
Skill installs — Trust pending Momentum not matched
$ pip install lmcacheCopy Prompt Caching
An AI gateway offering built in semantic and exact prompt caching for cost optimization.
Best match because its name, category, capabilities, or owner matches “caching”. Adoption and trust break close ties.
Stars 7.0k
Downloads —
Updated 2026-08-03
Skill installs — Trust pending Momentum not matched
Prompt Caching
Portkey-AI ★ 13k Tool stable Fast open-source AI gateway with automated retries, semantic caching, and budget rules.
Best match because its name, category, capabilities, or owner matches “caching”. Adoption and trust break close ties.
Stars 13k
Downloads 4.1k
Updated 2026-05-25
Skill installs — Trust pending Momentum not matched
$ npm i @portkey-ai/gatewayCopy LLM Gateways
messkan ★ 243 Tool experimental Provider agnostic proxy optimized for semantic prompt caching and latency reduction.
Best match because its name, category, capabilities, or owner matches “caching”. Adoption and trust break close ties.
Stars 243
Downloads —
Updated 2026-07-31
Skill installs — Trust pending Momentum not matched
Prompt Caching
zilliztech ★ 8.1k Tool slowing Semantic cache for storing LLM responses to bypass redundant API calls.
Best match because its name, category, capabilities, or owner matches “caching”. Adoption and trust break close ties.
Stars 8.1k
Downloads 437k
Updated 2025-07-11
Skill installs — Trust pending Momentum not matched
$ pip install gptcacheCopy Prompt Caching
Open-source LLM observability platform with real-time token tracking, caching, and cost analytics.
Best match because its name, category, capabilities, or owner matches “caching”. Adoption and trust break close ties.
Stars 6.0k
Downloads 3.9k
Updated 2026-07-25
Skill installs — Trust pending Momentum not matched
$ npm i @helicone/heliconeCopy Observability & Cost
TensorOpsAI ★ 388 Tool experimental Gateway and management tool for calling, caching and monitoring multiple LLM providers.
Best match because its name, category, capabilities, or owner matches “caching”. Adoption and trust break close ties.
Stars 388
Downloads —
Updated 2026-07-29
Skill installs — Trust pending Momentum not matched
$ pip install llmstudio-monorepoCopy LLM Gateways
openrouter-ai Tool experimental Unified LLM gateway with dynamic price routing, prompt caching, and cost analytics.
Best match because its name, category, capabilities, or owner matches “caching”. Adoption and trust break close ties.
Stars —
Downloads —
Updated —
Skill installs — Trust pending Momentum not matched
LLM Gateways Model Routing