Redis launches managed semantic cache claiming 90% LLM API cost cut

single source · 1 articles · release · confidence: medium · first seen 2026-09-10 22:34 UTC

Redis has launched LangCache, a managed semantic cache (a store that matches differently worded versions of the same question and returns a saved answer rather than calling the model). The company says it cuts LLM API costs by up to 90% and serves cache hits up to 15x faster. The service sits between an application and its LLM provider, aimed at support assistants and retrieval-augmented generation (RAG) pipelines that see repeated intents. The announcement did not include pricing, a general availability date, or independent benchmark details. The article appeared on 10 September 2026.

What this means for you

If you pay per LLM API call for repeated intents, test LangCache in a pilot. Redis claims up to 90% cost cuts and 15x faster hits, but no pricing or GA date is public.

Key facts

  • ·Redis LangCache is a managed semantic cache. source
  • ·Redis claims LangCache cuts LLM API costs by up to 90%. source
  • ·Redis claims cache hits return up to 15x faster. source
  • ·The service is aimed at support assistants and RAG pipelines. source
  • ·The announcement was published on 10 September 2026. source

What the sources say

  • MarkTechPostPress-style introduction to Redis LangCache and its claimed savings for LLM API calls.

Sources

The original reporting. Follow these — they did the work.

← the wire