

GitHub
Transparent semantic cache for LLM API calls on Redis VS
GitHub이란?
Khazad is a transparent semantic cache for LLM API calls. It intercepts LLM HTTP traffic at the httpx transport layer and serves semantically-equivalent requests from a Redis 8 vector cache with zero code changes. Works with OpenAI, Anthropic, Gemini, Azure OpenAI, and Mistral. Model-aware and conversation-aware caching, full streaming support, TTL, and tunable similarity thresholds. Stop paying for the same prompt twice in dev, CI, demos, or production. Open source (MIT).
스크린샷
?
아직 댓글이 없어요. 가장 먼저 남겨보세요!
GitHub에 대한 X의 실제 대화
X에 게시

