AI快讯 / AI 开源项目

cachellm

Cut token costs by caching LLM responses locally for OpenAI-compatible providers, eliminating duplicate API calls.

原文来源AI 开源 Releases
查看官方原文

Cut token costs by caching LLM responses locally for OpenAI-compatible providers, eliminating duplicate API calls.