Live 24/7 on Cloudflare Global Edge

Semantic Cache & Token Saver API

Sub-millisecond vector similarity cache for OpenAI, Claude, and Gemini. Cut inference costs by 30-60%.


      

📚 Official Endpoints

POST /v1/cache/check (Similarity Match) POST /v1/cache/set (Store Prompt/Response) POST /v1/cache/batch-check (Batch Lookups) GET /v1/cache/stats (Live Savings Metrics) GET /v1/health (Healthcheck)