For developers
Does Caveman add latency to provider calls?
Does Caveman add latency to provider calls?
What happens if Caveman is down?
What happens if Caveman is down?
caveman start, caveman wrap) need no account and keep working if Cloud is unreachable. For hosted gateway traffic, if the provider is unreachable, the gateway returns its own error with a Caveman request ID so you can trace it. Cache outages cost you the hit, never the call. If an optimization fails during processing, the gateway falls back to the original request and continues.See Troubleshooting for exit codes and diagnosis steps.Can Caveman change my model output?
Can Caveman change my model output?
x-cave-optimize: compress, and then rewrites tool output the model sees, losslessly or with an encrypted original for recovery. record mode is always pure pass-through. Unknown modes fail closed to the most conservative handling.See Rollout Safety for how changes are gated before they reach production.Do you store my prompts and responses?
Do you store my prompts and responses?
x-cave-retention: metadata or x-cave-retention: zdr on any request to override per-call.See Data and Privacy for the full consent and retention model.Do you train on my data?
Do you train on my data?
Which providers and SDKs do you support?
Which providers and SDKs do you support?
How do I remove Caveman from my stack?
How do I remove Caveman from my stack?
x-cave-upstream-key header if you were using per-request keys. If you stored provider keys in Caveman, rotate them at your provider after removal. No code changes are required beyond the configuration revert.See Quickstart for the original one-line integration pattern.What is the difference between measured, inferred, and verified savings?
What is the difference between measured, inferred, and verified savings?
Why do my verified savings show zero?
Why do my verified savings show zero?
record mode (pass-through), the provider or model is not on the allow-list for verified methods, a blocker such as output-shape prevented count-baseline verification, the model is unpriced in the catalog, the call failed, or it was a cache hit with no counterfactual provider response to compare.See Troubleshooting for the complete list.How do I control which optimizations run?
How do I control which optimizations run?
x-cave-optimize: off for pass-through, x-cave-optimize: compress to ask for compression, x-cave-cache: semantic,ttl=900 for semantic cache with a 15-minute lifetime. Unrecognized values return 400 with a clear error code.See Control Optimizations for the full header catalog.What permissions does my API key need?
What permissions does my API key need?
trace:read_metadata. For billing columns, you need billing:read. For payload reads, you need payload:read plus trace:read_payload.See Authentication for key types and scopes.Can I run SQL against my telemetry?
Can I run SQL against my telemetry?
sql:read keys are limited to 101 rows, 128 KiB, and 15 seconds. Human queries get up to 10,000 rows, 8 MiB, and 30 seconds.See Query with SQL for schema details and examples.For engineering leaders
How does Caveman fit with our existing stack?
How does Caveman fit with our existing stack?
When does Caveman actually save money, and when does it not?
When does Caveman actually save money, and when does it not?
How do we govern spend and prevent runaway costs?
How do we govern spend and prevent runaway costs?
What governance and audit capabilities do we get?
What governance and audit capabilities do we get?
Can we run Caveman in our own cloud or on-prem?
Can we run Caveman in our own cloud or on-prem?
caveman start / wrap) also runs on your own machine with no account required, though it stays inference-only with no verified-savings claims.See Deployment Options for hosted, your cloud, and local deployment paths.How does tenant isolation work?
How does tenant isolation work?
What is the rollout safety model?
What is the rollout safety model?
Can we self-host without sending anything to Caveman's cloud?
Can we self-host without sending anything to Caveman's cloud?
caveman start / wrap) sends usage telemetry to Caveman by default, but you can opt out with caveman telemetry off, CAVEMAN_TELEMETRY=0, or DO_NOT_TRACK=1. The CLI telemetry includes an install ID and client IP, but never prompts, responses, code, or file paths. Customer-owned cloud deployments keep operational telemetry in your own ClickHouse instance.See Deployment Options for air-gapped and customer-cloud details.How do we evaluate Caveman before committing?
How do we evaluate Caveman before committing?
What happens to our data if we stop using Caveman?
What happens to our data if we stop using Caveman?
For finance and procurement
How is Caveman priced?
How is Caveman priced?
What spend numbers can we trust for reporting?
What spend numbers can we trust for reporting?
Why are our verified savings zero?
Why are our verified savings zero?
How do we report savings to stakeholders?
How do we report savings to stakeholders?
What is the total cost of ownership?
What is the total cost of ownership?
What legal documents and certifications are available?
What legal documents and certifications are available?
Who are the subprocessors, and where is data hosted?
Who are the subprocessors, and where is data hosted?
europe-west4 with ClickHouse Cloud in the same region. Subprocessors include Google Cloud, ClickHouse Cloud, Cloudflare, Resend, PostHog, Stripe, Supabase, Vercel, Google Workspace, Cal.com, and GitHub. Your upstream model providers reached via BYOK are your own processors, not Caveman subprocessors.See Security Review for the complete subprocessor table and region details.Can we negotiate custom terms?
Can we negotiate custom terms?
How do we handle procurement security questionnaires?
How do we handle procurement security questionnaires?