Stop paying twice.
When the exact request already exists, Cache returns the shared result instantly - without generating or charging for another answer.
Hello, my name is
The AI answer is already out there.
Why pay again?
Cache is a shared gateway for AI. Every public request can become a free result for the next person who asks the same thing.
For developers
Keep the API format you already use. Replace the upstream host with Cache and a matching public result is free to rerun anonymously.
Anonymous mode can only return an existing cache hit. If your exact request is new, the gateway will ask for authorization instead of running it.
A better default
When the exact request already exists, Cache returns the shared result instantly - without generating or charging for another answer.
Every successful miss can add one more reusable answer to the commons. The gateway becomes more valuable with every contribution.
Anyone can retrieve a matching public result without an account. No identity required, no platform lock-in attached.
How it works
Cache fingerprints the request, checks the public pool, and only calls the model when it needs something new.
Point an OpenAI-compatible completions request at the Cache gateway.
Matching is deterministic: the same model, prompt, and parameters resolve to the same cached result.
A hit comes back free. A miss runs normally and can become the next person’s hit.
Public means public
Requests and their results are contributed to a public cache that others can rerun. Anonymous access is not confidential storage - never send passwords, private keys, personal data, or anything you wouldn’t share openly.