Pay less. Ship the same.
Save 30–60% of your AI input tokens with a one-line, OpenAI-compatible base URL switch. Input tokens only — not a guaranteed cut to your provider bill.
free tier · no card · keys prefixed ta_Mint a ta_ key in the dashboard.
OPENAI_BASE_URL → the archive.
X-TA-* headers on every response.
Set Override OpenAI Base URL in Settings → Models. Chat and Composer route through the archive.
Choose the OpenAI Compatible provider, paste the base URL and your ta_ key.
Set apiBase in ~/.continue/config.json, or run npx token-archive init --target continue.
Set OPENAI_BASE_URL in ~/.hermes/.env — compresses Hermes → model traffic.
Add openai-api-base to .aider.conf.yml and use your ta_ key.
Add a custom OpenAI connection with the archive URL and your ta_ key.
Pass baseURL / base_url with your ta_ key. openai-python, openai-node, LangChain, LlamaIndex, curl.
Honest limits. The archive cannot route GitHub Copilot or Cursor Tab — those completions run on fixed, first-party endpoints with no base URL override. Everything else that speaks the OpenAI Chat Completions API is fair game.
Sign in with your email, mint a ta_ key, and the dashboard shows tokens saved per request as they land.