Skip to content
TokenGauge
Menu

Reference · 16 questions

Questions about SayGM and OpenRouter

Short answers on pricing, fees, privacy and compatibility. Longer treatment is in the guides.

§ 01

General

What is SayGM?

SayGM is an LLM gateway. One API key reaches Claude, GPT, Gemini and a set of open-weight models. There is no subscription and no fee on credit top-ups. Prices are capped at the model maker's list price, and the gateway runs inside an Intel TDX trusted execution environment that publishes remote attestation.

Is SayGM cheaper than OpenRouter?

On the frontier models in this index, yes, by 16 to 34 percent once OpenRouter's 5.5 percent top-up fee is included. The exception is open-weight models: the cheapest OpenRouter providers for Kimi K3 and DeepSeek V4.1 Flash cost less on input tokens. Check the current rate for the model you use.

How does SayGM compare to Vercel AI Gateway, Requesty or buying direct?

Vercel AI Gateway and the labs themselves charge list price for real-time calls; Requesty adds 5 percent unless you bring your own keys. SayGM bills 11 to 30.5 percent below list depending on the lab. Buying direct still wins for batch jobs, which cost half of list, and for contracts such as HIPAA BAAs in your own name.

Does SayGM work with Cursor, Cline and Claude Code?

Yes. SayGM exposes OpenAI-compatible, Anthropic-compatible and Gemini-compatible endpoints. Any tool that accepts a custom base URL works after changing the URL and the API key.

How many models does SayGM support?

About 65 at the time of writing: the current Claude, GPT and Gemini lines plus open-weight models such as Kimi K3, GLM 5.3 and DeepSeek V4.1. OpenRouter lists several hundred. If you depend on a niche or legacy model, confirm it is in the SayGM catalogue first.

Can SayGM see my prompts?

The gateway runs in a trusted execution environment sealed from operators and host machines. Confidential open-weight models (IDs ending in -tee) also run inside the enclave, so prompts are decrypted only there. Frontier providers still receive the prompt, as with a direct API call. SayGM strips account identity and can replace personal data before sending.

Is there a free tier or minimum spend?

No free tier. SayGM is pay as you go: load credits, spend them per token. There is no fee on the top-up itself.

What is Bittensor subnet 28?

SayGM runs on Bittensor subnet 28. Capacity providers bid to serve inference, and that bidding sets the below-list price. Protocol data is public on Taostats.

§ 02

Pricing and fees

Why are SayGM prices below list price?

Capacity providers on Bittensor subnet 28 bid for the right to serve requests. The bids set the per-token price, capped at the maker's list price. SayGM reports an average discount in the low twenties percent across recent epochs.

Does the price change?

Yes. Prices update each epoch and the current rate is on the SayGM pricing page. Figures on this site are a dated snapshot. Verify before budgeting.

Are there other fees?

SayGM lists no top-up fee, no subscription and no per-seat charge. The bill is the per-token price for each model called.

How does OpenRouter's fee work?

OpenRouter charges 5.5 percent on card credit purchases, with a 0.80 dollar minimum, or 5 percent by crypto. The fee is taken when credit is bought, before any inference, on top of provider list prices.

§ 03

Privacy

What is a trusted execution environment?

A trusted execution environment, or TEE, is a hardware-isolated region of a CPU. Code and data inside it are encrypted in memory and unreadable by the operating system, the hypervisor and the machine's operator. SayGM uses Intel TDX.

What is remote attestation?

Remote attestation is a hardware-signed report of which code is running inside the enclave. SayGM publishes the evidence and a third-party verifier checks it, so the running gateway code can be compared to what is claimed.

Which models are fully confidential?

Confidential open-weight models on SayGM, whose IDs end in -tee (for example the TEE variants of Kimi K3 and DeepSeek V4.1 Flash), run inside the enclave end to end. The same models on the open tier do not. Frontier models from Anthropic, OpenAI and Google run on those providers' hardware under their API terms.

How does this compare to OpenRouter's privacy?

OpenRouter's privacy rests on a published policy with opt-in logging. SayGM's gateway privacy rests on hardware isolation that can be checked through attestation. In both cases frontier prompts reach the model provider.