Back to the feed

Privatemode's encrypted AI free tier gives 5M tokens on sign-up and 1M more each month, no card needed

#Privatemode#Free credits#API deals#GLM

Privatemode is an end-to-end encrypted AI service with a free tier that needs no credit card. New sign-ups get a one-time 5M-token quota, then 1M input and 1M output tokens free every month. Both the web chat and the API are included, with models such as GLM-5.3, GLM-5.3-Flash, and gpt-oss-120b.

The briefing

Privatemode is an AI service built around end-to-end encryption. Your data is encrypted before it leaves your device, inference runs inside confidential-computing environments, and the service is hosted in the EU. The free tier needs no credit card and covers both the web chat and the API.

What you get

A new account gets a one-time initial quota of 5 million tokens once you create an organization. While that balance lasts, the free tier's monthly limits don't apply, and cached prompt tokens aren't deducted from it.

After the initial quota runs out, the free tier still gives you 1 million input tokens and 1 million output tokens a month, 100 minutes of speech-to-text a month, and up to 40 requests per minute. gpt-oss-120b counts at twice the rate, so the same quota goes half as far on that model.

That's not a lot for heavy coding, but it works well for private notes and sensitive documents you'd rather not send to a regular cloud AI service.

What you can use

Chat models are GLM-5.3, GLM-5.3-Flash, and gpt-oss-120b. Both GLM models support a 1M-token context, and GLM-5.3-Flash also accepts images. There's also DeepSeek-OCR-2 for OCR, Qwen3-Embedding 4B for embeddings, and Whisper large-v3 and Voxtral Mini 3B for speech-to-text.

The API is compatible with OpenAI's /v1/chat/completions and Anthropic's /v1/messages, and there are ready-made setup guides for Claude Code, OpenCode, VS Code, and JetBrains.

How to claim

  1. Sign up in the Privatemode portal and create an organization. The initial quota is added right away.
  2. To chat in the browser, sign in at chat.privatemode.ai, pick a model, and start.
  3. To use the API, create a key on the API keys page in the portal. JS/TS projects can install the official SDK (npm install privatemode-ai) and use the model glm-latest; other languages and third-party tools can run the Privatemode Proxy locally to handle verification and encryption.

Details: pricing, quotas and rate limits, and the API quickstart.

Is this still live?

Community signalAwaiting first signal