Free GLM 5.3 API: Use GLM 5.3 and GLM 5.3 Flash Without Paying
By Vibhek SoniUpdated 3 min read
Quick answer
Sign up at freetheai.org, create an API key, do the daily check-in, then call https://api.freetheai.org/v1/chat/completions with the model fta/zai/glm-5.3. Use fta/zai/glm-5.3-flash when you need images or faster replies.
Key takeaways
- GLM 5.3 is free on FreeTheAI as
fta/zai/glm-5.3, with a lighterfta/zai/glm-5.3-flash. - Any app that speaks the OpenAI API works; Claude Code and other Anthropic clients work too.
- GLM 5.3 thinks before it answers by default. You can turn that off per request.
- Send images to GLM 5.3 Flash, not GLM 5.3, which reads text only.
- The free tier gives 50 requests a day to start, after a quick daily check-in.
What you get for free
FreeTheAI serves two GLM models on its free tier. Both work through one key and one base URL, and both are listed with their current limits on the models page.
| Model ID | Best for | Images |
|---|---|---|
fta/zai/glm-5.3 | Coding, agents, long documents, roleplay | No (text only) |
fta/zai/glm-5.3-flash | Quick answers, background tasks, pictures | Yes |
Both models list a very large context window and up to about 131K output tokens on their model pages. Use the full model ID, including the fta/zai/ prefix, in every request.
Get a key and make your first call
- Create a free account at freetheai.org/signup with a Gmail, Outlook, Yahoo, or iCloud address and confirm your email.
- Open Dashboard > API keys and create a key. It is shown once, so copy all of it.
- Do today's check-in. Free models unlock until 00:00 UTC.
- Send a request to the base URL below.
curl https://api.freetheai.org/v1/chat/completions \
-H "Authorization: Bearer ftai_your_key_here" \
-H "Content-Type: application/json" \
-d '{"model": "fta/zai/glm-5.3", "messages": [{"role": "user", "content": "Say hello in five words."}]}'In an app, the base URL is https://api.freetheai.org/v1. Apps that ask for the full address (JanitorAI, for example) need https://api.freetheai.org/v1/chat/completions.
Turn thinking on or off
GLM 5.3 reasons before it answers unless you ask it not to. Thinking helps with code and hard questions, but it makes replies slower and uses more tokens. To turn it off for a request, send either of these fields:
{
"model": "fta/zai/glm-5.3",
"thinking": { "type": "disabled" },
"messages": [{ "role": "user", "content": "Summarize this in one line." }]
}Or use the OpenAI-style switch "reasoning_effort": "none". Leave both out to keep the model's default. The reasoning, when there is some, arrives in reasoning_content, separate from the reply text.
Images: use GLM 5.3 Flash
GLM 5.3 reads text only. A request that includes a picture gets a clear 400 error that says the model accepts text only. Send pictures to fta/zai/glm-5.3-flash instead, as standard OpenAI image_url content parts (a data URL or an https link).
Use it in your apps
- Roleplay apps: JanitorAI, SillyTavern, Chub AI, and RisuAI all work with
fta/zai/glm-5.3. - Coding agents: Claude Code, Cline, OpenCode, and Kilo Code. Tool calls work.
- Chat apps: Open WebUI, LibreChat, Chatbox, and more in the setup guides.
Common errors and fixes
| What you see | What to do |
|---|---|
| "Daily check-in required" | Check in at freetheai.org/checkin. It lasts until 00:00 UTC. |
| "Content filtered" | The model's filter blocked the prompt or its reply. Rephrase, or try fta/kimi/k3. |
| "This model is busy right now" | Wait a minute and retry, or switch to GLM 5.3 Flash. Busy refusals do not use a free request. |
| "This model accepts text only" | Remove the image, or use fta/zai/glm-5.3-flash. |
| Replies stop early | Raise max tokens in your app (or send 0 for the model's maximum) and turn on streaming. |
Frequently asked questions
Is the GLM 5.3 API really free?
Yes. FreeTheAI serves fta/zai/glm-5.3 and fta/zai/glm-5.3-flash on its free tier. You need a free account, an API key, and a quick daily check-in. No card is needed.
What is the GLM 5.3 model ID?
fta/zai/glm-5.3 for the full model and fta/zai/glm-5.3-flash for the lighter one, sent to https://api.freetheai.org/v1.
Can GLM 5.3 read images?
GLM 5.3 is text only. Use fta/zai/glm-5.3-flash for requests with pictures.
How do I stop GLM 5.3 from thinking?
Send "thinking": {"type": "disabled"} or "reasoning_effort": "none" in the request.
How many free requests do I get?
You start with 50 a day. Linking Discord adds 50 more, and donating an API key can add more. The count resets at 00:00 UTC.
Was this post helpful?
About the author
Vibhek Soni founded FreeTheAI and builds and runs its API. He answers the support tickets these posts come from.
Published
Try the free API
Create a free account, make an API key, and do the daily check-in. No credit card.
Create a free account