z-ai/glm-4.5-air:free
open_router/z-ai/glm-4.5-air:free
Context
—
Max output
16K
Input / 1M
$32.00
Output / 1M
$22.998
Modality
Chat
Cutoff
2323
About this model
qwen/qwen3-coder:free
Best suited for
- qwen/qwen3-coder:free
Capabilities
Vision
Accepts images alongside text in the same message.
System prompt
Honours a dedicated system role, separate from the user turn.
Supported parameters
creativity_levelControls the creativity of responses. Higher values (e.g., 0.7) increase creativity; lower values (e.g., 0.2) make responses more predictable.
max_tokensMax Tokens LimitSpecifies the maximum number of text units (tokens) allowed in a response, limiting its length.
probability_cutoffProbability Cutoff (Top P)Focuses on the most likely words based on a percentage of probability.
log_probabilityIf true, returns the log probabilities of each output token returned in the content of message.
repetition_penaltyThe `frequency_penalty` controls how often the model repeats itself, with higher positive values reducing repetition and negative values encouraging it.
novelty_penaltyDiscourages responses that are too similar to previous ones.
stopThis parameter tells the model to stop generating text when it reaches any of the specified sequences (like a word or punctuation)
response_typeDefines the format or type of the generated response.
parallel_tool_callsEnables parallel execution of tools, allowing multiple tools to run simultaneously.
streamSends the response in real-time as it's being generated.
Use this model
OpenAI-SDK compatible, with gateway fallback and routing across providers.
Other Open Router models
| Model | Context | Max output | Input | Output | Capabilities |
|---|---|---|---|---|---|
| deepseek/deepseek-chat-v3-0324:free Chat | — | 16K | Free | Free | System prompt |
| openai/gpt-4o Chat | 128K | 16K | $2.50 | $10.00 | Vision, Tools, System prompt |
| deepseek/deepseek-r1-0528-qwen3-8b:free Chat | — | 16K | Free | Free | System prompt |
| deepseek/deepseek-r1:free Chat | 164K | 16K | Free | Free | System prompt |
| google/gemma-3n-e2b-it:free Chat | — | 16K | Free | Free | System prompt |
| qwen/qwen3-coder:free Chat | — | 16K | Free | Free | System prompt |
| moonshotai/kimi-k2:free Chat | — | 16K | Free | Free | Vision, Tools, System prompt |
| tencent/hunyuan-a13b-instruct:free Chat | — | 16K | Free | Free | Vision, Tools, System prompt |
| mistralai/devstral-small-2505:free Chat | — | 16K | Free | Free | Vision, Tools, System prompt |
| cognitivecomputations/dolphin-mistral-24b-venice-edition:free Chat | — | 16K | Free | Free | Vision, Tools, System prompt |
| mistralai/mistral-small-3.2-24b-instruct:free Chat | — | — | Free | Free | Vision, System prompt |
| google/gemini-3-flash-preview Chat | — | 16K | $0.01 | $0.12 | Vision, Tools, System prompt |
| openai/gpt-5.4-mini Chat | — | 128K | $0.75 | $4.50 | Vision, Tools, System prompt |
| cognitivecomputations/dolphin3.0-mistral-24b:free Completion | — | 32K | Free | Free | System prompt |
Pricing & provider details
Open Router
open_router/z-ai/glm-4.5-air:free
- Input · per 1M
- $32.00
- Output · per 1M
- $22.998
- Cached input · per 1M
- —
- Context window
- —
- Max output
- 16,384 tokens
- Knowledge cutoff
- 2323
- Auto-router
- Not supported
Start building today
Route z-ai/glm-4.5-air:free — and every other model in the catalogue — through one endpoint, with failover built in.