GLM 5 API: Amazon Bedrock (US)'s chat / LLM model, one key away
GLM 5 via the ApexApi gateway
Z.AI's current flagship, a 200K tool-calling model positioned for agentic work at open-weight prices. Call it through ApexApi's unified, OpenAI-compatible API with one `ak-` key, alongside every other model in the catalog. It supports a 200,000-token context window, tool / function calling, and streaming responses. Pricing is $1.20 input and $3.84 output per 1M tokens, billed pay-per-use from your credit balance with no subscription.
GLM 5 is served through ApexApi's unified, OpenAI-compatible API. One key, one format, with smart routing and automatic failover. Provider: Amazon Bedrock (US).
Why GLM 5
200,000-token context
Reason over long documents, codebases, and multi-turn history in a single request.
Tool & function calling
Structured outputs and tool orchestration for agentic and automation workloads.
OpenAI-compatible
Drop-in with any OpenAI SDK. Change the base URL and key, keep your code.
Specifications
- Modality
- chat
- Context window
- 200K tokens
- Max output
- —
- Streaming
- Yes
- Tool / function calling
- Yes
- Vision input
- —
GLM 5: measured reliability
From real calls ApexApi made to GLM 5 during health sweeps over the last 30 days. Our own measurements, not vendor claims.
- Success rate
- 100%
- Avg latency
- 6.8s
- Calls measured
- 1
- Last checked
- Aug 30
See how GLM 5 compares against every model we serve on the model rankings.
GLM 5 pricing
- A million tokens in and a million out costs $5.04 here, all-in.
- For lighter work, GLM 4.7 Flash from the same maker costs 8× less — one string apart on the same key.
Getting started with GLM 5
OpenAI-compatible. Point your client at ApexApi, swap your key, and call it. No new SDK to learn.
from openai import OpenAI
client = OpenAI(
base_url="https://api.apexapi.dev/v1",
api_key="<YOUR_APEXAPI_KEY>",
)
response = client.chat.completions.create(
model="zai/glm-5",
messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)GLM 5 use cases
What teams build with GLM 5 on ApexApi.
Agents & automation
Tool-using agents that orchestrate multi-step workflows and call APIs.
Coding assistants
Code generation, refactoring, reviews, and multi-file reasoning.
RAG & knowledge apps
Summarization and long-document Q&A grounded in your own data.
GLM 5 vs other models
How it compares, and on ApexApi you can switch between any of them with a single string change, same key, same endpoint.
GLM 5 vs Claude Opus 4.7
Compare GLM 5 against this model on price, capabilities, and latency. A/B them with a single key on ApexApi.
Learn more about Claude Opus 4.7 API →GLM 5 vs GPT-5.5
Compare GLM 5 against this model on price, capabilities, and latency. A/B them with a single key on ApexApi.
Learn more about GPT-5.5 API →GLM 5 vs Gemini 3.1 Pro Preview
Compare GLM 5 against this model on price, capabilities, and latency. A/B them with a single key on ApexApi.
Learn more about Gemini 3.1 Pro Preview API →GLM 5 API: FAQ
How do I call GLM 5 through ApexApi?
Create an ApexApi key, point your client at https://api.apexapi.dev/v1, and POST to /v1/chat/completions with the model set to `zai/glm-5`. The API is OpenAI-compatible, so an existing OpenAI SDK works once you change the base URL and key.
How much does GLM 5 cost on ApexApi?
$1.20 per million input tokens and $3.84 per million output tokens, all-in with no separate platform fee. You buy credits and pay per call, with no subscription or minimum.
What is the context window for GLM 5?
200,000 tokens.
Does GLM 5 support tool calling and image input?
GLM 5 supports tool / function calling and streaming responses, and it does not support image (vision) input.
What should I use instead of GLM 5?
From the same maker, GLM 4.7 Flash costs less for lighter work. Switching is one string: it runs on the same endpoint and the same key, so there is nothing to re-integrate.
Ready to build with GLM 5?
One API key. Every major AI model. Pay only for what you use.
Get your API key, free to start