Z AI (Zhipu AI) logo

GLM-4.5-Flash (retirement announced) Free API on Z AI (Zhipu AI)

Free API Verified
Catalog profile Reasoning provider catalog metadata

GLM-4.5-Flash (retirement announced) — free model from Z AI (Zhipu AI).

Free APIOpenAI compatibleReasoningTool callingTextReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://open.bigmodel.cn/api/paas/v4
Model ID
glm-4.5
API format OpenAI Chat Completions
Technical Details

GLM-4.5-Flash (retirement announced) specifications

Provider catalog
Context window 128K
Max output 96K
Status Online
Family glm-flash
Knowledge cutoff 2025-04
Released Jul 28, 2025
Last updated Sep 8, 2026
Free listing since Jul 28, 2025
Input text, reasoning
Output text
Capabilities reasoning, tool calling
AI Recommendation

Should you use GLM-4.5-Flash (retirement announced)?

GLM-4.5-Flash (retirement announced) is listed for chat workloads and supports a 128K context window.

Use it when Z AI (Zhipu AI)'s free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Reasoning mode listed
  • Tool calling support
  • Works with OpenAI-style SDKs

Watch outs

  • Free-tier rate limits apply
  • Vision support is not listed
Pricing

GLM-4.5-Flash (retirement announced) pricing per 1M tokens

Free tier listed
Input $0.6 per 1M tokens
Output $2.2 per 1M tokens
Free access Available Z AI (Zhipu AI)
Rate limit 1 concurrent request provider policy
Typical Use Cases

GLM-4.5-Flash (retirement announced) use cases

Chat

GLM-4.5-Flash (retirement announced) is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

GLM-4.5-Flash (retirement announced) free API FAQ

Is GLM-4.5-Flash (retirement announced) free to use?

GLM-4.5-Flash (retirement announced) is listed with free API access on Z AI (Zhipu AI), subject to the provider's quota and account policy.

What is the GLM-4.5-Flash (retirement announced) model ID?

The model ID shown in this catalog is glm-4.5.

What are the GLM-4.5-Flash (retirement announced) free tier rate limits on Z AI (Zhipu AI)?

The listed free tier limit is 1 concurrent request. Limits can change per account tier, so confirm against the provider dashboard.

What context window does GLM-4.5-Flash (retirement announced) support?

The listed context window is 128K tokens with up to 96K output tokens.

More about GLM-4.5-Flash (retirement announced)

GLM-4.5-Flash (retirement announced) — free model from Z AI (Zhipu AI).

For API keys, setup steps, and provider-level limits, see the Z AI (Zhipu AI) provider page.