Should you use glm-4.7-flash?
glm-4.7-flash is listed for chat workloads and supports a 131K context window.
Use it when Cloudflare Workers AI's free tier is enough for evaluation, demos, or light production traffic.
https://api.cloudflare.com/client/v4/accounts/{account_id}/ai/run @cf/zai-org/glm-4.7-flash glm-4.7-flash is listed for chat workloads and supports a 131K context window.
Use it when Cloudflare Workers AI's free tier is enough for evaluation, demos, or light production traffic.
We only found this glm-4.7-flash listing on Cloudflare Workers AI in the current catalog. Use the related models section for nearby alternatives.
| Provider | Model listing | Access | Context | API | Limits |
|---|---|---|---|---|---|
| | @cf/zai-org/glm-4.7-flash | Free tier | 131K | Native | 10K neurons/day (shared) |
| Model | Provider | Context | Access |
|---|---|---|---|
| Mistral 7B | Cloudflare Workers AI | 33K | Free tier |
| Qwen 1.5 7B | Cloudflare Workers AI | 33K | Free tier |
| @cf/meta/llama-3.3-70b-instruct-fp8-fast | Cloudflare Workers AI | 24K | Free tier |
| @cf/meta/llama-4-scout-17b-16e-instruct | Cloudflare Workers AI | 131K | Free tier |
| @cf/openai/gpt-oss-120b | Cloudflare Workers AI | 128K | Free tier |
glm-4.7-flash is tagged for chat in this catalog and works with OpenAI-compatible client libraries.
glm-4.7-flash is listed with free API access on Cloudflare Workers AI, subject to the provider's quota and account policy.
The model ID shown in this catalog is @cf/zai-org/glm-4.7-flash.
The listed free tier limit is 10K neurons/day (shared). Limits can change per account tier, so confirm against the provider dashboard.
The listed context window is 131K tokens with up to 131K output tokens.
Efficient GLM model for fast reasoning, coding, and agent workflows
For API keys, setup steps, and provider-level limits, see the Cloudflare Workers AI provider page.