Cloudflare Workers AI logo

glm-4.7-flash Free API on Cloudflare Workers AI

Free API Verified
Catalog profile Reasoning provider catalog metadata

Efficient GLM model for fast reasoning, coding, and agent workflows

Free APIOpenAI compatibleReasoningTool callingJSON modeTextReasoning
API Details

Copy-ready connection details

Use these values in your client or SDK.
Base URL
https://api.cloudflare.com/client/v4/accounts/{account_id}/ai/run
Model ID
@cf/zai-org/glm-4.7-flash
API format Cloudflare Workers AI native run + OpenAI Chat Completions
Technical Details

glm-4.7-flash specifications

Provider catalog
Context window 131K
Max output 131K
Status Online
Knowledge cutoff 2025-04
Released Jan 19, 2026
Last updated Sep 8, 2026
Free listing since Jan 19, 2026
Input text, reasoning
Output text
Capabilities reasoning, tool calling, structured output
AI Recommendation

Should you use glm-4.7-flash?

glm-4.7-flash is listed for chat workloads and supports a 131K context window.

Use it when Cloudflare Workers AI's free tier is enough for evaluation, demos, or light production traffic.

Best for
  • Chat
Strengths & Weaknesses

Strengths

  • Long context window
  • Reasoning mode listed
  • Tool calling support
  • Structured JSON output

Watch outs

  • Free-tier rate limits apply
  • Vision support is not listed
Pricing

glm-4.7-flash pricing per 1M tokens

Free tier listed
Input $0.0605 per 1M tokens
Output $0.4 per 1M tokens
Access Available Cloudflare Workers AI
Rate limit 10K neurons/day (shared) provider policy
Typical Use Cases

glm-4.7-flash use cases

Chat

glm-4.7-flash is tagged for chat in this catalog and works with OpenAI-compatible client libraries.

FAQ

glm-4.7-flash free API FAQ

Is glm-4.7-flash free to use?

glm-4.7-flash is listed with free API access on Cloudflare Workers AI, subject to the provider's quota and account policy.

What is the glm-4.7-flash model ID?

The model ID shown in this catalog is @cf/zai-org/glm-4.7-flash.

What are the glm-4.7-flash free tier rate limits on Cloudflare Workers AI?

The listed free tier limit is 10K neurons/day (shared). Limits can change per account tier, so confirm against the provider dashboard.

What context window does glm-4.7-flash support?

The listed context window is 131K tokens with up to 131K output tokens.

More about glm-4.7-flash

Efficient GLM model for fast reasoning, coding, and agent workflows

For API keys, setup steps, and provider-level limits, see the Cloudflare Workers AI provider page.