Skip to main content

Try GLM 5.2 in the Workbench

Run this model interactively, tune parameters, and compare outputs.
Model ID: zai-org-glm-5-2 zai-org/GLM-5.2 is a large Mixture-of-Experts language model from Z AI built for long-horizon coding and agentic tasks. It introduces a 1M-token context window and multi-effort coding capabilities that improve sustained performance across long, multi-step workflows. Its IndexShare architecture and improved MTP layer reduce per-token FLOPs while increasing speculative decoding lengths, boosting efficiency without sacrificing quality. GLM-5.2 supports function calling and is well suited for agentic engineering, code assistance, and tasks that require reasoning over very long contexts.
MetricValue
Parameter Count743 billion
Mixture of ExpertsYes
Context Length1,040,000 tokens
MultilingualYes
Quantized*Unknown
*Quantization is specific to the inference provider and the model may be offered with different quantization levels by other providers.

Example request

Use the Workbench as a request builder: configure parameters for this model in the UI, then open the API tab to copy the exact cURL or Python call.
curl -X POST https://hub.oxen.ai/api/ai/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $OXEN_API_KEY" \
  -d '{
  "model": "zai-org-glm-5-2",
  "messages": [
    {
      "role": "user",
      "content": "Hello, what can you do?"
    }
  ]
}'

Fetch model details

The models endpoint returns the full model object, including its json_request_schema.
curl -H "Authorization: Bearer $OXEN_API_KEY" https://hub.oxen.ai/api/ai/models/zai-org-glm-5-2

Request parameters

This model follows the standard OpenAI chat completions request body. See the chat completions reference for the full parameter list.