UnverifiedZ.ai official developer documentation (docs.z.ai)
GLM-5.2 Official Documentation: Overview and API Quick Start (docs.z.ai)
The official standard integration configuration for GLM-5.2 is: model name `glm-5.2`, a 1M context window / 128K maximum output, `thinking.type: enabled` + `reasoning_effort: max`, and `temperature: 1.0`. You can copy the curl / Python examples directly to make your first call and review the typical use cases identified by the official documentation..
Content typeConfiguration
InputsAPI key, model ID, request body, smoke output
StatusZ.AI API with GLM-5.2
API configurationreasoningAgent
UnverifiedZ.ai official developer documentation (docs.z.ai, Get Started / Migrate)
Official Configuration Guide for Migrating from GLM-5.1 / GLM-5 / GLM-4.x to GLM-5.2
The official GLM-5.2 migration checklist and parameter configuration: change the model ID to `glm-5.2`; use the default `temperature` of 1.0 or default `top_p` of 0.95 (tune only one of the two); enable thinking by default; use `high` or `max` for `reasoning_effort`; configure streaming and streaming tool calls (`stream=true` + `tool_stream=true`) as specified by the official guidance; and use the included Python migration example directly..
Content typeConfiguration
Inputsold request, new model ID, streaming flags, regression task
StatusZ.AI API migration from GLM-5.1/5/4.x
API configurationAgentCoding
UnverifiedZ.ai Official Developer Documentation (docs.z.ai, Capabilities / Thinking Mode)
GLM-5.2 Thinking Mode Configuration: Default Thinking / Interleaved Thinking / Preserved Thinking / Turn-level Thinking (Official)
The official documentation states that thinking is enabled by default for GLM-5.2 (as with GLM-5.1/5/4.7), and provides four thinking modes: default thinking, interleaved thinking (thinking between tool calls), preserved thinking (retaining reasoning content across turns with `clear_thinking: false`), and turn-level thinking (an independent switch for each turn). It also highlights a key constraint for Agent integrations: historical `reasoning_content` must be returned unchanged..
Content typeConfiguration
Inputsreasoning_content history, clear_thinking, turn policy, tool result
StatusGLM-5.2 tool-calling conversation
reasoningAgentAPI configuration
UnverifiedMistral AI official documentation (docs.mistral.ai)
Using GLM-5.2 (zai-glm-5-2) Through Mistral: Third-Party Hosting Configuration and Pricing
Mistral now hosts GLM-5.2 as a third-party open model (Public Preview, model ID `zai-glm-5-2`, 1M context / 128k output, with no modifications), so it can be accessed directly across the Mistral ecosystem (including Vibe CLI) using that ID, at $1.4 / $0.14 (cached input) / $4.4 (output) per million tokens..
Content typeConfiguration
InputsMistral model ID, billing tier, context workload, smoke call
StatusMistral hosted zai-glm-5-2 route
API configurationcostAgent
UnverifiedX.com (Twitter), @arena (official Arena.ai account)
Arena.ai Frontend Coding Head-to-Head: 10 Single-shot Generation Examples Comparing GLM-5.2 (Max) and Claude Opus 4.8 (Thinking)
Arena.ai gave the same set of frontend coding prompts to GLM-5.2 (Max) and Claude Opus 4.8 (Thinking) for single-shot generation and recorded the results side by side. The comparison shows that GLM-5.2 (Max) ranks higher than Opus 4.8 (Thinking) in Code Arena: Frontend community voting, and that these prompts can be copied directly for testing frontend visualizations or React site generation..
Content typeWorkflow
Inputstask objective, source material, output format, acceptance check
StatusModel-compatible prompt harness
Agent
Unverifiedrentry.org (the author's self-hosted prompt library, recommended by the SillyTavernAI community)
GLM-5.2 Role-Playing (RP) System Prompt: Evening-Truth Complete Dark-Version Prompt
A complete role-playing system prompt tuned specifically for GLM 5.2 that can be pasted directly into frontends such as SillyTavern (including "dark version" content rules and post-history rules), accompanied by the author's tested sampling parameters and assessment of GLM 5.2's capabilities in character consistency, initiative, and writing style..
Content typeWorkflow
Inputstask objective, source material, output format, acceptance check
StatusModel-compatible prompt harness
Agent
UnverifiedX.com (Twitter), @CossackWang (C0ss4ck, a security-focused developer)
X Field Test: GLM-5.2 (Code Audit) + GPT-5.5 (Context Collection) Firmware Source Code Audit Workflow
A real, successfully run “multi-model division-of-labor audit workflow”: GLM-5.2 handled code auditing while GPT-5.5 handled context collection, scanning firmware source code and finding high- and critical-severity vulnerabilities at a cost of about ¥1,000. The author later reproduced an almost identical vulnerability set with dsh minimal mode + DeepSeek-V4-Pro-0813 (only two were missing), bringing the cost down to ¥4. This provides a reusable reference for model selection and cost planning for “LLM-assisted source code auditing.”.
Content typeWorkflow
Inputstask objective, source material, output format, acceptance check
StatusModel-compatible prompt harness
Agent