Tabbit
ResourcesBlogModels
Tabbit LogoTabbit

Tabbit — The AI Browser that Works for You

Topics

  • AI Browser Resources
  • Agentic Browser Resources
  • Browser Downloads and Install Guides
  • Browser Comparisons
  • AI Browser Alternatives
  • Browser Productivity Resources

Popular Guides

  • AI Browser
  • Agentic Browser Download
  • Best AI Browser 2026: Top 9 Tested & Ranked
  • AI Browser Download
  • Free AI Browser
  • Best AI Browser 2026
  • AI Browser Comparison 2026
  • AI Browser for Windows
  • AI Browser for Mac
  • Chrome Alternative 2026

Events

  • Tabbit Skill Competition
  • KPOP SBTI Fandom Personality Test
  • Tabbit Campus Creator Program
  • fifi's Picks: AI Skills for Research Papers
  • User Survey

About

  • Tabbit Blog
  • Press & Media
English
简体中文English
Prompts and workflows

GLM-5.2 · configuration

Using GLM-5.2 (zai-glm-5-2) Through Mistral: Third-Party Hosting Configuration and Pricing

Mistral now hosts GLM-5.2 as a third-party open model (Public Preview, model ID `zai-glm-5-2`, 1M context / 128k output, with no modifications), so it can be accessed directly across the Mistral ecosystem (including Vibe CLI) using that ID, at $1.4 / $0.14 (cached input) / $4.4 (output) per million tokens..

Source not verifiedMistral hosted zai-glm-5-2 route

Prerequisites and inputs

  • Mistral model ID
  • billing tier
  • context workload
  • smoke call

Complete templates

Editorial adaptation:Using GLM-5.2 (zai-glm-5-2) Through Mistral: Third-Party Hosting Configuration and Pricing

Tabbit editorial adaptation; not the original source prompt
Use {{HOSTED_MODEL_ID}} at {{BILLING_TIER}} for {{SHORT_TASK}} and {{LONG_DOCUMENT}}; log {{INPUT_TOKENS}}, {{CACHE_TOKENS}}, {{OUTPUT_TOKENS}}, and {{ACTUAL_COST}}.

Replace before running: {{HOSTED_MODEL_ID}}, {{BILLING_TIER}}, {{SHORT_TASK}}, {{LONG_DOCUMENT}}, {{INPUT_TOKENS}}, {{CACHE_TOKENS}}, {{OUTPUT_TOKENS}}, {{ACTUAL_COST}}

In Mistral Public Preview, pin zai-glm-5-2, the billing tier, and the 1M-context option. Run one short task and one long-document task, recording input, cached, output tokens, and actual cost. If the hosted price or snapshot changes, mark the old data and recheck; do not present hosted rates as Z.AI pricing.

Read the source research notes

One-sentence takeaway

Mistral now hosts GLM-5.2 as a third-party open model (Public Preview, model ID zai-glm-5-2, 1M context / 128k output, with no modifications), so it can be accessed directly across the Mistral ecosystem (including Vibe CLI) using that ID, at $1.4 / $0.14 (cached input) / $4.4 (output) per million tokens.

Use cases

  • Suitable tasks: scenarios where development is already happening in the Mistral ecosystem (Mistral API, Vibe CLI, Mistral Code, and so on) and GLM-5.2 needs to be switched in or added directly; integrations that require function calling, structured outputs, Predicted Outputs, prefix caching, or batching.

  • Unsuitable tasks: scenarios requiring Z.ai-exclusive capabilities (GLM Coding Plan quotas, tool_stream streaming tool calls, GLM-5.2[1m] integration with Claude Code, and so on); scenarios with latency or regional requirements (third-party hosting nodes differ from the official ones).

  • Applicable model version: GLM-5.2 (v5.2, Mistral-side model ID zai-glm-5-2; a Reddit user verified that setting active_model = "zai-glm-5-2" in Vibe CLI's config.toml works).

  • Applicable clients, agents, or APIs: Mistral API (/v1/chat/completions, /v1/conversations, /v1/batch); Vibe CLI.

  • Recommended reasoning tier and parameters: No special parameters required; Mistral explicitly states that the model is "served without Mistral modifications," and its thinking behavior is consistent with the official version.

Ready-to-use content

Specifications and pricing (official page)

ItemValue
Model IDzai-glm-5-2
StatusPublic Preview (third-party open model)
Context1M tokens
Maximum output128k tokens
Input price$1.4 / M tokens
Cached input price$0.14 / M tokens
Output price$4.4 / M tokens
Supported featuresChat Completions, Function Calling, Structured Outputs, Predicted Outputs, Prefix, Batching

Vibe CLI integration (r/MistralAI user-tested configuration)

Set the following in Vibe CLI's config.toml:

active_model = "zai-glm-5-2"

(Source: the r/MistralAI post "Will we get GLM-5.2 in Vibe CLI?"; a user reported, "this worked for me." The official documentation does not directly provide this config.toml snippet; it is only a community-verified configuration.)

Notes and limitations

  • This page reflects Mistral's hosting perspective: it is intended for developers already using the Mistral ecosystem; prices and capabilities from the official Z.ai API (such as GLM Coding Plan quota multipliers and tool_stream) cannot be compared or substituted directly.

  • A Reddit discussion (r/MistralAI 1vm88ur) mentioned that the model "seems to be available through Mistral hosting," but it may not appear in Vibe CLI's default list and must be configured manually—the tested configuration is shown above.

Source and dates

Mistral AI official documentation (docs.mistral.ai) · Source date: 2026-08-06 · Edited: 2026-09-20

Read the original source
Variable checklist

Still to replace: 8

{{HOSTED_MODEL_ID}}{{BILLING_TIER}}{{SHORT_TASK}}{{LONG_DOCUMENT}}{{INPUT_TOKENS}}{{CACHE_TOKENS}}{{OUTPUT_TOKENS}}{{ACTUAL_COST}}

Related prompts

GLM-5.2 Official Documentation: Overview and API Quick Start (docs.z.ai)Official Configuration Guide for Migrating from GLM-5.1 / GLM-5 / GLM-4.x to GLM-5.2GLM-5.2 Thinking Mode Configuration: Default Thinking / Interleaved Thinking / Preserved Thinking / Turn-level Thinking (Official)GLM-5.2 Role-Playing (RP) System Prompt: Evening-Truth Complete Dark-Version Prompt

Related reviews

NIST CAISI's Independent Capability Assessment of Z.ai GLM-5.2Semgrep IDOR Benchmark: GLM-5.2 Results with a Prompt-Only Setup in Security Code AuditingGLM-5.2 Official Release Notes and Complete Benchmark Table (Z.ai Blog)Reddit Blind Code Review: GLM-5.2's Production-Readiness Score and Multi-Judge Recheck

Read the full analysis

Overview · English

GLM-5.2: What It Is, What It Costs, and Where It Fits

A sourced GLM-5.2 overview covering the June 2026 release, 1M context, open-weight deployment, API pricing boundaries, coding evidence and a safer pilot path.

GLM-5.2

Use GLM-5.2 in Tabbit

Run this guide in the environment listed above. Downloading does not transfer the template or establish model availability for your account.