Tabbit
ResourcesBlogModels
Tabbit LogoTabbit

Tabbit — The AI Browser that Works for You

Topics

  • AI Browser Resources
  • Agentic Browser Resources
  • Browser Downloads and Install Guides
  • Browser Comparisons
  • AI Browser Alternatives
  • Browser Productivity Resources

Popular Guides

  • AI Browser
  • Agentic Browser Download
  • Best AI Browser 2026: Top 9 Tested & Ranked
  • AI Browser Download
  • Free AI Browser
  • Best AI Browser 2026
  • AI Browser Comparison 2026
  • AI Browser for Windows
  • AI Browser for Mac
  • Chrome Alternative 2026

Events

  • Tabbit Skill Competition
  • KPOP SBTI Fandom Personality Test
  • Tabbit Campus Creator Program
  • fifi's Picks: AI Skills for Research Papers
  • User Survey

About

  • Tabbit Blog
  • Press & Media
English
简体中文English
Reviews and evidence

GLM-5.2 · Community source · Personal experience

r/openrouter Discussion: GLM-5.2 Free Endpoint (Hosted by Decart): Availability and Limitations

u/FreeTruck7609 discovered a free GLM endpoint hosted by Decart on OpenRouter (launched "about a day ago"). It had reliability problems at first, then stabilized at close to 100% availability..

Unverified: the original source could not be rechecked. Historical figures below are not current verified results.

Community sourcePersonal experienceEdited 2026-09-20

Test conditions

Model/version
GLM-5.2; source title “r/openrouter Discussion: GLM-5.2 Free Endpoint (Hosted by Decart): Availability and Limitations”, with no cross-version merge.
Task/harness
Summary of key content u/FreeTruck7609 discovered a free GLM endpoint hosted by Decart on OpenRouter (launched "about a day ago"). It had reliability problems at first, then stabilized at close to 100% availability.。 The complete task set, runtime parameters, and review procedure are not fully public.
Sample/date
Source note reviewed 2026-09-20; undisclosed sample count, repeats, and raw logs remain unknown.

Key data and applicable tasks

Summary of key content

u/Free_Truck_7609 discovered a free GLM endpoint hosted by Decart on OpenRouter (launched "about a day ago"). It had reliability problems at first, then stabilized at close to 100% availability.

Important limitations revealed in the comments

  • Quantization: u/RepulsiveRaisin7 pointed out that the endpoint uses FP4 quantization ("Ok it is FP4, but still...").

  • No tool calling: u/DominikPlays: "No tool calling, so it's useless"; u/Ysn123987: "It does not allow any tool use; it's just a simple chatbot" — essentially unusable for agent or coding purposes.

  • Context: u/Ok_Philosophy_4031 pointed out that it has a 128k context window (far smaller than the official 1M).

  • Rate limits and stability: u/CraftyCheeseburger ran into rate limiting on Decart's side; u/Urmomsmellgood: "now its gone" (the endpoint has disappeared or been taken offline).

  • Questions about the motivation: u/Adventurous_Bus_437 and u/Eyelbee questioned the motivation behind offering it for free (stress testing/data); u/fyndor said the free endpoint was "kicked out just as it got started, so it can't do anything serious."

Evidence highlights and scope

  • What the conclusion points to: The free GLM-5.2 endpoint on OpenRouter (Decart) is a restricted version — FP4 quantization, a 128k context window, and no tool calling. It is suitable for single-turn casual chat or small tests, but not for agent or coding workflows; the endpoint is also unstable (rate limiting and takedown).

  • Limitations: Everything comes from community reports and personal experience, with no official confirmation; free endpoints can change at any time. Its capabilities differ dramatically from the official API/Coding Plan (1M context, tool support, and tool_stream), so the free endpoint's performance must not be treated as GLM-5.2's true capability.

  • Use case: A reminder of where to try GLM-5.2 for free: use the official API or GLM Coding Plan for full capabilities. It also reinforces the shared conclusion of document 05 (local deployment with Colibri) and document 06 (differences between PoeAI-hosted versions): the channel determines the experience.

Key quotes from the original

"No tool calling, so it's useless."

"It did not allow any tool use. Just a simple chatbot like on www."

"Free endpoints are relatively useless, unless you want to do some simple single prompt test."

What this supports

  • Supports the source-specific observation in “r/openrouter Discussion: GLM-5.2 Free Endpoint (Hosted by Decart): Availability and Limitations”: Summary of key content u/FreeTruck7609 discovered a free GLM endpoint hosted by Decart on OpenRouter (launched "about a day ago"). It had reliability problems at first, then stabilized at cl

What this does not support

  • Does not support a general capability or production-rate claim from “r/openrouter Discussion: GLM-5.2 Free Endpoint (Hosted by Decart): Availability and Limitations”; the source lacks a controlled task set, provider snapshot, and repeated independent retest.

Method, limits, and reproduction

The figures, task set, reasoning tier, and client conditions apply only to the listed source and collection snapshot. Different versions, harnesses, or providers must not be compared directly; undisclosed parameters remain unknown.

For a reproduction, fix the model version, provider or client, reasoning tier, tools, task-set version, sample count, and collection date, and record failures, retries, and human corrections. Full steps are in the source notes below.

Original source

Reddit r/openrouter · u/FreeTruck7609 (main post); u/RepulsiveRaisin7, u/DominikPlays, u/Ysn123987, u/fyndor, and others (comments) · Original publication date Unknown · Site edit date 2026-09-20

Open original source

GLM-5.2

Compare GLM-5.2 in Tabbit

Download the Tabbit client to check model access

Read the full analysis

Overview · English

GLM-5.2: What It Is, What It Costs, and Where It Fits

A sourced GLM-5.2 overview covering the June 2026 release, 1M context, open-weight deployment, API pricing boundaries, coding evidence and a safer pilot path.

Related reviews

Reddit Blind Code Review: GLM-5.2's Production-Readiness Score and Multi-Judge RecheckA Reddit VPS Manager blind review compared five models under one specification; Qwen 3.7 Plus first used a fixed 25-point rubric, followed by GPT Codex and Gemini 3.1 Pro rechecks; the sample is one project.GLM-5.2 Official Release Notes and Complete Benchmark Table (Z.ai Blog)Z.ai’s 2026-06-16 release positions GLM-5.2 as a 1M-context long-horizon flagship and reports 81.0 on Terminal-Bench 2.1 and 62.1 on SWE-Bench Pro; it also discloses training-stage reward-hacking risk.NIST CAISI's Independent Capability Assessment of Z.ai GLM-5.2NIST CAISI published its assessment on 2026-07-17 after completing it on 2026-07-08: GLM-5.2 was similar to GPT-5.2 overall and Opus 4.6 on cyber capability, while safeguards were mixed for agentic exploits and biological questions.Semgrep IDOR Benchmark: GLM-5.2 Results with a Prompt-Only Setup in Security Code AuditingSemgrep’s 2026-06-22 IDOR benchmark held dataset, evaluation, and prompt constant: GLM-5.2 reached 39% F1 in a Pydantic AI prompt-only harness at about $0.17 per vulnerability; this is not a general cyber score.GLM-5.2 Official Documentation: Overview and API Quick Start (docs.z.ai)The official standard integration configuration for GLM-5.2 is: model name `glm-5.2`, a 1M context window / 128K maximum output, `thinking.type: enabled` + `reasoning_effort: max`, and `temperature: 1.0`. You can copy the curl / Python examples directly to make your first call and review the typical use cases identified by the official documentation..GLM-5.2 Thinking Mode Configuration: Default Thinking / Interleaved Thinking / Preserved Thinking / Turn-level Thinking (Official)The official documentation states that thinking is enabled by default for GLM-5.2 (as with GLM-5.1/5/4.7), and provides four thinking modes: default thinking, interleaved thinking (thinking between tool calls), preserved thinking (retaining reasoning content across turns with `clear_thinking: false`), and turn-level thinking (an independent switch for each turn). It also highlights a key constraint for Agent integrations: historical `reasoning_content` must be returned unchanged..Official Configuration Guide for Migrating from GLM-5.1 / GLM-5 / GLM-4.x to GLM-5.2The official GLM-5.2 migration checklist and parameter configuration: change the model ID to `glm-5.2`; use the default `temperature` of 1.0 or default `top_p` of 0.95 (tune only one of the two); enable thinking by default; use `high` or `max` for `reasoning_effort`; configure streaming and streaming tool calls (`stream=true` + `tool_stream=true`) as specified by the official guidance; and use the included Python migration example directly..Using GLM-5.2 (zai-glm-5-2) Through Mistral: Third-Party Hosting Configuration and PricingMistral now hosts GLM-5.2 as a third-party open model (Public Preview, model ID `zai-glm-5-2`, 1M context / 128k output, with no modifications), so it can be accessed directly across the Mistral ecosystem (including Vibe CLI) using that ID, at $1.4 / $0.14 (cached input) / $4.4 (output) per million tokens..