Tabbit
ResourcesBlogModels
Tabbit LogoTabbit

Tabbit — The AI Browser that Works for You

Topics

  • AI Browser Resources
  • Agentic Browser Resources
  • Browser Downloads and Install Guides
  • Browser Comparisons
  • AI Browser Alternatives
  • Browser Productivity Resources

Popular Guides

  • AI Browser
  • Agentic Browser Download
  • Best AI Browser 2026: Top 9 Tested & Ranked
  • AI Browser Download
  • Free AI Browser
  • Best AI Browser 2026
  • AI Browser Comparison 2026
  • AI Browser for Windows
  • AI Browser for Mac
  • Chrome Alternative 2026

Events

  • Tabbit Skill Competition
  • KPOP SBTI Fandom Personality Test
  • Tabbit Campus Creator Program
  • fifi's Picks: AI Skills for Research Papers
  • User Survey

About

  • Tabbit Blog
  • Press & Media
English
简体中文English
Reviews and evidence

GLM-5.2 · Community source · Personal experience

r/PoeAI Discussion: Did GLM-5.2 Suddenly “Get Dumber”? — Community Evidence of Quality Differences Across Hosting Platforms

u/Friendly-Play-1953 posted in r/PoeAI that GLM-5.2 on Poe had consistently worked very well for creative writing, but over the past two or three days it had “lost half its IQ overnight”—even forgetting basic details across posts (the sa.

Unverified: the original source could not be rechecked. Historical figures below are not current verified results.

Community sourcePersonal experienceEdited 2026-09-20

Test conditions

Model/version
GLM-5.2; source title “r/PoeAI Discussion: Did GLM-5.2 Suddenly “Get Dumber”? — Community Evidence of Quality Differences Across Hosting Platforms”, with no cross-version merge.
Task/harness
Summary of key content u/Friendly-Play-1953 posted in r/PoeAI that GLM-5.2 on Poe had consistently worked very well for creative writing, but over the past two or three days it had “lost half its IQ overnight”—even forge The complete task set, runtime parameters, and review procedure are not fully public.
Sample/date
Source note reviewed 2026-09-20; undisclosed sample count, repeats, and raw logs remain unknown.

Key data and applicable tasks

Summary of key content

u/Friendly-Play-1953 posted in r/PoeAI that GLM-5.2 on Poe had consistently worked very well for creative writing, but over the past two or three days it had “lost half its IQ overnight”—even forgetting basic details across posts (the same character was blond in one post and had brown hair in the next), and feeling “like a model from 18 months ago.” The prompts and character cards had not changed at all.

Key information from the comments

  • u/jack9556: Asked which provider was being used; said they had not noticed any problems with the Jas-hosted version, and shared Jas's Poe bot link: poe.com/GLM-5.2-JasInTech.

  • u/Friendly-Play-1953: Said they were using the official Z.ai-hosted version (custom bots can only select official models).

  • u/TheIncarnated (key comment):

"GLM ramped up to release 5.3, they definitely made it dumber on Poe, and removed the ability for it to run tools. I had… This is a necessary excerpt; read the original source for full context. (While preparing for the 5.3 release, GLM was clearly made dumber on Poe, and its tool-use capability was removed; switching to OpenRouter restored normal operation.)

  • Emerging community consensus: The same model performs differently across hosting providers (official Poe / Jas / OpenRouter), and quality fluctuations are strongly associated with platform-side configuration or downgrades, rather than necessarily reflecting degradation in the model itself.

Evidence highlights and scope

  • What the conclusion points to: Community observations about whether GLM-5.2 has degraded suggest that the hosting channel is the decisive variable (the official Poe endpoint was reportedly downgraded, while Jas and OpenRouter worked normally). There is also speculation (unverified) that the model was intentionally or unintentionally downgraded on the eve of the 5.3 release.

  • Limitations: Everything is based on personal experience and speculation; there is no quantitative data (no A/B scores using the same prompts). The time frame is mid-August 2026, around the release of GLM-5.3 (August 14), so the observations may be time-sensitive.

  • Use case: A reminder for model selection: when using GLM-5.2 through an aggregator or free endpoint, pay attention to differences between hosting providers and version routing. If quality suddenly drops, switch providers to verify the issue before dismissing the model outright.

Key quotes from the original

"It feels like something changed somewhere very quickly, has anyone else noticed this? Same prompts, same character card… This is a necessary excerpt; read the original source for full context.

"GLM ramped up to release 5.3, they definitely made it dumber on Poe, and removed the ability for it to run tools. I had… This is a necessary excerpt; read the original source for full context.

What this supports

  • Supports the source-specific observation in “r/PoeAI Discussion: Did GLM-5.2 Suddenly “Get Dumber”? — Community Evidence of Quality Differences Across Hosting Platforms”: Summary of key content u/Friendly-Play-1953 posted in r/PoeAI that GLM-5.2 on Poe had consistently worked very well for creative writing, but over the past two or three days it had “lost hal

What this does not support

  • Does not support a general capability or production-rate claim from “r/PoeAI Discussion: Did GLM-5.2 Suddenly “Get Dumber”? — Community Evidence of Quality Differences Across Hosting Platforms”; the source lacks a controlled task set, provider snapshot, and repeated independent retest.

Method, limits, and reproduction

The figures, task set, reasoning tier, and client conditions apply only to the listed source and collection snapshot. Different versions, harnesses, or providers must not be compared directly; undisclosed parameters remain unknown.

For a reproduction, fix the model version, provider or client, reasoning tier, tools, task-set version, sample count, and collection date, and record failures, retries, and human corrections. Full steps are in the source notes below.

Original source

Reddit r/PoeAI · u/Friendly-Play-1953 (original post); u/TheIncarnated, u/jack9556, and others (comments) · Original publication date Unknown · Site edit date 2026-09-20

Open original source

GLM-5.2

Compare GLM-5.2 in Tabbit

Download the Tabbit client to check model access

Read the full analysis

Overview · English

GLM-5.2: What It Is, What It Costs, and Where It Fits

A sourced GLM-5.2 overview covering the June 2026 release, 1M context, open-weight deployment, API pricing boundaries, coding evidence and a safer pilot path.

Related reviews

Reddit Blind Code Review: GLM-5.2's Production-Readiness Score and Multi-Judge RecheckA Reddit VPS Manager blind review compared five models under one specification; Qwen 3.7 Plus first used a fixed 25-point rubric, followed by GPT Codex and Gemini 3.1 Pro rechecks; the sample is one project.GLM-5.2 Official Release Notes and Complete Benchmark Table (Z.ai Blog)Z.ai’s 2026-06-16 release positions GLM-5.2 as a 1M-context long-horizon flagship and reports 81.0 on Terminal-Bench 2.1 and 62.1 on SWE-Bench Pro; it also discloses training-stage reward-hacking risk.NIST CAISI's Independent Capability Assessment of Z.ai GLM-5.2NIST CAISI published its assessment on 2026-07-17 after completing it on 2026-07-08: GLM-5.2 was similar to GPT-5.2 overall and Opus 4.6 on cyber capability, while safeguards were mixed for agentic exploits and biological questions.Semgrep IDOR Benchmark: GLM-5.2 Results with a Prompt-Only Setup in Security Code AuditingSemgrep’s 2026-06-22 IDOR benchmark held dataset, evaluation, and prompt constant: GLM-5.2 reached 39% F1 in a Pydantic AI prompt-only harness at about $0.17 per vulnerability; this is not a general cyber score.GLM-5.2 Official Documentation: Overview and API Quick Start (docs.z.ai)The official standard integration configuration for GLM-5.2 is: model name `glm-5.2`, a 1M context window / 128K maximum output, `thinking.type: enabled` + `reasoning_effort: max`, and `temperature: 1.0`. You can copy the curl / Python examples directly to make your first call and review the typical use cases identified by the official documentation..GLM-5.2 Thinking Mode Configuration: Default Thinking / Interleaved Thinking / Preserved Thinking / Turn-level Thinking (Official)The official documentation states that thinking is enabled by default for GLM-5.2 (as with GLM-5.1/5/4.7), and provides four thinking modes: default thinking, interleaved thinking (thinking between tool calls), preserved thinking (retaining reasoning content across turns with `clear_thinking: false`), and turn-level thinking (an independent switch for each turn). It also highlights a key constraint for Agent integrations: historical `reasoning_content` must be returned unchanged..Official Configuration Guide for Migrating from GLM-5.1 / GLM-5 / GLM-4.x to GLM-5.2The official GLM-5.2 migration checklist and parameter configuration: change the model ID to `glm-5.2`; use the default `temperature` of 1.0 or default `top_p` of 0.95 (tune only one of the two); enable thinking by default; use `high` or `max` for `reasoning_effort`; configure streaming and streaming tool calls (`stream=true` + `tool_stream=true`) as specified by the official guidance; and use the included Python migration example directly..Using GLM-5.2 (zai-glm-5-2) Through Mistral: Third-Party Hosting Configuration and PricingMistral now hosts GLM-5.2 as a third-party open model (Public Preview, model ID `zai-glm-5-2`, 1M context / 128k output, with no modifications), so it can be accessed directly across the Mistral ecosystem (including Vibe CLI) using that ID, at $1.4 / $0.14 (cached input) / $4.4 (output) per million tokens..