Tabbit
ResourcesBlogModels
Tabbit LogoTabbit

Tabbit — The AI Browser that Works for You

Topics

  • AI Browser Resources
  • Agentic Browser Resources
  • Browser Downloads and Install Guides
  • Browser Comparisons
  • AI Browser Alternatives
  • Browser Productivity Resources

Popular Guides

  • AI Browser
  • Agentic Browser Download
  • Best AI Browser 2026: Top 9 Tested & Ranked
  • AI Browser Download
  • Free AI Browser
  • Best AI Browser 2026
  • AI Browser Comparison 2026
  • AI Browser for Windows
  • AI Browser for Mac
  • Chrome Alternative 2026

Events

  • Tabbit Skill Competition
  • KPOP SBTI Fandom Personality Test
  • Tabbit Campus Creator Program
  • fifi's Picks: AI Skills for Research Papers
  • User Survey

About

  • Tabbit Blog
  • Press & Media
English
简体中文English
Reviews and evidence

Ox Alpha · Community source · Editorial analysis

Jonathan Turner: Ox Alpha Fingerprint Comparisons and Identity Boundaries

The article compares Ox Alpha with public GLM, Gemini, DeepSeek, Kimi, and MiMo using tokenizer behavior, video-token budgets, and server errors, strongly pointing to a GLM-family model on a Z.ai serving stack while explicitly stopping short of naming a product or developer.

Unverified: the original source could not be rechecked. Historical figures below are not current verified results.

Community sourceEditorial analysisEdited 2026-09-20

Test conditions

Model/version
Ox Alpha; exact snapshot follows the source
Source date
2026-08-25
Method/client
Source-specific public post; client and provider conditions follow the source
Review state
Dynamic source not reopened on 2026-09-20; values remain unverified

Key data and applicable tasks

One-sentence takeaway

The article compares Ox Alpha with public GLM, Gemini, DeepSeek, Kimi, and MiMo using tokenizer behavior, video-token budgets, and server errors, strongly pointing to a GLM-family model on a Z.ai serving stack while explicitly stopping short of naming a product or developer.

Test environment

  • Object: OpenRouter stealth/ox-alpha.

  • Input specification: 1,048,576-token context, text/image/video input, and text output; the public model card lists an anonymous Stealth provider.

  • Data sources: Measurements attributed to YFarmX, unclecode, Ben Davis, and Security Kid; the author explicitly says they were not rerun in this article.

  • Comparisons: Public GLM-5.3, GLM-5V-Turbo, Gemini, DeepSeek, Kimi, and MiMo.

Results data

MeasurementOx Alpha comparisonInterpretation boundary
Prompt tokens on 50 strings50/50 matches GLM-5.3; Gemini 11/50, DeepSeek 9/50, Kimi 9/50, MiMo 18/50Inherited community measurement; tokenizer match is not weight or product match
Video-token budgetFour clips match GLM-5V-Turbo at 296, 296, 884, and 1,064 respectivelyn=4; encoder/budget similarity does not prove the same model
Server errorsThe article says Java/Spring errors and business codes 1210/1214 are byte-identical to public GLM-5.3 errorsShared serving stack is a hosting clue, not proof of weight ownership
Product-card comparisonOx is 1M multimodal; public GLM-5.3 is 1M text-only, while GLM-5V-Turbo is 200K vision“Same family” and “same public SKU” must remain separate

Conclusion

The article's value is the reproducible numbers and explicit boundaries, not the claim that Ox Alpha is a named model. The stronger current conclusion is a GLM-style tokenizer/video encoder and a Z.ai PaaS hosting clue; the exact checkpoint, operator, and developer remain undisclosed.

Reproduction steps

  1. Fix the 50 strings and tokenization method, query Ox Alpha and public models, and save prompt_tokens.

  2. Fix the four video clips, dimensions, and frame rates; record visual-token or server-reported budgets and expand to far more than four samples.

  3. Save raw HTTP errors, response headers, timestamps, and model versions; compare error bytes rather than screenshot similarity.

  4. Score tokenizer, video budget, server stack, and product card separately. Update the identity conclusion only after a first-party claim, weight match, or larger-sample evidence appears.

Limitations

  • The article says the measurements came from other people's work and were not rerun by its author.

  • Public model cards, anonymous-provider data, and community fingerprints involve inference; “GLM family” must not be rewritten as “GLM-5.3 Flash confirmed.”

  • The article's state is as of 2026-08-25; the server, model card, and anonymous preview can change.

What this supports

  • The article compares Ox Alpha with public GLM, Gemini, DeepSeek, Kimi, and MiMo using tokenizer behavior, video-token budgets, and server errors, strongly pointing to a GLM-family model on a Z.ai serving stack while explicitly stopping short of naming a product or developer.

What this does not support

  • “Jonathan Turner: Ox Alpha Fingerprint Comparisons and Identity Boundaries” lacks a unified task set, complete method, or version isolation (source date the source date); its observation cannot be generalized to universal capability or a current fact.

Method, limits, and reproduction

The figures, task set, reasoning tier, and client conditions apply only to the listed source and collection snapshot. Different versions, harnesses, or providers must not be compared directly; undisclosed parameters remain unknown.

For a reproduction, fix the model version, provider or client, reasoning tier, tools, task-set version, sample count, and collection date, and record failures, retries, and human corrections. Full steps are in the source notes below.

Original source

X Article · Jonathan Turner (@JonathanNTurner), with David and the house · Original publication date 2026-08-25 · Site edit date 2026-09-20

Open original source

Ox Alpha

Compare Ox Alpha in Tabbit

Download the Tabbit client to check model access

Read the full analysis

Overview · English

Ox Alpha Explained: From Stealth Preview to GLM-5.3-Flash

Ox Alpha was the anonymous name for Z.ai GLM-5.3-Flash. Here are the verified specs, access boundaries, preview timeline and safe testing decision.

Related reviews

Cline Test: Ox Alpha and Fable Both Fixed a Real Repository Bug, with About 3x Less OutputCline says Ox Alpha and Fable both correctly fixed one real bug in its repository; Ox Alpha used about three times fewer output tokens and repeated less reasoning. This is an efficiency signal worth retesting, not a general capability ranking.OpenCode Go Entry: Free Period and Load FeedbackOpenCode announced Ox Alpha on OpenCode Go for six days of near-unlimited free use outside Go usage; public replies also report mid-run stops, roughly 20 tokens/s, and overload, so convenience and service stability must be evaluated separately.OpenRouter Record: Ox Alpha Specs, Availability, and Anonymous ProviderOpenRouter describes Ox Alpha as a reasoning model for coding, sustained agentic work, and production workloads, listing free access, a 1,048,576-token context, up to 131,072 output tokens, and text/image/video input; the provider is only named Stealth, the developer remains anonymous, and the page publishes no reproducible capability benchmark.OpenCode Official Observation: 26T Ox Alpha Tokens in Four DaysOpenCode reports that Ox Alpha processed 26T tokens in four days, showing heavy real-world use of the preview but saying nothing by itself about model quality, individual quotas, or availability.OpenCode 1.18.21: Automatic Retries for Ox Alpha StopsOpenCode recommends upgrading to 1.18.21 when Ox Alpha produces network errors; the release automatically retries unknown stops, improving client resilience but not repairing provider outages or rate limits.Ox Alpha Same-Session Typecheck Audit and Custom-Instruction WritebackWhen Ox Alpha continues to report errors after repeated typechecks in OpenCode, ask it to audit the errors in the same session, then write verified repair principles back into custom instructions to create a project-specific feedback loopOx Alpha + Three.js Nan Lian Garden: A Single-Prompt 3D Scene Example“Build a 3D version of Nan Lian Garden with Three.js” is a short task suitable for checking spatial layout, rendering stability, and the debugging loop, but the original post publishes only a task summary, not the full prompt.SVG structure smoke testA one-line SVG prompt for a dragon riding a bicycle is a quick way to check Ox Alpha's handling of structural relationships, physical plausibility, and executable SVG code.