Tabbit
ResourcesBlogModels
Tabbit LogoTabbit

Tabbit — The AI Browser that Works for You

Topics

  • AI Browser Resources
  • Agentic Browser Resources
  • Browser Downloads and Install Guides
  • Browser Comparisons
  • AI Browser Alternatives
  • Browser Productivity Resources

Popular Guides

  • AI Browser
  • Agentic Browser Download
  • Best AI Browser 2026: Top 9 Tested & Ranked
  • AI Browser Download
  • Free AI Browser
  • Best AI Browser 2026
  • AI Browser Comparison 2026
  • AI Browser for Windows
  • AI Browser for Mac
  • Chrome Alternative 2026

Events

  • Tabbit Skill Competition
  • KPOP SBTI Fandom Personality Test
  • Tabbit Campus Creator Program
  • fifi's Picks: AI Skills for Research Papers
  • User Survey

About

  • Tabbit Blog
  • Press & Media
English
简体中文English
Reviews and evidence

LongCat 2.0 · Official source · Platform telemetry

OpenRouter Channel Data: LongCat-2.0 Pricing, Measured Performance, and Third-Party Benchmarks (Artificial Analysis)

The OpenRouter page provides a third-party view beyond the official figures: LongCat-2.0 is listed at $0.30/$1.20 per 1M tokens (with a 60% discount at collection time), while the actual weighted transaction price for input was only $0.03872/M (88.9% cache-hit rate); throughput was P50 29 tok/s, three-day availability 99.93%, and tool-call error rate 0.90%, with real traffic mainly coming from Hermes Agent (7.77B tokens) and Claude Code (3.31B tokens).

Unverified: the original source could not be rechecked. Historical figures below are not current verified results.

Official sourcePlatform telemetryEdited 2026-09-20

Test conditions

Model/version
LongCat-2.0; source date: 2026-07-20.
Harness/task
Model and pricing; Model ID: `meituan/longcat-2.0` (page version 20260720); 48B active / 1.6T total parameters MoE; 1M context; text only.
Sample/gaps
Limitations noted: Artificial Analysis's Coding Index 45.3 (better than 49% of models) is clearly below the impression created by the official SWE-bench Pro score of 59.5. The benchmark sets differ, and the third-party index does not rank LongCat particularly highly, which is an important correction when judging "which tasks it suits."; A single provider (AtlasCloud) means there is no multi-route redundancy; the 99.93% availability is a snapshot covering roughly the past three days.

Key data and applicable tasks

One-sentence takeaway

The OpenRouter page provides a third-party view beyond the official figures: LongCat-2.0 is listed at $0.30/$1.20 per 1M tokens (with a 60% discount at collection time), while the actual weighted transaction price for input was only $0.03872/M (88.9% cache-hit rate); throughput was P50 29 tok/s, three-day availability 99.93%, and tool-call error rate 0.90%, with real traffic mainly coming from Hermes Agent (7.77B tokens) and Claude Code (3.31B tokens).

Key data (collected 2026-08-18)

Model and pricing

  • Model ID: meituan/longcat-2.0 (page version 20260720); 48B active / 1.6T total parameters MoE; 1M context; text only.

  • List price: IN $0.30 / OUT $1.20 per 1M (60% off for a limited time); launch price $0.30/$1.20 (confirmed the same day by r/AIToolsPerformance).

  • Actual transactions: weighted-average input $0.03872/M, output $1.20/M (caching and discounts make the actual price far lower than the list price); provider AtlasCloud, cache-hit rate 88.92%, token share 100% (single provider).

  • The price-history chart shows the effective input price fluctuating between $0 and $0.08/1M from July 24 to August 18.

Performance and availability

  • Throughput: P50 29 tok/s (all-time average); P99 54 / P95 46 / P90 42 / P75 35 tok/s.

  • Latency: P50 1.82s; E2E P50 13.75s (P99 159.74s).

  • Tool-call error rate: 0.90% (AtlasCloud).

  • Uptime over 3 days: 100.00%; availability: 99.93%; OpenRouter's own availability: 99.85%.

Third-party benchmarks (as displayed by Artificial Analysis)

BenchmarkScoreDescription
Coding Index45.3Better than 49% of comparison models
GPQA Diamond78.0%Graduate-level scientific reasoning
HLE (Humanity's Last Exam)33.7%Broad difficult problems
AA-LCR62.7%Long-context reasoning
GDPval-AA26.5%Economically valuable tasks
CritPt2.6%Research-level physics reasoning
SciCode35.4%Scientific-computing programming
AA-Omniscience Accuracy29.6%Knowledge question-answering accuracy
AA-Omniscience Non-hallucination rate24.6%Anti-hallucination ratio when the answer is not correct

Real-world application traffic (Apps ranking, reflecting production use)

  1. Hermes Agent — 7.77B tokens (open-source Agent from Nous Research)

  2. Claude Code — 3.31B tokens

  3. pi — 2.27B tokens

  4. Halluna — 1.84B tokens

  5. TokenTool — 1.49B tokens

Review and scope

  • The data is platform real-time telemetry and can be checked against the original URL at any time; however, discounts, availability, and benchmark scores change over time, so citations should include the collection date.

  • Difference from official pricing: OpenRouter's list price ($0.30/$1.20) and the official direct-connection price (¥5/¥20 per 1M at the original price) are not the same billing system, and cache-hit pricing on OpenRouter substantially dilutes the actual cost; they cannot be directly converted and compared.

  • Artificial Analysis's Coding Index 45.3 (better than 49% of models) is clearly below the impression created by the official SWE-bench Pro score of 59.5. The benchmark sets differ, and the third-party index does not rank LongCat particularly highly, which is an important correction when judging "which tasks it suits."

  • A single provider (AtlasCloud) means there is no multi-route redundancy; the 99.93% availability is a snapshot covering roughly the past three days.

  • The free endpoint meituan/longcat-2.0:free ($0.00/1M) also exists in the OpenRouter and Nous Portal directories and is an entry point for low-cost trials (see Prompt Directory 03 and 06).

What this supports

  • The data is platform real-time telemetry and can be checked against the original URL at any time; however, discounts, availability, and benchmark scores change over time, so citations should include the collection date.
  • Difference from official pricing: OpenRouter's list price ($0.30/$1.20) and the official direct-connection price (¥5/¥20 per 1M at the original price) are not the same billing system, and cache-hit pricing on OpenRouter substantially dilutes the actual cost; they cannot be directly converted and compared.

What this does not support

  • Artificial Analysis's Coding Index 45.3 (better than 49% of models) is clearly below the impression created by the official SWE-bench Pro score of 59.5. The benchmark sets differ, and the third-party index does not rank LongCat particularly highly, which is an important correction when judging "which tasks it suits."
  • A single provider (AtlasCloud) means there is no multi-route redundancy; the 99.93% availability is a snapshot covering roughly the past three days.

Method, limits, and reproduction

The figures, task set, reasoning tier, and client conditions apply only to the listed source and collection snapshot. Different versions, harnesses, or providers must not be compared directly; undisclosed parameters remain unknown.

For a reproduction, fix the model version, provider or client, reasoning tier, tools, task-set version, sample count, and collection date, and record failures, retries, and human corrections. Full steps are in the source notes below.

Original source

OpenRouter (third-party model routing platform) · OpenRouter platform telemetry + Artificial Analysis evaluation data · Original publication date 2026-07-20 · Site edit date 2026-09-20

Open original source

LongCat 2.0

Compare LongCat 2.0 in Tabbit

Download the Tabbit client to check model access

Read the full analysis

Overview · English

LongCat 2.0: what changed, where to use it, and what the price misses

LongCat 2.0 combines 1M context, open weights, and low provider pricing with real questions about tooling, data terms, and operational cost.

Related reviews

LongCat-2.0 Official Model Card: Specifications and Official Benchmarks (Including Comparison Tables with Gemini/GPT-5.5/Claude Opus)The official model card is the primary authoritative source for judging LongCat-2.0's suitable tasks: it scores 59.5 on SWE-bench Pro, ahead of GPT-5.5 (58.6) and Gemini 3.1 Pro (54.2), and reaches 70.8 on Terminal-Bench 2.1. However, it trails GPT-5.5 and Claude Opus 4.8 on several benchmarks including BrowseComp, GPQA, and IFEval—in short, it is strong at coding and agent tasks, but not a leader in retrieval and general reasoning.LongCat-2.0 Official Technical Blog: Architecture, Training on Domestic Compute, and Inference Deployment (Release Notes)The official technical blog provides the complete technical foundation for LongCat-2.0 (LSA sparse attention, N-gram Embedding, 6D parallel training on domestic compute, and prefill-decode disaggregated deployment), making it useful for assessing the model's intended long-context and Agent capabilities, as well as reproducing the official benchmarks and deployment path.eesel Independent Review: LongCat-2.0's Agent Reliability and Hard Blockers to Production DeploymentThis independent review separates LongCat-2.0 into two questions: "can the model complete Agent work?" and "can the product enter enterprise production?" Public user reports support it as an inexpensive, stable coding executor, but its context specifications, tool contract, and data-governance documentation are insufficient to pass a sensitive-data production review.AlphaSignal Deep Dive: Owl Alpha's True Identity and a Reality Check on LongCat-2.0's Official ClaimsThis deep-dive review, published the day after launch, establishes the most important background fact — the anonymous free model "Owl Alpha," which ran on OpenRouter for two months, was LongCat-2.0 (processing approximately 10 trillion tokens per month and ranking first on the Hermes Agent leaderboard) — while breaking down the official claims one by one: the narrow SWE-bench Pro win over GPT-5.5 was an official self-test, the weights were still only "planned" when the review was published, and the open-source fielLongCat-2.0 API Platform Quick Start (Official Quick Start + Chat Completions Reference + Pricing)The LongCat Claude Code guide configures a compatible endpoint and keeps the first task in a disposable worktree.LongCat-2.0 Chat Template and Tool-Calling Configuration (Official Hugging Face Model Card)The official model card’s chat template and tool-call examples are converted into a local inference configuration check.Claude Code Integration with LongCat-2.0 (Official Documentation)The official LongCat integration guide configures a named client and keeps the first run observable and reversible.Hermes Agent Integration with LongCat-2.0 (Official Documentation + Nous Portal Free Entry)The official LongCat guide “Hermes Agent Integration with LongCat-2.0 (Official Documentation + Nous Portal Free Entry)” configures a named client and keeps the first run observable and reversible.