The OpenRouter page provides a third-party view beyond the official figures: LongCat-2.0 is listed at $0.30/$1.20 per 1M tokens (with a 60% discount at collection time), while the actual weighted transaction price for input was only $0.03872/M (88.9% cache-hit rate); throughput was P50 29 tok/s, three-day availability 99.93%, and tool-call error rate 0.90%, with real traffic mainly coming from Hermes Agent (7.77B tokens) and Claude Code (3.31B tokens).
Model ID: meituan/longcat-2.0 (page version 20260720); 48B active / 1.6T total parameters MoE; 1M context; text only.
List price: IN $0.30 / OUT $1.20 per 1M (60% off for a limited time); launch price $0.30/$1.20 (confirmed the same day by r/AIToolsPerformance).
Actual transactions: weighted-average input $0.03872/M, output $1.20/M (caching and discounts make the actual price far lower than the list price); provider AtlasCloud, cache-hit rate 88.92%, token share 100% (single provider).
The price-history chart shows the effective input price fluctuating between $0 and $0.08/1M from July 24 to August 18.
Throughput: P50 29 tok/s (all-time average); P99 54 / P95 46 / P90 42 / P75 35 tok/s.
Latency: P50 1.82s; E2E P50 13.75s (P99 159.74s).
Tool-call error rate: 0.90% (AtlasCloud).
Uptime over 3 days: 100.00%; availability: 99.93%; OpenRouter's own availability: 99.85%.
| Benchmark | Score | Description |
|---|---|---|
| Coding Index | 45.3 | Better than 49% of comparison models |
| GPQA Diamond | 78.0% | Graduate-level scientific reasoning |
| HLE (Humanity's Last Exam) | 33.7% | Broad difficult problems |
| AA-LCR | 62.7% | Long-context reasoning |
| GDPval-AA | 26.5% | Economically valuable tasks |
| CritPt | 2.6% | Research-level physics reasoning |
| SciCode | 35.4% | Scientific-computing programming |
| AA-Omniscience Accuracy | 29.6% | Knowledge question-answering accuracy |
| AA-Omniscience Non-hallucination rate | 24.6% | Anti-hallucination ratio when the answer is not correct |
Hermes Agent — 7.77B tokens (open-source Agent from Nous Research)
Claude Code — 3.31B tokens
pi — 2.27B tokens
Halluna — 1.84B tokens
TokenTool — 1.49B tokens
The data is platform real-time telemetry and can be checked against the original URL at any time; however, discounts, availability, and benchmark scores change over time, so citations should include the collection date.
Difference from official pricing: OpenRouter's list price ($0.30/$1.20) and the official direct-connection price (¥5/¥20 per 1M at the original price) are not the same billing system, and cache-hit pricing on OpenRouter substantially dilutes the actual cost; they cannot be directly converted and compared.
Artificial Analysis's Coding Index 45.3 (better than 49% of models) is clearly below the impression created by the official SWE-bench Pro score of 59.5. The benchmark sets differ, and the third-party index does not rank LongCat particularly highly, which is an important correction when judging "which tasks it suits."
A single provider (AtlasCloud) means there is no multi-route redundancy; the 99.93% availability is a snapshot covering roughly the past three days.
The free endpoint meituan/longcat-2.0:free ($0.00/1M) also exists in the OpenRouter and Nous Portal directories and is an entry point for low-cost trials (see Prompt Directory 03 and 06).
LongCat 2.0