Gemini 3.8 Flash

Gemini 3.8 Flash review navigator

Official benchmarks, independent analysis, and community reports about Gemini 3.8 Flash, clearly separated from Tabbit's own testing.

11 source-checked resourcesOfficial · Media · Community

Official

1 source-checked resources

Media

6 source-checked resources
MediaArtificial Analysis (official model pages, methodology, and release article; the official X account was used to discover and cross-check the release post)

Gemini 3.8 Flash: Artificial Analysis Intelligence, Speed, Pricing, and Latency

One-sentence takeaway Gemini 3.8 Flash's speed and latency vary substantially by reasoning tier: Artificial Analysis's current v4.3 pages record 285.9 tok/s and 17.36 seconds to the first answer token for high, 246.3 tok/s and 8.29 seconds for medium, and 265.。

MediaAI IQ (AIIQ, Liberated Software LLC)

Gemini 3.8 Flash: AI IQ Capability Benchmarks and Task Boundaries

One-sentence takeaway AIIQ's current model page gives Gemini 3.8 Flash an AI IQ of 125, ranked 14/131, and lists scores from 8 source benchmarks. Academic reasoning has 5/6 coverage, programmatic reasoning only 1/6, and reliability 2/7, while abstract reasonin。

MediaVals AI

Vals AI Finance Agent v2: Professional Finance Agent Benchmark for Gemini 3.8 Flash

One-sentence takeaway On the overall Vals AI Finance Agent v2 tasks, Gemini 3.8 Flash achieved 61.44% ± 0.13 Partial Credit, ranking No. 1 among the 58 systems listed on the page; its strict All-Pass score was 49.69% ± 0.42, ranking No. 2. It is suitable for e。

MediaSimpleBench official leaderboard and project

SimpleBench: Gemini 3.8 Flash on Everyday Reasoning and Language-Trap Questions

One-sentence takeaway The current SimpleBench leaderboard lists Gemini 3.8 Flash at No. 4 with 82.4% (AVG@5); it is below the highest human score of 95.4% and Claude Fable 5.1's 86.6%, and also below the human baseline of 83.7% (by 1.3 percentage points). This。

MediaVals AI

Gemini 3.8 Flash: Vals AI Harvey Legal Agent Benchmark Re-evaluation

One-sentence takeaway On Vals AI's Harvey's Legal Agent Benchmark v1 held-out test set, Gemini 3.8 Flash achieved a strict Task Pass Rate of 10.00% ± 2.42%, ranking 11th out of 59 systems; the page also shows a Criteria Pass Rate of 90.19% ± 0.36%. It passes m。

MediaCursor / Anysphere (CursorBench 3.2)

Gemini 3.8 Flash: CursorBench Multi-file Agent Tasks and Cost Boundaries

One-sentence takeaway On the current CursorBench 3.2 leaderboard, Gemini 3.8 Flash High scores 69.2% (9th among the 60 listed configurations), at an average of $2.38, 81,524 tokens, and 161 steps per task; Medium scores 67.0% (14th), at $1.93, 61,603 tokens, a。

Community

4 source-checked resources
CommunityX.com (Twitter), @arena (Arena.ai official account)

Arena.ai: Gemini 3.8 Flash (High) Agent, Text, and WebDev Rankings

One-sentence takeaway In Arena.ai's 2026-09-03 launch snapshot, Gemini 3.8 Flash (High) ranked 14th on Agent Arena (net improvement +5.94%), 7th on Text Arena (1,494 points), and 18th on Code Arena: WebDev (1,567 points). It was competitive on the cost-perform。

CommunityReddit, r/googleantigravity

Gemini 3.8 Flash: Reddit Short Tasks in a Serious Code Repository and Autonomy Boundaries

One-sentence takeaway The author used Gemini 3.8 Flash at the high tier in Antigravity for about an hour on deliberately underspecified short tasks in a serious code repository. They found it better than 3.7 Flash at planning and working from real code, and le。

CommunityX.com (Twitter)

Gemini 3.8 Flash: Four-Model 3D Rocket Launch Comparison and Generation Boundaries

One-sentence takeaway The author placed Gemini 3.8 Flash, GPT 5.6 Sol, Claude Opus 5, and Kimi K3 under the same brief and had each generate a 3D rocket-launch scene once. The author concluded that Flash did produce a runnable scene, but lost this comparison o。

CommunityReddit / r/GeminiAI

Gemini 3.8 Flash: Cross-Version Game Generation Comparison with the Same Creative Prompt

One-sentence takeaway One user used the same zero-shot creative prompt in Google Antigravity to have Gemini 3.1 Pro, 3.6 Flash, 3.7 Flash, and 3.8 Flash create a simple game combining Y City and Minecraft. The author felt that 3.7 and 3.8 Flash worked longer a。

Gemini 3.8 Flash

Use and compare models in Tabbit

Official benchmarks, independent analysis, and community reports about Gemini 3.8 Flash, clearly separated from Tabbit's own testing.