GPT-5.5 Outcome-Oriented and Verification-Driven Agent Prompt
OpenAI's GPT-5.5 guidance puts the goal, success criteria, allowed side effects, evidence rules, and delivery fields first, then asks the agent to verify the result with targeted checks.
Prepare
task goal, source or reference material, runtime constraints, acceptance criteria
Runtime
GPT-5.5 client or API; confirm the live model ID, tools, permissions, and version before execution.
GPT-5.5 Responses API Reasoning and Multi-turn State Configuration
This official configuration guide focuses on Responses API reasoning levels and multi-turn state transfer, with response items preserved across tool calls.
Prepare
task goal, source or reference material, runtime constraints, acceptance criteria
Runtime
GPT-5.5 client or API; confirm the live model ID, tools, permissions, and version before execution.
GPT-5.5 Official Benchmarks, Pricing, and Safety Boundaries
OpenAI's release page attributes GPT-5.5 results in tool-heavy coding, browsing, and cross-software agents to specific harnesses; the tables cannot be reproduced without the snapshot and tools.
Evidence
Vendor report
Boundary
It does not support treating those figures as an independently reproduced win rate, a guarantee for another snapshot, or a production SLA.
GPT-5.5 Vellum Cross-Model Benchmarking and Vendor Data Boundaries
Vellum aggregates public GPT-5.5, Claude, and Gemini scores and shows substantial task-to-task variation; it is a cross-source digest, not a controlled rerun.
Evidence
Independent measurement
Boundary
It supports task-based candidate screening, not a single overall score, current pricing, or a safety-critical deployment guarantee.