OpenAI GPT-5.6 Sol brings OpenAI flagship GPT-5.6 model for advanced reasoning, coding, scientific research, cybersecurity, and agentic workflows to Amazon Bedrock.
OpenAI GPT-5.6 Sol is the flagship model in the GPT-5.6 family and is the strongest model in the series. It is designed for complex work that benefits from deep reasoning, tool use, and multi-step execution across software engineering, professional knowledge work, scientific research, computer use, and cybersecurity.
Sol is well suited for teams building advanced AI applications and agentic workflows that require strong technical performance and long-horizon reasoning. It can support coding workflows, scientific and analytical work, vulnerability research, patch development, debugging, security education, and defensive testing.
Through Amazon Bedrock, organizations can build with GPT-5.6 Sol in the AWS environment where many enterprise workloads already run. Teams can access OpenAI capabilities alongside the AWS services, security controls, identity systems, and procurement processes they already rely on.
Highlights
Flagship GPT-5.6 model for advanced reasoning, coding, scientific research, cybersecurity, and complex agentic workflows.
Designed for high-value technical and professional work where teams need deep reasoning, tool use, and multi-step execution.
Available through Amazon Bedrock so organizations can build with OpenAI capabilities inside their existing AWS environment and workflows.
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
You pay only for what you use, measured in tokens. Pricing splits by token role: input, output, cached input, cache reads, and cache writes. Each role comes in processing modes: standard, priority, flex, and batch. Batch and flex suit high-volume or delay-tolerant jobs; priority suits time-sensitive work. Dimensions also vary by context length, with separate long-context rates, and by scope, with regional and global variants. Cache writes include a 30-minute retention option. You combine these attributes to match your workload, and charges add up across whichever token types your usage triggers.
Top-of-mind questions for buyers
What counts as one billing unit across these token dimensions?
Each unit is one token processed by the model. Tokens are chunks of text, roughly a few characters each. Input tokens cover your prompt, output tokens cover the generated reply, and cached tokens cover reused prompt content. You are charged separately for each token type your request triggers.
How do the cache-related dimensions affect my total cost?
Cache writes charge when you store prompt content for reuse, with a 30-minute retention option available. Cache reads and cached input charge when the model reuses stored content instead of reprocessing it. These charges bill independently and add to your input and output token charges on the same usage.
When would I pick batch or flex modes over standard or priority?
Batch and flex modes meter the same tokens but suit delay-tolerant, high-volume jobs where response speed matters less. Standard handles typical workloads. Priority meters the same tokens for time-sensitive requests needing faster handling. You choose the mode per request, and charges follow whichever mode you select.
developers.openai.com
Helpful?
Vendor refund policy
All sales are final. Fees are non-refundable.
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
SaaS delivers cloud-based software applications directly to customers over the internet. You can access these applications through a subscription model. You will pay recurring monthly usage fees through your AWS bill, while AWS handles deployment and infrastructure management, ensuring scalability, reliability, and seamless integration with other AWS services.
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
Use Daybreak Blue to access GPT-5.6 Sol for specialized cyber capabilities for verified defenders conducting advanced, authorized cybersecurity work to Amazon Bedrock.
OpenAI GPT-5.6 Terra brings a balanced GPT-5.6 model for everyday work, software engineering, knowledge workflows, and scalable AI applications to Amazon Bedrock.
OpenAI GPT-5.5 brings OpenAIs most capable model for complex professional work to coding, analysis, software operation,and long-running agentic tasks through OpenAI APIs and Amazon Bedrock.
Be the first to review this product. We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.