Teams are building AI into the places where work already happens, across documents, spreadsheets, presentations, software systems, tools, and business workflows. GPT-5.4 brings frontier reasoning, strong coding performance, computer-use capabilities, and long-context support to applications that need to operate across those environments.
GPT-5.4 helps developers build AI applications and production workflows that can interpret context, interact with tools, operate software environments, and verify outputs across multiple steps. It is well suited for professional workflows that require reliable reasoning and action across complex business systems.
Through OpenAI APIs and Amazon Bedrock, organizations can build reliable AI applications and production workflows with GPT-5.4s frontier reasoning capabilities and the AWS controls they already use.
Highlights
Access GPT-5.4 for frontier reasoning, coding, computer use, document and spreadsheet workflows, presentations, software operation, and professional knowledge work.
Build AI applications and agents with long-context reasoning, tool use, computer interaction, and workflow execution across complex environments.
Support more reliable professional workflows with a model designed to plan, execute, and verify complex tasks, while using Amazon Bedrock for AWS-native access, governance, billing, and eligible commitment drawdown.
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
You pay per token used, with no upfront commitment. Pricing splits into three token types: input, output, and cached input. Each type is billed separately. Rates vary by four processing modes: standard, priority, flex, and batch. A long-context option applies when prompts are large. Many dimensions carry an AWS Region prefix (such as USE1 or EUW2), so your rate depends on where you run the model. Global cross-region dimensions are also offered. Cache-read tokens are priced on their own. Your total cost scales with token volume, chosen mode, context length, and Region.
Top-of-mind questions for buyers
What counts as one billing unit, and how are tokens measured?
Each unit is a token, the small text fragment a model reads or writes. Input tokens cover your prompt, output tokens cover the model's reply, and cached input tokens cover reused prompt content. The system counts tokens separately for each type and bills each at its own rate.
How do the standard, priority, flex, and batch processing modes change what I pay?
Each mode is a separate rate for the same token. You pick a mode per request based on speed and workload needs. Priority targets faster handling, flex and batch suit less time-sensitive or high-volume jobs. Your token count stays the same; only the per-token rate shifts with the mode you choose.
How do the different token charges combine into my total bill?
Charges apply independently and add together. Input, output, and cached input tokens each bill at their own rate, then sum per request. Output tokens often carry higher rates than input tokens. Cached input and cache-read tokens lower cost when prompt content repeats. Your Region and chosen mode set which rates apply.
developers.openai.com+1
Helpful?
Vendor refund policy
All sales are final. Fees are non-refundable.
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
SaaS delivers cloud-based software applications directly to customers over the internet. You can access these applications through a subscription model. You will pay recurring monthly usage fees through your AWS bill, while AWS handles deployment and infrastructure management, ensuring scalability, reliability, and seamless integration with other AWS services.
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
OpenAI GPT-5.5 brings OpenAIs most capable model for complex professional work to coding, analysis, software operation,and long-running agentic tasks through OpenAI APIs and Amazon Bedrock.
OpenAI GPT-5.6 Sol brings OpenAI flagship GPT-5.6 model for advanced reasoning, coding, scientific research, cybersecurity, and agentic workflows to Amazon Bedrock.
OpenAI GPT-5.6 Terra brings a balanced GPT-5.6 model for everyday work, software engineering, knowledge workflows, and scalable AI applications to Amazon Bedrock.
Be the first to review this product. We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.