
OpenAI GPT-5.6 Sol, Terra, and Luna are now generally available on Amazon Bedrock
Quick Answer
OpenAI's GPT-5.6 models—Sol, Terra, and Luna—are now available on Amazon Bedrock, offering advanced capabilities for various workloads.
Quick Take
Sol excels in reasoning with an 80-point score on the Artificial Analysis Coding Agent Index, while Terra and Luna provide balanced and fast performance at lower costs, enhancing productivity across AI applications.
Key Points
- GPT-5.6 Sol scores 80 points on the Artificial Analysis Coding Agent Index, outperforming its predecessor.
- Terra offers superior performance for everyday tasks at a lower cost compared to GPT-5.5.
- Luna is optimized for high-volume inference tasks, focusing on latency and cost efficiency.
- Amazon Bedrock's inference engine supports bursty traffic and ensures data residency compliance.
- GPT-5.6 features robust safety measures, including real-time misuse classifiers and hardware-enforced security.
DeepSignal Analysis
What happened
OpenAI's GPT-5.6 models, named Sol, Terra, and Luna, are now available on Amazon Bedrock. Sol is noted for its advanced reasoning capabilities, scoring 80 points on the Artificial Analysis Coding Agent Index, while Terra and Luna are designed for balanced performance and cost efficiency.
Key evidence
- GPT-5.6 Sol achieved an 80-point score on the Artificial Analysis Coding Agent Index, outperforming the next-best model by 2.8 points.
- Terra offers superior performance compared to GPT-5.5 at a lower cost, suitable for everyday production tasks.
- Luna is optimized for high-volume inference tasks, emphasizing low latency and cost per token.
Why it matters
The introduction of these models on Amazon Bedrock enhances the capabilities available for various AI applications, particularly in coding, cybersecurity, and genomics. Sol's performance metrics suggest significant improvements in efficiency and cost-effectiveness, which could influence how organizations approach AI workloads. The tiered model structure allows users to select the appropriate model based on their specific needs, potentially optimizing resource allocation.
What to watch
📖 Reader Mode
~5 min readArtificial Intelligence
Build with the smartest family of models from OpenAI yet, on Amazon Bedrock’s next-generation inference engine.
Organizations scaling autonomous agents and AI-powered products need frontier intelligence that performs reliably across hundreds of steps, from coding agents shipping production code to cyber security research probing novel attack surfaces to genomics workflows analyzing entire gene sequences end-to-end. These workloads run on sensitive data, demand consistent throughput under unpredictable load, and operate in environments where data residency and security are non-negotiable.
Today, GPT-5.6 Sol, Terra, and Luna from OpenAI are generally available on Amazon Bedrock, bringing the smartest family of models from OpenAI yet to Amazon Bedrock’s next-generation inference engine built for high-performance, security and reliability. GPT-5.6 sets a new standard for intelligence and efficiency, so you can solve harder problems in less time and with more intelligence per token. Pricing matches OpenAI first-party rates, and usage counts toward your existing AWS commitments.
Frontier intelligence from flagship reasoning to fast inference
GPT-5.6 introduces a new naming system from OpenAI: the number identifies the generation, while Sol, Terra, and Luna identify durable capability tiers that can advance on their own cadence.
- GPT-5.6 Sol is the flagship reasoning model and the most powerful from OpenAI to date. According to OpenAI, Sol sets a new state of the art on the Artificial Analysis Coding Agent Index at 80 points (2.8 above the next-best model) while using less than half the output tokens, taking less than half the time, and costing about one-third less, and scores 73.5% on ExploitBench for cybersecurity research versus 47.9% for GPT-5.5 at a comparable output-token budget. On Agents’ Last Exam, an evaluation of long-running professional workflows across 55 fields, Sol sets a new high of 53.6, outperforming the next-best model by 13.1 points. At medium reasoning effort, it still leads by 11.4 points at roughly one-quarter the estimated cost. Sol also introduces max reasoning effort, which lets you dial up compute for complex work. Use Sol for autonomous coding agents, vulnerability research, drug discovery workflows, and tasks that require deep multi-step reasoning.
- GPT-5.6 Terra is the balanced model for everyday production work. It delivers superior performance to GPT-5.5 at a lower cost. Use Terra for code generation, content workflows, structured data extraction, and general-purpose agentic tasks that need strong reasoning without flagship pricing.
- GPT-5.6 Luna is the fast and affordable model. Use Luna for high-volume inference tasks like classification, summarization, routing, and real-time applications where latency and cost per token matter most.
With the three tiers, you can right-size model capability and cost to each workload. GPT-5.6 models complete tasks with fewer output tokens than their predecessors, delivering stronger performance per dollar.
An inference engine built for scale
Agent traffic is often bursty: one user request can trigger hundreds of model calls, and demand can change quickly as usage grows. Amazon Bedrock’s next-generation inference engine pools capacity to absorb demand spikes while isolating each customer’s throughput. This reduces the need to choose between shared capacity and predictable application performance. In-Region inference keeps requests within the AWS Region that you specify, helping teams meet strict data-residency requirements.
Agentic and multi-step workloads also repeat much of their context between calls. System instructions, tool definitions, and reference files often stay the same while only the latest input changes. GPT-5.6 on Amazon Bedrock introduces prompt caching with explicit cache breakpoints to take advantage of that repetition. You mark the reusable part of a prompt with a cache breakpoint, and Amazon Bedrock reuses the processed context on subsequent requests that share it, so each call pays only for the new work. Cached input is billed at a 90 percent discount and stays available for reuse for at least 30 minutes. This is long enough to cover the burst of calls a single agent run generates without compounding cost as workloads scale.
Robust safety meets hardware-enforced security
More capable models require stronger safeguards. GPT-5.6 features OpenAI’s most robust safety stack to date, shaped by their most extensive evaluation period yet, combining human red teaming with large-scale automated testing. The safety stack includes model-level refusals for prohibited activity, real-time misuse classifiers, continuous monitoring, and account-level enforcement for persistent patterns.
On Amazon Bedrock, these protections sit on top of hardware-level security. Amazon Bedrock uses a zero-operator access (ZOA) security model enforced at the chip, so no AWS operators can access your prompts or completions. Every model call runs under your AWS Identity and Access Management (IAM) policies, inside your virtual private cloud (VPC), and is logged in AWS CloudTrail. Data perimeter policies prevent exfiltration across account and network boundaries. As required by the model-provider, classifier-flagged traffic data will be retained for up to 30 days for automated abuse detection.
More ways to put GPT-5.6 on Amazon Bedrock to work
Alongside GPT‑5.6, OpenAI launched ChatGPT Work, an agent in ChatGPT for larger, multi-step tasks. The updated ChatGPT desktop app for Mac and Windows brings Chat, Work, and Codex together in one experience, with two dedicated agents for different kinds of work:
- Work: The new Work agent is built for bigger, multi-step tasks. It can gather information across apps and files, use the web, create finished materials such as sheets, slides, docs, and Sites, and stay with complex projects for hours. Users can follow its progress, change direction, and approve important actions.
- Codex: Codex remains the powerful coding agent for developers and technical professionals. It can work with local files, repositories, terminals, developer tools, and development environments to write features, fix bugs, run tests, and open pull requests.
Users can configure the app to use GPT-5.6 through the Responses API on Amazon Bedrock. Run the smartest family of models from OpenAI on Amazon Bedrock and get frontier intelligence with the security and scale of AWS.
Get started today
GPT-5.6 Sol is available in the following AWS Regions: US East (N. Virginia) and US East (Ohio). GPT-5.6 Terra and Luna are available in US East (N. Virginia), US East (Ohio), and US West (Oregon).
Get started with Sol, Terra, and Luna in the Amazon Bedrock Console or programmatically through the Responses API. To learn more, see the Amazon Bedrock documentation and visit the Amazon Bedrock product page.
About the authors
— Originally published at aws.amazon.com
Want this in your inbox every morning?
Daily brief at your local 8am — bilingual EN/中文, free.
More from AWS Machine Learning
See more →
Building an agentic app deployer with Amazon Bedrock and AWS Lambda
PDI Technologies developed PDI Brew, enabling non-technical employees to create web applications on AWS without developer involvement, leveraging Amazon Bedrock for AI capabilities. This agentic app deployer streamlines internal tool delivery, removing traditional bottlenecks in deployment pipelines.

