Amazon Web Services (AWS), an Amazon.com, Inc. company (NASDAQ: AMZN), today announced the general availability of OpenAI GPT-5.6 Terra and Luna models on Amazon Bedrock in India with in-country inferencing. With this capability, OpenAI model inference on Amazon Bedrock is processed on AWS infrastructure, giving organisations the ability to leverage frontier models locally. This launch enables customers across financial services, healthcare, public sector, start-ups, and other industries to access OpenAI models locally within India while trusting the security, governance and operational controls they already rely on in AWS.
GPT-5.6 Terra is a balanced, general-purpose model for everyday enterprise workloads that delivers superior performance to the previous generation at a lower cost, while GPT-5.6 Luna is the fastest and most cost-efficient model, purpose-built for high-volume, latency-sensitive tasks such as summarisation, classification etc. Following OpenAI's latest round of price reductions, GPT-5.6 Luna now costs up to 80% less and GPT-5.6 Terra up to 20% less. These models bring frontier intelligence to India at a dramatically lower cost, with the new pricing reflected on Amazon Bedrock.
“What we're seeing in India is a shift from AI as experimentation to AI as operating leverage. Businesses want models that can help teams write better software, work through complex information, and make higher-quality decisions in the flow of work. For many organisations in India, especially in regulated sectors, local inferencing is a key part of adopting AI with confidence. Making our models available through Amazon Bedrock gives more organizations a practical way to bring that capability into production,” says Nitin Bawankule, Head of Enterprise Sales, India, OpenAI.
“India's enterprises are moving rapidly from AI experimentation to production-scale deployment, and they need frontier intelligence that runs where their data already lives. By bringing OpenAI advanced models to India, we are giving customers the ability to build transformative AI applications with in-country inferencing, and governance controls that organisations in India demand - all on the trusted AWS infrastructure they already use," says Satinder Pal Singh, Director, Solution Architecture, AWS India and South Asia.
Choice and flexibility through a unified API
With Amazon Bedrock, customers can now access OpenAI models alongside a wide selection of foundation models available on AWS - all through a single API, with no infrastructure changes. This allows organisations the choice to find the best-suited model for their needs while keeping their development workflow and infrastructure consistent. This also lets teams match each workload to the most cost-effective tier – routing everyday tasks to Terra and high-volume jobs to Luna.
Enterprise security, governance, and compliance
Every call to OpenAI models on Amazon Bedrock benefits from the enterprise-grade controls customers already use across AWS. Organisations can apply existing governance, security and operational practices as they bring advanced AI into production environments.
Use cases driving demand in India
With OpenAI advanced models now available for in-country inference in India, organizations across India can apply frontier AI to workloads to solve real business challenges at scale:
- Agentic coding: Codex can help complete tasks end to end, like building features, complex refactors, migrations, and more enabling organisations to ship products faster while keeping costs predictable.
- Data analysis: Allows business teams to go from raw data to finished reports autonomously - generating documents, spreadsheets, and actionable insights without manual intervention - helping organisations make faster, data-driven decisions.
- Agent workflows and ChatGPT work: Handles complex, multi-step business processes, giving India's banking, financial services, and healthcare sectors the automation they need without compromising on privacy.
- Production inference: Enables Indian businesses to access frontier intelligence reliably, at scale, and with the operational cost efficiency to sustain it at scale
- Software development: Engineering teams can use GPT-5.6 Terra for code generation, review, and testing across large codebases.
- Regulated workflow automation: Organisations in banking, financial services, and healthcare can build multi-step agents on Amazon Bedrock to automate processes such as KYC verification, regulatory filings, and document processing.
- High-volume customer applications: GPT-5.6 Luna enables organisations to deploy AI-powered chatbots, voice assistants, and support systems at significantly lower cost per interaction - designed for businesses serving millions of users daily where cost efficiency at scale is essential.
- Sensitive data analysis: Teams in financial services, healthcare, and the public sector can generate forecasts, synthesize unstructured records, and produce actionable reports using frontier models.









