Global AI, Malaysia Angle / AI news for Malaysia
OpenAI's GPT-6 Sol and Luna arrive on Azure and AWS at lower prices
Two cheaper GPT-6 models are now sold through the big clouds, with Luna priced for high-volume chores. But neither cloud says whether Malaysian businesses can keep that processing inside the country.

In brief
- OpenAI released GPT-6 Sol and GPT-6 Luna on 22 September 2026, and both went on general sale the same day through Microsoft Foundry and Amazon Bedrock.[1][2][3]
- Microsoft lists Luna at US$0.10 per million input tokens and US$0.50 per million output tokens on its Global tier, a twentieth of Sol's rate.[2]
- Neither cloud's announcement confirms that the models run in a Malaysian region. Microsoft names only US and EU data zones.[2][3]
What was announced
OpenAI on 22 September 2026 introduced GPT-6 Sol and GPT-6 Luna, two models it describes as bringing frontier intelligence to everyday work with different balances of capability and cost. They sit below GPT-6 Astra, the most capable and most expensive model in the family.[1][3]
On the same day, Microsoft and Amazon Web Services said both models were generally available on their AI platforms, Microsoft Foundry and Amazon Bedrock. For many companies, that means the models can be bought through an existing cloud account rather than a separate OpenAI contract.[2][3]

Two models for two kinds of work
AWS describes Sol as a model for demanding tasks that recur through the week: writing features, debugging, reviewing code, analysing data and completing multistep work across tools. It says OpenAI's own internal factuality test found Sol made about half as many factual mistakes as its predecessor, GPT-5.6 Sol. That figure comes from OpenAI and has not been independently checked.[3]
Luna is the smaller, faster sibling. Microsoft recommends it for extraction, summarisation, routing requests and routine customer conversations, while AWS stresses workloads that run thousands of times a day. Both clouds suggest mixing models in one system, for example Luna to sort incoming requests, Sol to investigate harder cases and Astra only where deeper reasoning changes a decision.[2][3]

How the price ladder works
Microsoft published full rates for its Global Standard tier. Sol costs US$2 per million input tokens and US$10 per million output tokens for short prompts. Luna costs US$0.10 and US$0.50. Astra costs US$10 and US$50. Rates roughly double for long-context requests, and Microsoft's US and EU data-zone deployments carry a premium over Global.[2]
AWS did not print a rate card in its post, saying only that both models come at significantly lower API pricing than their GPT-5.6 predecessors. Microsoft urged buyers to look at cost per task rather than price per token, because a model that needs fewer retries can end up cheaper even at a higher rate.[3][2]
Caching and data controls
Both models support prompt caching, which lets repeated instructions, policies or reference documents be reused instead of processed again on every call. On Microsoft's Global tier, cached input for Luna is listed at US$0.01 per million tokens. For a support desk that sends the same policy text with each question, that is where much of the saving is likely to sit.[2][3]
AWS says inference data on Bedrock is not used for model training and customers do not have to opt in to share data with OpenAI. Traffic flagged by its abuse classifiers can be kept for up to 30 days, and zero data retention is available on request through an AWS account team.[3]
What the announcements leave out
We could not read OpenAI's own launch page, which refuses automated readers, so its title, date and summary were confirmed from OpenAI's official news feed. OpenAI's direct API prices for Sol and Luna are therefore not reported here, and they may differ from Microsoft's cloud rates.[1]
Location is the bigger gap for Malaysian buyers. Microsoft says standard deployment covers all 28 of its Global regions plus US and EU data zones, but it does not list the regions or offer an Asian data zone. AWS points readers to its documentation for supported regions without naming any in the post.[2][3]
Why Malaysia should care
AWS has run a Malaysian cloud region since August 2024, yet neither the AWS nor the Microsoft announcement says the new models are served from Malaysia. Firms with data-location duties should confirm the serving region before sending customer records.
Malaysian SMEs
Routine sorting and summarising jobs now have a much cheaper model tier.[2][3]
Practical move: Test Luna on one repetitive task before paying for a bigger model.
What Malaysians can do now
- Pick one high-volume chore, such as tagging customer enquiries, and run a week of it on Luna with a person checking a sample of the results.
- Before sending personal data, ask Microsoft or AWS which region will serve your GPT-6 requests and get the answer in writing.
- Put repeated instructions and policy text in cached prompts, then check your next bill to see whether the cached rate is being applied.
What we still do not know
What we still do not know
- Whether GPT-6 Sol or Luna can be served from AWS's Malaysian region or from any Microsoft region in Malaysia.
- What OpenAI charges for the two models on its own API.
- How Sol and Luna perform on independent tests rather than OpenAI's internal evaluations.
Sources
- 1.Introducing GPT-6 Sol and Luna OpenAI, 22 September 2026
- 2.GPT-6 Astra, Sol, and Luna: For production agents in Microsoft Foundry Microsoft Azure, 22 September 2026
- 3.Bring more intelligence to everyday work with GPT-6 Sol and GPT-6 Luna on Amazon Bedrock Amazon Web Services, 22 September 2026
- 4.Now open: AWS Asia Pacific (Malaysia) Region Amazon Web Services, 21 August 2024
- 5.File:Wenatchee washington azure datacenter aerial 1.jpg Wikimedia Commons, 22 August 2025
- 6.File:Amazon AWS us-west-2 beach AZ.jpg Wikimedia Commons, 11 August 2023
- 7.File:Bangsar South night view (230319).jpg Wikimedia Commons, 19 March 2023


