OpenAI models coming to Amazon Bedrock: Interview with OpenAI and AWS CEOs
stratechery.com
stratechery.com
This paper isn't the exact same scenario, since it's an auditable open weight llama model, but shows the symptoms of this: https://arxiv.org/pdf/2410.20247
Model performance consistency is important not because you want inference determinism (which you can actually get by setting tempetature to zero and applying a static seed). The `another axis of non-determinism` can be illustrated by the question "if I move from openrouter to bedrock, will gpt-5.5 perform the same?", to which the answer is no, at least not necessarily.
This is important because workflows that used to work on one platform might degrade or outright not work on another, even using the same model, which you have to account when deciding which provider to use.
For Anthropic, it can vary based on model and time. For Opus 4.7, Bedrock is the clear winner in TPS by leaps: https://artificialanalysis.ai/models/claude-opus-4-7/provide...
I wonder if this is directly linked to the split up with Microsoft. Just from my anecdata, OpenAI is getting completely ignored in serious enterprise deployments because what they offer on Azure sucks and there is no other corporate friendly way to get it. They probably saw themselves getting destroyed in enterprise and realised it was existential to be able to compete with Anthropic on AWS.
It was also very clear the OAI and MS teams held each other in contempt (not relevant, but was interesting and grew in the immediate aftermath of the Altman drama).
So why were we using it? OpenAI don’t really have an enterprise go to market, bedrock still relied on Claude 2, and we weren’t willing to YOLO on clickthroughs.
Once Claude 3 came out, we jumped ship. That sucked too, although I hear it’s gotten better though.
So yeah Azure sucked ass and plenty of outages or latency, like 3min for first byte while usually it was max 30sec to 1min, if not even faster (memory is a bit fussy)
secondarily its an industrialization of software development. the hiring process is where they try define the labor as a replaceable component. grab the best cogs you can annually for the lowest price, run them in the machine for 2-4 years and swap most of them out before they get too expensive, or specialized or uppity.
Why would I care if AWS asks their engineers to work a little harder on a project
...anyone with a brain at AWS knows that supporting OpenAI's latest models on Bedrock is simply good for AWS. That context is rather important!
There's always some carrot with the stick, even if an imaginary one!
Openai hasn't been publishing innovations for quite a while.
They're both just stealing ideas from pimono extensions
[0] https://business.columbia.edu/sites/default/files-efs/imce-u...
[1] https://www.reuters.com/business/autos-transportation/compan...
We will see if this changes the equation, but it feels like OpenAI is pretty far behind and playing catch up on all fronts. Though to be honest, "pretty far behind" is like 2-8 weeks in the AI world, so it may not matter a ton, it's mostly perception. And for me and my information bubble, perception of OpenAI is rock-bottom due to Sam Altman. From appearing unethical to appearing unhinged with demands from fabs and everything else, I'm not a fan.
[0]: https://platform.claude.com/docs/en/build-with-claude/claude...
I think that when people are worried about ZDR, what they really worry about is data governance. From what I’ve seen there’s a general distrust of OpenAI. AWS may keep your data around (without formal ZDR) but the concern of governance (using your data to train without your consent) seems like it would be much lower, because any breach of contract at AWS would have potential to destroy trust in what’s already a massively profitable company, so the incentives just aren’t there.
I’m not claiming OpenAI is training on API data. Just that they don’t have as strong of an incentive not to as AWS.
And for OpenAI, there is a May 2025 preservation order in NYT v. OpenAI. The court is forcing OpenAI to retain ChatGPT output logs indefinitely, including chats users have deleted that would normally be purged within 30 days [2]. That makes it a non starter for HIPAA/GDPR bound orgs.
> Update on October 22, 2025:
> After months of litigation, we are no longer under a legal order to retain consumer ChatGPT and API content indefinitely. Our obligations under the earlier order ended on September 26, 2025.
> We’ve returned to our standard data retention practices :
> Deleted ChatGPT conversations and Temporary Chats will be automatically deleted from our systems within 30 days (opens in a new window).
> API data will also be automatically deleted after 30 days.
As for Musk ever being "the left"'s "hero" -- that's amazing, that's what Pauli would call 'not even wrong'.
If by "the best bet" you mean slightly less shitty bet then maybe.
Nuff said
It has the most "UNIX" feel of a simple app that you compose the just right flow from and nothing more
What OP is referring to is Anthropic aligning with corporate terms and conditions early, positioning themselves to be effectively resold by AWS rather than requiring orgs to procure them directly. This is huge in the enterprise world because the processes to get broad approval are generally far smaller and shorter for "just another AWS service" compared to a whole new vendor.
Oai language models are largly irrelevant at this point imo.
You can just run "air gapped" inference?
Is this only of interest to enterprise customers already on AWS (who want "air gapped" behavior)? Is there any other use case for this?
This will be more expensive than calling OpenAI directly, right?
If this ends up similar to Claude on Bedrock, it's the same price.
But it also is for Devs in a company who already have a blanket agreement with Amazon, but would have an uphill battle signing an agreement with openAI.
I think cost is fairly similar. As for who wants it, enterprise stuff is a big thing, but it also seems very reliable. Have never had any rate restrictions or had it just be down, which it sounds like a lot of people have issues with for Anthropic's servers.
https://docs.aws.amazon.com/bedrock/latest/userguide/bedrock...
This HN post itself has 4 simultaneous announcement links; not a coincidence.
There are billions of investor money on the line if the wrong thing is said at the wrong time, it needs to be carefully crafted and staged.
0. https://www.cnbc.com/video/2026/04/28/tech-shares-fall-after...
It’s good to see more model diversity on Bedrock, but this smells more a rushed play by OpenAI to stop the market share bleeding vs a concerted strategy.
But with Codex and GPT5.5 having generally good reviews, when it goes on Bedrock, it would be very simple for us to just try it out. Obviously there's the general friction of knowledge and comfort of Claude Code vs Codex, but that seems possible to overcome depending on the differential of the models and also of the features. Like if Codex gets remote control for Bedrock, that'd be a big advantage over Claude Code on it.
Since the product doesn't seem to be available yet, and the other links are all press releases, we'll leave the interview up as the main link.
Microsoft Azure has been the worst interms of maintaining a highly available service and also managing predictable latency.
Their azure customer support is bad. Not ready for any real enterprise cloud offering. They behave like Comcast customer support.
It was absolutely idiotic to lock it down to Azure. It wasn't meant to be an iphone+at&t combo where the phone is an end all be all.
A cloud product depends on a lot of services and nobody would switch cloud providers for a candy.
AI is kind of like the ultimate corporation drug. They are all on it. And can't get rid of it - ever again.