https://mp.weixin.qq.com/s/V4xhEIy8xDXSMDPrPkmUAQ
Generally looks like a Sol/Fable tier model, better across the board than Opus 4.8.
(Edit) English blogpost is up now: https://www.kimi.com/blog/kimi-k3
https://mp.weixin.qq.com/s/V4xhEIy8xDXSMDPrPkmUAQ
Generally looks like a Sol/Fable tier model, better across the board than Opus 4.8.
(Edit) English blogpost is up now: https://www.kimi.com/blog/kimi-k3
Forget about their pricing but the companies that do have means to host such models fully on-prem are also the same companies that are paying tens of millions of $ in inference cost every month, and are by extension the biggest customers of OAI and Anthropic
I don't want to cheer against my country, but we've given up on open source. The way Anthropic and OpenAI treat their customers as adversaries is embarrassing.
I will cheer for China, for Kimi, and for z.ai until we have something in the same category.
[1] I'd even be fine with open weights, fair source, or anything that let us have direct access to the weights. Even if that came with stipulations. Don't hide the weights from us.
The argument on our side wins - if America or the West don't do open source, China will. And that means -- with certainty -- that China wins the market.
Every politician and VC should hear that loud and clear.
After using it for a few hours, I believe these benchmarks.
(I mantain a client with llama.cpp and 101 models across 14 companies by http)
Having said that, the safety system on Fable makes it an extremely unattractive model. It feels that half of the time you're paying double for Opus level performance.
I finally bumped into a task that Codex would refuse to work on.
Was I attempting to reverse-engineer a GPU driver? Yes. Was I trying to hack into the DoD? No.
I wasn't doing anything wrong, but that's not what OpenAI's safety mechanisms thought.
With Oracle being junk before this, more will follow.
Now they are betting with Project Stargate but it also seems to be crumbling down.
But don't forget that they literally hold the biggest databases, both in commercial and open source, that is, Oracle Database and MySQL. Plus Oracle Java they literally controls at least 30% of the internet's software infrastructure.
And also with a good team of attorneies enforcing the licenses, they can squeeze so much money at the cost of morality.
Also recently they downgraded the always free OCI ARM instance from 4C24G to 2C12G without telling anyone.
They're drowning in debt and risk is increasing. If these US models don't keep holding up their valuation will tank further and some will recall the loans or ask for different terms.
The DeepSeek incident has already shown it, this is a reminder.
This would drive down Anthropic's margins, but drive up demand for datacenter and GPU capacity. It's not that people would be using fewer GPUs, they'd just shift demand from high priced token vendors to direct GPU rental, which benefits datacenter companies while hurting Anthropic.
Correction: Lots of organizations are refusing to use Anthropic Fable because they have forced opt-in data collection as part of their privacy policy, even for Enterprise.
Not everyone's going to care about Anthropic requiring data collection (a similar debate plays out with regards to "pay or consent" on website tracking), just as not everyone cares about China with regards to security/IP issues (if they did, a lot more would be banned besides occasionally-Huawei).
This is such a common omission: the Chinese models are open, you can host them yourself on your premises. So privacy and independence.
while I am skeptical that this is happening atm, there are probably many industries where the risk does not seem worthwhile
Maybe I just don't have any imagination.
Because in 2026 we still believe USA is more trustworthy than China?
I have an RF engineering background, a nice mmWave vector network analyzer can easily land in that ballpark.
If the business value is there, companies will pay for it.
One can run open weights in an exclusive TEE'd GPU too, which still comes out cheaper than closed weight LLMs. Ex: https://chutes.ai/pricing
These customers exist (e.g. US military) but there's not enough of them to justify a trillion dollar valuation.
Anthropic's valuation is predicated on growth. If they start going backwards and losing customers to open models, it hurts their ability to gather investment and with it the ability to train new models, leading to a death spiral.
The best they can hope for is that the US gives them state aid to compete with China, however their relationship with the current administration is not great.
The reality this demonstrates: most US companies don't give even 2 shits about their IP, and are fine willingly handing it to Anthropic et al. Those that do care largely must care contractually. For group 2, they're either using Chinese today or aren't using AI at all. Those are the only two valid options, there is no secret "use US but self host it" third option.
https://www.youtube.com/watch?v=LSlV206xPqM
These real world examples show it's one tier away.
Everybody can agree that K3 doesn't clearly surpass Fable. However, inevitably there will be a time in the future when a Chinese AI company releases a model that's better than any US model.
K3 isn't the knockout blow but it's the 2nd knockdown that makes everyone in the arena realize that the fighter is not winning the fight.
For code editing Cursor editor tooling is even better.
Given the pricing, it suggests that this model is much more efficient/competent than previous-gen OS/distilled models.
https://nitter.net/synthwavedd/status/2077537805715005724#m
(As an aside, I don't know how it was professional of Arena to unmask an unreleased cloaked model on their platform. Also practically, upstream could have been A/B testing multiple variants under same endpoint, casting validity of such pre-announcement tests into question)