PS: Have not tried this but Deepseev4 Flash (not even Deepseekv4 Pro version) with set to "high" has pretty much Claud Opus 4.7 level of capabilities and is lightening fast and dirty cheap. Hours and hours of conversation barely costs few cents.
PS: Have not tried this but Deepseev4 Flash (not even Deepseekv4 Pro version) with set to "high" has pretty much Claud Opus 4.7 level of capabilities and is lightening fast and dirty cheap. Hours and hours of conversation barely costs few cents.
Very disproportionate intelligence-to-cost ratio.
I'm leveraging this temporary anomaly and using it as my coding workhorse.
I can easily run it in a 8 bit quant with the 4 x 48GB Radeon Pro W7900 GPUs I snagged for 2k each before the memory squeeze.
A 158B parameter model, especially in an architecture as efficient as DS4 is not that hard to drive currently if you got in before the craze, and will be relatively easy to drive with future hardware generations.
A big caveat here is that many US companies (particularly in sensitive industries, like defense) will likely not want to (or not be allowed to) use Chinese models for anything of substance.
You're referring to a very small subset of the American population. It's ironic because you seem to be claiming Americans are closed-minded here but I think that may actually describe your mindset as well.
I can only conclude that people who claim they are aren't doing anything close to the edge of what these models are capable of or any niche things.
I would say DSv4Pro is around the same level as Sonnet.