Nvidia RTX 5000 Ada 32GB Workstation GPU Review
servethehome.com
servethehome.com
"fewer", not "less".
I’m sorry.
Edit: as someone who can’t even figure out the markdown to italicise right now, maybe I should have kept my nose out!
Interesting. What is it?
(See, for example, “literally,” or “comprise,” or “hopefully.”)
Why not have the word "no" mean "yes" and "no" while we're at it?
- to adhere firmly and closely or loyally and unwaveringly
- to divide by or as if by a cutting blow : split
https://www.merriam-webster.com/dictionary/cleave
So cleave is it's own opposite. It causes no confusion because context is everything. q.v. homonyms.
Edit: "The word "cleave" has two opposite meanings - either to stick together or to split apart. Are there any other words that do the same thing?" https://www.theguardian.com/notesandqueries/query/0,5753,-13...
The problem with defending the purity of the English language is that English is about as pure as a cribhouse whore. We don't just borrow words; on occasion, English has pursued other languages down alleyways to beat them unconscious and rifle their pockets for new vocabulary.
- James Nicoll
"Literally" has meant figuratively for centuries[0]. It's a language construct known as a contronym[1,2].Yes, English is not strictly typed, and doesn't conform to a formal spec or mathematical proof. A word can have multiple definitions. A word can have contradictory definitions. The definition of a word can change over time without needing to submit to an approval process. A word can mean something different the next town over. Dialects and creoles and slang exist in wild, flagrant abandon and disregard for your rules. And all of it is valid.
Despite all of this, people manage to be able to comprehend one another, even when using literally to not literally mean literally.
Even the user who used "less" instead of "fewer" upthread that started this. Everyone understood exactly what they meant. The two words mean literally the same thing. But some people insist on maintaining a meaningless formality and insisting upon rules that don't matter, or rather insisting that even in casual conversations, their rules must take precedence over everyone else's.
[0]https://www.thecut.com/2018/01/the-300-year-history-of-using...
[1]https://en.wikipedia.org/wiki/Contronym
[2]https://www.quora.com/Etymology-When-did-the-word-literally-...
You don't need to jail or slap someone in the face for using less instead of fewer but you don't want gynecologist to mean oncologist and cancer to mean gonorrhea from one week to another otherwise nobody knows what we are talking about.
In English there is both essential plumbing and pointless ornament, and I wish for that distinction to be recognised because confusing the two is damaging.
I'm a non native speaker and surely make a lot of mistakes, and I would be OK with such a correction
The person you replied to used perfectly common and comprehensible English, so much so that you knew exactly how to "correct" them.
So the comment was grand and your prescriptivist response adds nothing of value. Hopefully you can see why this is true and reflect on why you felt the need to "correct" it.
There we go. The "response" was prescriptivist. The person posting it (me) and yourself are both somewhere in between the two. Now I think you're starting to get it.
Regardless, the original response was a just a joke referring to Stannis Baratheon (from Game of Thrones) saying it. I think some people read it that way. You didn't.
Absurd pricing, ridiculous vram offering, I'm sure they're trying very hard to find a way to stop AI like SD or LLMs to run on their gamer cards at this point.
It's reached a point where not only has ATI/AMD essentially caught up to them in rasterization, but they're frankly a better offer at every price point against the 4XXX generation for pure gaming, with only DLSS and brand recognition keeping nvidia ahead.
The equipment from ASML and the know how of TSMC and Samsung is far far out of reach of china now.
Intel can't even seem to get down to 5nm.
A wonderful thing would be a GPU with open source drivers and a standartized API. AMD could do that and it would kill nVidia marketshare in a few years, but they don't do that because they want the server gpu money at a very inflated margin. Maybe Intel will do it but that depends who runs the company. With open source often corporations publish a feture stripped solution and keep the full drivers closed source. Sad reality we live in.
If 300M people in the US, let alone 7B in the world, need even a 4060 to run basic business workloads, NVidia is sitting pretty.
That's why you have an 8 GB 4060 Ti then a 16 GB one then the 4070 Ti with 12 GB ... It's a mess.
They artificially segment the market because "pro" users are willing to pay almost 10k for a good card. So why not offer just that?
So now you have gamers who are pissed off at their 4060 Ti or 4070 Ti barely having enough vram for 1440p at modern quality of gaming, and AI enthousiast wondering why buy they "pro" card when a 4090 is not even 2k€.
essentially, they're trying to virtually segment a market that is not segmented, to take advantage of the much higher buying power in one of the segment.
The 4090 wins in every benchmark for 1/3rd the price. Why would anybody buy this card? Is 8 GB more VRAM and lower power consumption really worth that much when the performance is so lackluster?
(Not saying you shouldn't game on a workstation, but a workstation card will be worse at it especially for its price)
I guess what's unique with Ada is that they're using her first name? Though most official sources call it Ada Lovelace in full.
I think the complaint is more with the consumer card being 4xxx but this is 5000 both on the same architecture.
Quadro RTX 4000
RTX A4000
RTX 4000 Ada
Unfortunately they’ve had 3 separate naming conventions in 3 successive generations. Those 4000 series cards are in the same position in the lineup for each generation.
Add to it the card variants, and there's a chance that you might still end up with the wrong part if your purchaser isn't careful.
They keep removing features
When I tested the A6000 against the H100, there wasn’t that big of a boost from the newer card. Perhaps GPU operations weren’t the bottleneck in that case.
Yes, but the point of a review with benchmarks is that it is expensive and time-consuming for a customer to acquire the hardware just to run their own benchmarks on.
Stable Diffusion and various LLMs are available pretty easily.
A simple benchmark that this version of stable diffusion/llm was used with these settings and this is how long it took to produce image/we got this many tokens/sec would be a nice comparison that you are in a good position to do with access to all the hardware.
And anyways, it’s not that expensive. You can rent the same gpu from cloud providers for a few dollars per hour. If you are serious about buying a GPU it is an extremely small cost in terms of time and money compared to the price of the gpu itself.
The alternative is trusting an online review running some training or inference code which is likely not comparable to what most people are doing.
The big three cloud providers don't offer things like the RTX 5000 32GB or the 4090 do they?
AWS will happily rent me a H100, A100, V100, K80, A10G, T4, or M60. Not to mention a Trainium, Inferentia, Inferentia2, Gaudi or Qualcomm AI 100.
And don't forget to benchmark each one of those in a 1, 2, 4, and 8 GPU configuration; and with a variety of batch sizes; and with and without distributed training. Remember to work out the performance-per-dollar for on-demand, reserved, and spot instances, times three different cloud providers. Now, to compare to on-premise pricing we start by calling our air conditioning and backup generator vendors...
Starting a VM with a cloud provider you already know and use may be extremely fast, but adopting a new cloud provider isn't.
Even if your organisation is so unbureaucratic you can get a new provider set up without any due diligence work, every cloud provider comes with their own oddities. Will they have a wacky set of default firewall rules? Will they only offer Debian not Ubuntu, if you want an instance with the CUDA drivers already installed? Will they insist you learn what a 'provisioned iops' is before letting you start an instance? Will some GPUs only be available in certain regions? Will accounts have quota limits that vary by region, availability zone and GPU type, some of which default to zero?
If you're making, say, an ML-based on-premise CCTV system and you need to run several large ResNets at the same time? And you don't want to go rack-mounted, as some sites don't have a data centre? And you want the longer lifecycle and guaranteed spare parts availability of an enterprise product line? This could be the card for you.
Admittedly it's a rip-off, but the Workstation/Quadro line always has been.
Honestly I'm not sure how healthy the workstation market is right now - with the rise of work-from-home and hybrid working, I don't see many people using huge desktops any more. And when Adobe puts a powerful generative AI feature into Photoshop, they don't expect users to upgrade to powerful GPUs - they run it in the cloud, so it works for users with puny GPUs and Adobe can get that sweet sweet recurring revenue.
"Customer may not [...] provide commercial hosting services with the SOFTWARE. [...] The SOFTWARE is not licensed for datacenter deployment, except that blockchain processing in a datacenter is permitted."
So if you were e.g. an ad agency artist using a 4090 at your desk to generate images for commercial use - that's fine, because it's not in a datacenter, and the commercial services you're providing with it aren't hosting.
I've never heard of nvidia taking any enforcement action, or defining precisely when an office becomes a datacenter - I suspect this is mostly to ensure the big cloud providers don't offer 4090s.
Reading my comment again, obviously it wrong and I should have been more careful with the wording.
The worst thing that can happen is Nvidia declining any warranty repairs on these cards as they've been used outside of the intended use. But it's definitely not a legal issue.
And EULAs in general are just not enforceable outside of the US.