107 karma · joined May 22, 2025
https://github.com/MoonshotAI/Kimi-K2/blob/main/docs/deploy_...
"The smallest deployment unit for Kimi-K2 FP8 weights with 128k seqlen on mainstream H200 or H20 platform is a cluster with 16 GPUs with either Tensor Parallel (TP) or "data parallel + expert parallel" (DP+EP)."
16 GPUs costing ~$30k each. No one is running a ~$500k server at home.
This sentence is 100% AI generated: "It’s a symbiotic cycle: retro inspires, modern sustains. Commodore isn’t returning. It’s evolving, with purpose."
Frivolous flagging - as you are doing - could eventually get your account privileges removed.
I use Grok, ChatGPT, and Gemini. They are all excellent, state of the art, and have their unique strengths and weaknesses.
"b. Disclaimer. PRE-GA OFFERINGS ARE PROVIDED “AS IS” WITHOUT ANY EXPRESS OR IMPLIED WARRANTIES OR REPRESENTATIONS OF ANY KIND. Pre-GA Offerings (i) may be changed, suspended or discontinued at any time without prior notice to Customer and (ii) are not covered by any SLA or Google indemnity. Except as otherwise expressly indicated in a written notice or Google documentation, (A) Pre-GA Offerings are not covered by TSS, and (B) the Data Location Section above will not apply to Pre-GA Offerings."
I am done with this thread. We are going around in circles.
And I mean I genuinely do not understand what you are trying to say. Couldn't parse it.
It's a preview model - for testing only, not for production. Really not that complicated.
https://cloud.google.com/products?hl=en#product-launch-stage...
"At Preview, products or features are ready for testing by customers. Preview offerings are often publicly announced, but are not necessarily feature-complete, and no SLAs or technical support commitments are provided for these. Unless stated otherwise by Google, Preview offerings are intended for use in test environments only. The average Preview stage lasts about six months."
The Plasma web search plugin includes ~100 search engines / services OOTB.
In terms of desktop environment features there is no comparison. The variety of options and widgets in Linux DE’s extends far beyond the basic shit available in Windows.
It took Windows several decades to get virtual desktops (Windows 11). Linux had that in 1993.
They are busy doing their work and prefer their competitors (other developers) to not use these tools.
Those predictive text systems are usually Markov models. LLMs are fundamentally different. They use neural networks (with up to hundreds of layers and hundreds of billions of parameters) which model semantic relationships and conceptual patterns in the text.
Here's something interesting from the conclusion of the paper:
"An interesting and promising direction for future work that leverages the inherent differentiability, would be to apply RenderFormer to inverse rendering applications."
That means generate a 3D scene from 2D images.