WebGPU GPT Model Demo
kmeans.org
kmeans.org
Also instead of just "Update Chrome to v113" the domain owner could sign up for an origin trial https://developer.chrome.com/origintrials/#/view_trial/11821...
"Find an error initializing the WebGPU device Error: Cannot initialize runtime because of requested maxBufferSize exceeds limit. requested=1024MB, limit=256MB. This error may be caused by an older version of the browser (e.g. Chrome 112). You can try to upgrade your browser to Chrome 113 or later."
Releasing April 26th when Chrome 113 hits stable. Open source NPM library you can add to any project.
Preview here: https://twitter.com/fleetwood___/status/1646608499126816799?...
> WebGPU is supported in your browser!
> Uncaught (in promise) DOMException: WebGPU is not yet available in Release or Beta builds.
Anyone using Chromium care to chime in?If no one chimes in I might set up a Chromium browser up just to take a look at this, seems pretty cool.
Could this code also be used to train models or only for inference?
What I'm getting at, is could I take the WGSL and using rust wgpu create a mini ChatGPT that runs on all GPU's?
How do ChatGPT on GPT-3.5 / GPT-4 compare?
Thats slower than Vicuna 7B (aka LLaMa 7000M, GPT 3ish? model) on linux on the same machine, where I get about 3.5 tokens/sec and 97% usage. So... yeah, performance is not so great yet.
Another annoying constraint but specific to wgpu (Rust's implementation of WebGPU) is that it does not support f16 yet (which IS in the spec), only through SPIR-V passthrough...
google-chrome --enable-unsafe-webgpu --enable-features=Vulkan,UseSkiaRenderer --enable-dawn-features=disable_robustness
GPU doesn't work in --headless though.