Janus Pro 1B running 100% locally in-browser on WebGPU
old.reddit.com
old.reddit.com
> 860M UNet and 123M text encoder
[0] https://github.com/CompVis/stable-diffusion/blob/main/README...
Comparisons with other SoTA (Flux, Imagen, etc):
https://imgur.com/a/janus-flux-imagen3-dall-e-3-comparisons-...
Also it looks like octopuses are suffering the “six finger hand” syndrome with their arms from all models.
I bet there has been a lot of testing that what looks "by default" much more attractive for the general people. It is also a selling point, when low effort produces something visually amazing.
Just three years ago this would have been world-changing.
Why JanusPro? Decoupled Visual Encoding: Separates image generation/understanding pathways, eliminating role conflicts in visual processing while maintaining a unified backbone 2.
Hardware Agnostic: Runs efficiently on consumer GPUs (even AMD cards), with users reporting 30% faster inference vs. NVIDIA equivalents 2.
Ethical Safeguards: Open-source license restricts military/illegal use, aligning with responsible AI development
please checkout the website: https://januspro-ai.com/
The burger looks good.
a 1B model should be able to run in the RAM constraints of a phone(?) if this is supported soon this would actually be wild. Local LLMs in the palm of your hands
Local LLMs in the palm of your hands
https://apps.apple.com/us/app/mlc-chat/id6448482937Q:These models running in WebGPU all seem to need nodejs installed. I that for just the local 'server side', can you not just use a python http server or tomcat for this and wget files?