HNHacker News
TopNewBestAskShowJobs

polisteps

150 karma · joined November 2, 2020

submissionscomments
polisteps··on Jev Reproductions Tracker
tracking all jev reproduction efforts
polisteps··on ZeroLabs – 100x cheaper than ElevenLabs (free forever locally) with open models
While LLM and agents dominate the discourse, folks didn't realize that open Source has either caught u, approached or surpassed quality in voice tasks (text-to-speech, transcription, sound effect generation, etc). This demo running on Hugging Face helps recalibrate that
polisteps··on LoRA Roulette
Two random LoRAs are loaded into SDXL, can you find a way to combine them?
polisteps··on Karlo, the first open source DALL-E 2 replication is here
GitHub: https://github.com/kakaobrain/karlo

diffusers lib integration: https://github.com/huggingface/diffusers/releases/tag/v0.11....

polisteps··on Dreambooth training UI for training a model for less than US$0.80
Probably the cheapest UI, by hacking a bit Hugging Face Spaces. Also works locally. Of course Google Colab is still free but I think the UI / pre-curation of hyperparameters may be worth it?
polisteps··on Diffusers 0.7.0: Stable Diffusion up to 2x faster on any hardware (including M1)
Pipelines also load faster, more schedulers are out, and the community contributed a bunch!
polisteps··on The Ethereum merge is done
Which is fairer imo than paying for the whole site. I prefer creator control on which content is paid and which content is not.
polisteps··on 1 week of Stable Diffusion
Hi, I'm the creator of multimodal.art, I didn't overlook it, but there's no "specialized" NSFW content maker to be highlighted - this Vice articles just show people using the model in different iterations to generate NSFW content; you don't need a specialized notebook/tool for that, a few ones on the post can do it (others have a NSFW filter that comes in by default).

Additionally it is important to note that model was licensed under the OpenRAIL-M LICENSE which is not as permissive as an MIT license and forbids certain outputs to be shared or purposes to be built as apps

polisteps··on Multimodal Art
Hi everyone! I am the creator of this website multimodal.art. Our goal is to break the barries of entry for this tech, as well as to inform people about the potentil of that technology. We also developed MindsEye, an open source pilot to many AI art models (VQGAN+CLIP, Guided Diffusion and Latent Diffusion) https://multimodal.art/mindseye
polisteps··on Rendering photo-realistic glass in the browser
Who else read the title and though it would render photo realistic glasses as filters in your face? :-)