While LLM and agents dominate the discourse, folks didn't realize that open Source has either caught u, approached or surpassed quality in voice tasks (text-to-speech, transcription, sound effect generation, etc). This demo running on Hugging Face helps recalibrate that
Probably the cheapest UI, by hacking a bit Hugging Face Spaces. Also works locally. Of course Google Colab is still free but I think the UI / pre-curation of hyperparameters may be worth it?
Hi, I'm the creator of multimodal.art, I didn't overlook it, but there's no "specialized" NSFW content maker to be highlighted - this Vice articles just show people using the model in different iterations to generate NSFW content; you don't need a specialized notebook/tool for that, a few ones on the post can do it (others have a NSFW filter that comes in by default).
Additionally it is important to note that model was licensed under the OpenRAIL-M LICENSE which is not as permissive as an MIT license and forbids certain outputs to be shared or purposes to be built as apps
Hi everyone! I am the creator of this website multimodal.art. Our goal is to break the barries of entry for this tech, as well as to inform people about the potentil of that technology. We also developed MindsEye, an open source pilot to many AI art models (VQGAN+CLIP, Guided Diffusion and Latent Diffusion) https://multimodal.art/mindseye