[HN Gopher] Search over 5M+ Stable Diffusion images and prompts
___________________________________________________________________
Search over 5M+ Stable Diffusion images and prompts
Author : headalgorithm
Score : 163 points
Date : 2022-08-26 06:39 UTC (16 hours ago)
(HTM) web link (lexica.art)
(TXT) w3m dump (lexica.art)
| bravura wrote:
| How do you submit new images and prompts?
| criddell wrote:
| https://huggingface.co/spaces/stabilityai/stable-diffusion
| isaacfrond wrote:
| https://beta.dreamstudio.ai/
| [deleted]
| chrismorgan wrote:
| Stable Diffusion was finally enough to push me to install drivers
| for the NVIDIA 3060 on my laptop which has sat completely unused
| (never powered on once I figured out how to _not_ power it on!)
| since I got it (I'd have preferred no dGPU at the time, but
| wanted other features of the laptop that are just about never
| sold without a fancy dGPU for some reason). Pretty hefty
| requirements, as a casual layman, even though I know this is
| smaller and more accessible than just about everything in the
| past. I think I ended up at around 9GB downloaded (which will
| cost me almost $2 in concrete terms) and 23GB of disk space used
| (including things like nvidia-dkms, nvidia-utils, cuda and
| python-pytorch-opt-cuda; all the relevant Arch packages came to
| about 14GB).
|
| I'm using https://github.com/basujindal/stable-diffusion to run
| it since I only have 6GB of VRAM.
|
| I'm having fun. But I haven't had much luck getting it to draw
| the quick brown fox jumping over the lazy dog; a few steps in
| there are often the shapes of two animals, but it is consistently
| reduced to just a fox after a bit more. Extensions to the prompt
| (like reminding it that there are two animals, and trying to
| separate the two concepts) can improve it a bit, but it still
| tends to forget there are two animals, or if it gets two, to draw
| two foxes, or a dog-fox hybrid and a lazy fox. I imagine I could
| vastly improve my results with img2img and giving it a basic
| sketch with placeholders for two distinct animals.
|
| It also has a surprisingly poor idea of what an echidna is.
| moritonal wrote:
| Is it okay to ask what your situation is that 9GB cost $2?
| GrayShade wrote:
| Mobile data usage costs?
| tlholaday wrote:
| The Jack Chick Tract somehow makes me want to use regex to parse
| html:
|
| https://lexica.art/prompt/64a384bd-d1b2-4f79-8921-f71737c70d...
| immmmmm wrote:
| Are there diffusion models trained on face datasets only?
| albertzeyer wrote:
| And yesterday we had the post on OpenArt:
| https://news.ycombinator.com/item?id=32586439
| cercatrova wrote:
| I've been using this recently, it works great if you don't know
| what you want to make.
|
| Also look into CLIP Interrogator, it does image to text
| basically, turning an image you like into what its prompt could
| be. However, it won't provide everything for you, just the main
| description of content.
| tiborsaas wrote:
| Thanks for the interrogator tip, It's so cool that it works in
| reverse.
| malshe wrote:
| Because the title doesn't give the warning- the website is NSFW!
| jxcole wrote:
| For the curious: it's not intentionally NSFW but the filters
| aren't very good (if there are any) so you do see occasionally
| inappropriate images.
| benob wrote:
| What is the license for those pictures? Neither lexica nor
| openart talk about it.
| criddell wrote:
| https://huggingface.co/spaces/CompVis/stable-diffusion-licen...
| wnkrshm wrote:
| Same license treatment as the training data. Edit: Nobody cares
| about licenses since they don't want to be asked about how they
| licensed the training data.
| ilaksh wrote:
| Where is this coming from? Everything people submit is public?
| pavlov wrote:
| Feels like shiny concept art for games is facing a similar moment
| as portrait painting in late 19th century. Why pay someone to
| paint in this generic commercial style when you can get a
| meaningful automatic result at the push of a button?
|
| Most other styles of illustration seem safer for the moment
| because they rely more on the illustrator's personality. (I'm not
| talking about the kind of stuff you buy on Fiverr, but
| professional designers who mainly get work through their
| networks.)
| woojoo666 wrote:
| What styles of illustration would you say are safe?
| ralfd wrote:
| > Hyperrealistic mixed media image of matt damon bald head
| resembles !!uncircumcised penis!!, stunning 3d render inspired
| art by istvan sandorfi and greg rutkowski, perfect facial
| symmetry, realistic, highly detailed attributes and atmosphere,
| dim volumetric cinematic lighting, 8k octane extremely hyper-
| detailed render, post-processing, masterpiece
|
| Why the penis prompt?
| isaacfrond wrote:
| Will this get me into trouble?
|
| https://lexica.art/prompt/f472f43f-cc61-4a58-a2a8-f599bb598d...
| Aachen wrote:
| For anyone else who has no idea what they're looking at, from
| https://duckduckgo.com/?q=stable+diffusion :
|
| > Stable Diffusion is a text-to-image model [...] It is a
| breakthrough in speed and quality meaning that it can run on
| consumer GPUs.
| wdfx wrote:
| meta: this site is causing all sorts of graphical glitches when
| scrolling on a pixel 6 android 13 Firefox. There are flickering
| boxes filled with pixel junk between each item and in the header
| Samin100 wrote:
| Could you take a screenshot or better yet -- a screen
| recording? I'm not really sure what might be causing that.
| wdfx wrote:
| Here's a screenshot https://imgur.com/a/gpStPID
| Samin100 wrote:
| Ah, that might be because of the backdrop-blur on the
| navigation bar. Weird that it's causing graphical issues on
| your phone.
| null_object wrote:
| I'm thinking if Greg Rutkowski can get himself removed from the
| AI-prompts half of this art will disappear overnight.
| isaacfrond wrote:
| He seems very popular in prompts. But how is he? He doesn't
| even have a wikipedia page.
| whywhywhywhy wrote:
| Isn't the data set lifted from Artstation for a lot of this
| stuff?
|
| https://www.artstation.com/Rutkowski He's a fantasy concept
| artist, not a like wikipedia page having artist.
| MonkeyMalarky wrote:
| I think its just a meme from users seeing his name in other
| prompts. In many cases it's used in combinations with other
| artists whose styles aren't at all similar, it doesn't make
| sense other than users spamming it to get "good" results?
| ralfd wrote:
| Here is his Twitter:
|
| https://twitter.com/GrzegorzRutko14/with_replies
|
| But he doesn't seem to comment on ai art.
| corysama wrote:
| He has given multiple live presentations on the Midjourney
| Discord server. He's quite happy that his work is helping
| lots of people make great new art.
| isaacfrond wrote:
| Page with his work is here:
|
| https://www.artstation.com/Rutkowski
| hulahoof wrote:
| I was just wondering this, does including his name have the
| same effect as the trending on art station stuff ?
| dchuk wrote:
| Is it possible to economically/efficiently run this on a 14" MBP?
| Or do you need an nvidia graphics card to actually run this
| thing?
| ManuelKiessling wrote:
| Takes about 10-20 minutes to create 5 images on my 16" M1 Pro
| 32 GiB MacBook Pro. It takes around 1 minute on my desktop
| system using a 6 GiB VRAM RTX 3070 Ti.
| avocado2 wrote:
| web demo for stable diffusion:
| https://huggingface.co/spaces/stabilityai/stable-diffusion
|
| can also run it in colab (includes img2img):
| https://colab.research.google.com/drive/1NfgqublyT_MWtR5Csmr...
|
| and web ui for stable diffusion runs locally (includes
| gfpgan/realesrgan and alot of other features):
| https://github.com/hlky/stable-diffusion-webui
| spywaregorilla wrote:
| I find it hilarious how many of these prompts are using "unreal
| engine 5" to get a good image.
|
| There's a lot (or honestly maybe a small amount) of work to be
| done to improve these prompt interfaces. Raw projections of your
| queries into the embedding space is honestly pretty dumb. Like,
| it'd be nice if we could start by settling the embeddings into
| images that are "good".
| addandsubtract wrote:
| There's a rating feature on the website that let's you rate the
| results. It's greyed out for me so I'm not sure if it's a timed
| feature or a premium thing, but it's there.
| spywaregorilla wrote:
| Not this website. The prompt interface to the embedding
| model.
|
| There's no reason one should have to spam "good" sounding
| phrases like "high quality" into the prompt to get a good
| image. Direct embeddings of the prompt are stupid.
| whywhywhywhy wrote:
| What people call "prompt engineering" is just knowing to do
| that.
| spywaregorilla wrote:
| Yeah and what I'm saying is that this rapidly rising
| "skill" is just nonsense. This is not a reasonable way to
| interact with the embedding space. We will not be doing
| prompt engineering, hopefully within the "near" future.
|
| This is like copy pasting by highlighting, clicking
| 'edit' and scrolling down to copy/paste all with the
| mouse.
| galangalalgol wrote:
| Do dall-e2 and stable diffusion models regularly get
| retrained? If so, as they get retrained using the ouputs
| of the models scraped from sites like this will we see
| some sort of mode collapse?
|
| Is it reasonable for hobbyists to retrain these models
| with reduced or custom image sets or would that require a
| lot of money in compute?
| jenny91 wrote:
| The way it mangles faces is actually super creepy.
|
| These shots from an exit-less, claustrophobic NYC subway with
| mangled faceless things is the stuff of nightmares:
|
| https://lexica.art/?q=new+york+subway
| ALittleLight wrote:
| I feel like there needs to be a model that fixes faces to clean
| this up. Humans are so attuned to faces that I can imagine it
| would take a specialized model to render convincing faces.
| Maybe there could be a layer to identify and occlude existing
| pseudo-faces generated by Stable Diffusion and another model to
| populate the occlusion.
| meowface wrote:
| Many are doing exactly this with Stable Diffusion + GFPGAN
| (https://github.com/TencentARC/GFPGAN) as a post-processing
| model.
| thomashop wrote:
| This is already done e.g. in Majesty Diffusion. You run
| another amount of iterations using a face Restauration GAN
| zamfi wrote:
| NYC in the '80s was a hard time. This is pretty much in line
| with my recollections...
| isaacfrond wrote:
| Dupe:
|
| Discover stable diffusion prompts with Lexica (lexica.art)
| https://news.ycombinator.com/item?id=32594107
|
| Seems the same idea as openart: OpenArt: "Pinterest" for Dalle-2
| images and prompts (openart.ai)
| https://news.ycombinator.com/item?id=32586439
___________________________________________________________________
(page generated 2022-08-26 23:02 UTC)