[HN Gopher] Search over 5M+ Stable Diffusion images and prompts
       ___________________________________________________________________
        
       Search over 5M+ Stable Diffusion images and prompts
        
       Author : headalgorithm
       Score  : 163 points
       Date   : 2022-08-26 06:39 UTC (16 hours ago)
        
 (HTM) web link (lexica.art)
 (TXT) w3m dump (lexica.art)
        
       | bravura wrote:
       | How do you submit new images and prompts?
        
         | criddell wrote:
         | https://huggingface.co/spaces/stabilityai/stable-diffusion
        
         | isaacfrond wrote:
         | https://beta.dreamstudio.ai/
        
         | [deleted]
        
       | chrismorgan wrote:
       | Stable Diffusion was finally enough to push me to install drivers
       | for the NVIDIA 3060 on my laptop which has sat completely unused
       | (never powered on once I figured out how to _not_ power it on!)
       | since I got it (I'd have preferred no dGPU at the time, but
       | wanted other features of the laptop that are just about never
       | sold without a fancy dGPU for some reason). Pretty hefty
       | requirements, as a casual layman, even though I know this is
       | smaller and more accessible than just about everything in the
       | past. I think I ended up at around 9GB downloaded (which will
       | cost me almost $2 in concrete terms) and 23GB of disk space used
       | (including things like nvidia-dkms, nvidia-utils, cuda and
       | python-pytorch-opt-cuda; all the relevant Arch packages came to
       | about 14GB).
       | 
       | I'm using https://github.com/basujindal/stable-diffusion to run
       | it since I only have 6GB of VRAM.
       | 
       | I'm having fun. But I haven't had much luck getting it to draw
       | the quick brown fox jumping over the lazy dog; a few steps in
       | there are often the shapes of two animals, but it is consistently
       | reduced to just a fox after a bit more. Extensions to the prompt
       | (like reminding it that there are two animals, and trying to
       | separate the two concepts) can improve it a bit, but it still
       | tends to forget there are two animals, or if it gets two, to draw
       | two foxes, or a dog-fox hybrid and a lazy fox. I imagine I could
       | vastly improve my results with img2img and giving it a basic
       | sketch with placeholders for two distinct animals.
       | 
       | It also has a surprisingly poor idea of what an echidna is.
        
         | moritonal wrote:
         | Is it okay to ask what your situation is that 9GB cost $2?
        
           | GrayShade wrote:
           | Mobile data usage costs?
        
       | tlholaday wrote:
       | The Jack Chick Tract somehow makes me want to use regex to parse
       | html:
       | 
       | https://lexica.art/prompt/64a384bd-d1b2-4f79-8921-f71737c70d...
        
       | immmmmm wrote:
       | Are there diffusion models trained on face datasets only?
        
       | albertzeyer wrote:
       | And yesterday we had the post on OpenArt:
       | https://news.ycombinator.com/item?id=32586439
        
       | cercatrova wrote:
       | I've been using this recently, it works great if you don't know
       | what you want to make.
       | 
       | Also look into CLIP Interrogator, it does image to text
       | basically, turning an image you like into what its prompt could
       | be. However, it won't provide everything for you, just the main
       | description of content.
        
         | tiborsaas wrote:
         | Thanks for the interrogator tip, It's so cool that it works in
         | reverse.
        
       | malshe wrote:
       | Because the title doesn't give the warning- the website is NSFW!
        
         | jxcole wrote:
         | For the curious: it's not intentionally NSFW but the filters
         | aren't very good (if there are any) so you do see occasionally
         | inappropriate images.
        
       | benob wrote:
       | What is the license for those pictures? Neither lexica nor
       | openart talk about it.
        
         | criddell wrote:
         | https://huggingface.co/spaces/CompVis/stable-diffusion-licen...
        
         | wnkrshm wrote:
         | Same license treatment as the training data. Edit: Nobody cares
         | about licenses since they don't want to be asked about how they
         | licensed the training data.
        
       | ilaksh wrote:
       | Where is this coming from? Everything people submit is public?
        
       | pavlov wrote:
       | Feels like shiny concept art for games is facing a similar moment
       | as portrait painting in late 19th century. Why pay someone to
       | paint in this generic commercial style when you can get a
       | meaningful automatic result at the push of a button?
       | 
       | Most other styles of illustration seem safer for the moment
       | because they rely more on the illustrator's personality. (I'm not
       | talking about the kind of stuff you buy on Fiverr, but
       | professional designers who mainly get work through their
       | networks.)
        
         | woojoo666 wrote:
         | What styles of illustration would you say are safe?
        
       | ralfd wrote:
       | > Hyperrealistic mixed media image of matt damon bald head
       | resembles !!uncircumcised penis!!, stunning 3d render inspired
       | art by istvan sandorfi and greg rutkowski, perfect facial
       | symmetry, realistic, highly detailed attributes and atmosphere,
       | dim volumetric cinematic lighting, 8k octane extremely hyper-
       | detailed render, post-processing, masterpiece
       | 
       | Why the penis prompt?
        
         | isaacfrond wrote:
         | Will this get me into trouble?
         | 
         | https://lexica.art/prompt/f472f43f-cc61-4a58-a2a8-f599bb598d...
        
       | Aachen wrote:
       | For anyone else who has no idea what they're looking at, from
       | https://duckduckgo.com/?q=stable+diffusion :
       | 
       | > Stable Diffusion is a text-to-image model [...] It is a
       | breakthrough in speed and quality meaning that it can run on
       | consumer GPUs.
        
       | wdfx wrote:
       | meta: this site is causing all sorts of graphical glitches when
       | scrolling on a pixel 6 android 13 Firefox. There are flickering
       | boxes filled with pixel junk between each item and in the header
        
         | Samin100 wrote:
         | Could you take a screenshot or better yet -- a screen
         | recording? I'm not really sure what might be causing that.
        
           | wdfx wrote:
           | Here's a screenshot https://imgur.com/a/gpStPID
        
             | Samin100 wrote:
             | Ah, that might be because of the backdrop-blur on the
             | navigation bar. Weird that it's causing graphical issues on
             | your phone.
        
       | null_object wrote:
       | I'm thinking if Greg Rutkowski can get himself removed from the
       | AI-prompts half of this art will disappear overnight.
        
         | isaacfrond wrote:
         | He seems very popular in prompts. But how is he? He doesn't
         | even have a wikipedia page.
        
           | whywhywhywhy wrote:
           | Isn't the data set lifted from Artstation for a lot of this
           | stuff?
           | 
           | https://www.artstation.com/Rutkowski He's a fantasy concept
           | artist, not a like wikipedia page having artist.
        
           | MonkeyMalarky wrote:
           | I think its just a meme from users seeing his name in other
           | prompts. In many cases it's used in combinations with other
           | artists whose styles aren't at all similar, it doesn't make
           | sense other than users spamming it to get "good" results?
        
         | ralfd wrote:
         | Here is his Twitter:
         | 
         | https://twitter.com/GrzegorzRutko14/with_replies
         | 
         | But he doesn't seem to comment on ai art.
        
           | corysama wrote:
           | He has given multiple live presentations on the Midjourney
           | Discord server. He's quite happy that his work is helping
           | lots of people make great new art.
        
         | isaacfrond wrote:
         | Page with his work is here:
         | 
         | https://www.artstation.com/Rutkowski
        
         | hulahoof wrote:
         | I was just wondering this, does including his name have the
         | same effect as the trending on art station stuff ?
        
       | dchuk wrote:
       | Is it possible to economically/efficiently run this on a 14" MBP?
       | Or do you need an nvidia graphics card to actually run this
       | thing?
        
         | ManuelKiessling wrote:
         | Takes about 10-20 minutes to create 5 images on my 16" M1 Pro
         | 32 GiB MacBook Pro. It takes around 1 minute on my desktop
         | system using a 6 GiB VRAM RTX 3070 Ti.
        
       | avocado2 wrote:
       | web demo for stable diffusion:
       | https://huggingface.co/spaces/stabilityai/stable-diffusion
       | 
       | can also run it in colab (includes img2img):
       | https://colab.research.google.com/drive/1NfgqublyT_MWtR5Csmr...
       | 
       | and web ui for stable diffusion runs locally (includes
       | gfpgan/realesrgan and alot of other features):
       | https://github.com/hlky/stable-diffusion-webui
        
       | spywaregorilla wrote:
       | I find it hilarious how many of these prompts are using "unreal
       | engine 5" to get a good image.
       | 
       | There's a lot (or honestly maybe a small amount) of work to be
       | done to improve these prompt interfaces. Raw projections of your
       | queries into the embedding space is honestly pretty dumb. Like,
       | it'd be nice if we could start by settling the embeddings into
       | images that are "good".
        
         | addandsubtract wrote:
         | There's a rating feature on the website that let's you rate the
         | results. It's greyed out for me so I'm not sure if it's a timed
         | feature or a premium thing, but it's there.
        
           | spywaregorilla wrote:
           | Not this website. The prompt interface to the embedding
           | model.
           | 
           | There's no reason one should have to spam "good" sounding
           | phrases like "high quality" into the prompt to get a good
           | image. Direct embeddings of the prompt are stupid.
        
             | whywhywhywhy wrote:
             | What people call "prompt engineering" is just knowing to do
             | that.
        
               | spywaregorilla wrote:
               | Yeah and what I'm saying is that this rapidly rising
               | "skill" is just nonsense. This is not a reasonable way to
               | interact with the embedding space. We will not be doing
               | prompt engineering, hopefully within the "near" future.
               | 
               | This is like copy pasting by highlighting, clicking
               | 'edit' and scrolling down to copy/paste all with the
               | mouse.
        
               | galangalalgol wrote:
               | Do dall-e2 and stable diffusion models regularly get
               | retrained? If so, as they get retrained using the ouputs
               | of the models scraped from sites like this will we see
               | some sort of mode collapse?
               | 
               | Is it reasonable for hobbyists to retrain these models
               | with reduced or custom image sets or would that require a
               | lot of money in compute?
        
       | jenny91 wrote:
       | The way it mangles faces is actually super creepy.
       | 
       | These shots from an exit-less, claustrophobic NYC subway with
       | mangled faceless things is the stuff of nightmares:
       | 
       | https://lexica.art/?q=new+york+subway
        
         | ALittleLight wrote:
         | I feel like there needs to be a model that fixes faces to clean
         | this up. Humans are so attuned to faces that I can imagine it
         | would take a specialized model to render convincing faces.
         | Maybe there could be a layer to identify and occlude existing
         | pseudo-faces generated by Stable Diffusion and another model to
         | populate the occlusion.
        
           | meowface wrote:
           | Many are doing exactly this with Stable Diffusion + GFPGAN
           | (https://github.com/TencentARC/GFPGAN) as a post-processing
           | model.
        
           | thomashop wrote:
           | This is already done e.g. in Majesty Diffusion. You run
           | another amount of iterations using a face Restauration GAN
        
         | zamfi wrote:
         | NYC in the '80s was a hard time. This is pretty much in line
         | with my recollections...
        
       | isaacfrond wrote:
       | Dupe:
       | 
       | Discover stable diffusion prompts with Lexica (lexica.art)
       | https://news.ycombinator.com/item?id=32594107
       | 
       | Seems the same idea as openart: OpenArt: "Pinterest" for Dalle-2
       | images and prompts (openart.ai)
       | https://news.ycombinator.com/item?id=32586439
        
       ___________________________________________________________________
       (page generated 2022-08-26 23:02 UTC)