# Digital art creator algorithm website

**URL:** <https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081>\
**Category:** Cafe Society\
**Tags:** arts-crafts, ai\
**Created:** [April 1, 2022, 1:26am UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081 "2022-04-01T01:26:49Z")\
**Posts on this page:** 20\
**Page:** 91

<div class="post-metadata">

**Author:** ![Darren\_Garrison](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/darren_garrison/32/92_2.png) [@Darren\_Garrison](https://boards.straightdope.com/u/Darren_Garrison)\
**Post date:** [November 22, 2023, 11:15am UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1803 "2023-11-22T11:15:32Z")

</div>

Nice. But you can download the image without the watermark in the corner.

---

<div class="post-metadata">

**Author:** ![Darren\_Garrison](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/darren_garrison/32/92_2.png) [@Darren\_Garrison](https://boards.straightdope.com/u/Darren_Garrison)\
**Post date:** [November 22, 2023, 5:47pm UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1804 "2023-11-22T17:47:44Z")

</div>

Stable Diffusion video model released.

> **[stabilityai/stable-video-diffusion-img2vid · Hugging Face](https://huggingface.co/stabilityai/stable-video-diffusion-img2vid)**
>
> We’re on a journey to advance and democratize artificial intelligence through open source and open science.

---

<div class="post-metadata">

**Author:** ![Dr.Strangelove](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/dr.strangelove/32/6613_2.png) [@Dr.Strangelove](https://boards.straightdope.com/u/Dr.Strangelove)\
**Post date:** [November 23, 2023, 5:03am UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1805 "2023-11-23T05:03:09Z")

</div>

I downloaded the tensor weights for Stable Video Diffusion the other day, but I use ComfyUI as it’s the only front end I’ve found that totally suck, and they don’t seem to have added support yet. Hopefully soon.

---

<div class="post-metadata">

**Author:** ![Darren\_Garrison](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/darren_garrison/32/92_2.png) [@Darren\_Garrison](https://boards.straightdope.com/u/Darren_Garrison)\
**Post date:** [November 23, 2023, 4:42pm UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1806 "2023-11-23T16:42:04Z")

</div>

Image-generating adjacent, a lawsuit against LLMs has been largely tossed out.

> **[Sarah Silverman suffers legal blow](https://www.newsweek.com/sarah-silverman-lawsuit-meta-ai-1846340)**
>
> The comedian attempted to sue Meta over using her books to train its AI bots, but the judge was not convinced.

---

<div class="post-metadata">

**Author:** ![Jophiel](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/jophiel/32/66_2.png) [@Jophiel](https://boards.straightdope.com/u/Jophiel)\
**Post date:** [November 23, 2023, 4:45pm UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1807 "2023-11-23T16:45:05Z")

</div>

> [@Dr.Strangelove](#):
>
> I use ComfyUI as it’s the only front end I’ve found that totally suck

Bold choice to use the only front end that totally sucks 😉

---

<div class="post-metadata">

**Author:** ![Darren\_Garrison](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/darren_garrison/32/92_2.png) [@Darren\_Garrison](https://boards.straightdope.com/u/Darren_Garrison)\
**Post date:** [November 24, 2023, 1:47am UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1808 "2023-11-24T01:47:27Z")

</div>

[![](https://i.postimg.cc/y6crNB4Y/f9321241-25ac-4e89-a8a1-36841bbc2332.jpg) ](https://i.postimg.cc/y6crNB4Y/f9321241-25ac-4e89-a8a1-36841bbc2332.jpg)

---

<div class="post-metadata">

**Author:** ![Dr.Strangelove](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/dr.strangelove/32/6613_2.png) [@Dr.Strangelove](https://boards.straightdope.com/u/Dr.Strangelove)\
**Post date:** [November 24, 2023, 8:50am UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1809 "2023-11-24T08:50:52Z")

</div>

> [@Jophiel](#):
>
> Bold choice to use the only front end that totally sucks 😉

I’m a glutton for punishment, apparently. But seriously, ComfyUI is great for running local models. And it looks like support for Stable Video Diffusion just came out:

> **[GitHub - thecooltechguy/ComfyUI-Stable-Video-Diffusion: ComfyUI nodes for...](https://github.com/thecooltechguy/ComfyUI-Stable-Video-Diffusion)**
>
> ComfyUI nodes for Stable Video Diffusion. Contribute to thecooltechguy/ComfyUI-Stable-Video-Diffusion development by creating an account on GitHub.

I’m visiting the parents, though, so I won’t be able to try it for a few days.

---

<div class="post-metadata">

**Author:** ![Chronos](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/chronos/32/134_2.png) [@Chronos](https://boards.straightdope.com/u/Chronos)\
**Post date:** [November 24, 2023, 4:37pm UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1810 "2023-11-24T16:37:18Z")

</div>

Weird. The popcorn, pretzels, and jellybeans look great, the bird’s only problem is being implausibly still, and the dog is pretty good aside from being out of focus, but what the heck is going on with that toast? I’d think that any model that can do the other things well would also be able to handle toast, or at least do a better job of it than that.

---

<div class="post-metadata">

**Author:** ![Darren\_Garrison](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/darren_garrison/32/92_2.png) [@Darren\_Garrison](https://boards.straightdope.com/u/Darren_Garrison)\
**Post date:** [November 24, 2023, 4:44pm UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1811 "2023-11-24T16:44:00Z")

</div>

I had better toasts, but many images had a dog/table merger going on.

---

<div class="post-metadata">

**Author:** ![Darren\_Garrison](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/darren_garrison/32/92_2.png) [@Darren\_Garrison](https://boards.straightdope.com/u/Darren_Garrison)\
**Post date:** [November 28, 2023, 4:20am UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1812 "2023-11-28T04:20:07Z")

</div>

Hugging Face has a usable version of Stable Diffusion image to video now.

> **[Stable Video Diffusion - a Hugging Face Space by multimodalart](https://huggingface.co/spaces/multimodalart/stable-video-diffusion)**
>
> Discover amazing ML apps made by the community

I made one video last night, but spent more than 12 minutes in a queue waiting for it to run. Here’s the result:

[![](https://img.youtube.com/vi/gHQEKke0EMw/hqdefault.jpg "Stable Diffusion image to video test") ](https://www.youtube.com/watch?v=gHQEKke0EMw)

For comparison, here’s what I got with RunwayML Gen2 when that first released:

[![](https://img.youtube.com/vi/CtMooiPk-q8/maxresdefault.jpg "Runway image to video test") ](https://www.youtube.com/watch?v=CtMooiPk-q8)

(Here’s the source image, made with Bing/DE3):

[![](https://i.postimg.cc/Njx7DhP0/d91d123b-611f-4ff9-a81e-15f4d0b60492-1.jpg) ](https://i.postimg.cc/Njx7DhP0/d91d123b-611f-4ff9-a81e-15f4d0b60492-1.jpg)

---

<div class="post-metadata">

**Author:** ![Jophiel](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/jophiel/32/66_2.png) [@Jophiel](https://boards.straightdope.com/u/Jophiel)\
**Post date:** [November 28, 2023, 10:51pm UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1813 "2023-11-28T22:51:06Z")

</div>

Stable Diffusion released an early test of SDXL Turbo, a “one step” model. “Steps” being how many passes it makes when it renders an image, not how many tasks you need to complete to use it. A photorealistic image in Stable Diffusion is usually around 30-50 steps.

As a result, it was completing _two images per second_ on my RTX 3080Ti when I tested it. That’s hecka fast. It does have some major limitations though – it doesn’t really do photorealism, photoreal human faces are a disaster, it’s meant to run at 512x512 and adding more steps immediately blows the image out. It’s really just one step and setting CFG to 1 as well. It is a really cool hint at how the tech is progressing though and I tried a bunch of stuff and had some “great” results when you remember: two images per second.

[![](https://i.imgur.com/HulqCMO.png) ](https://i.imgur.com/HulqCMO.png)  
[![](https://i.imgur.com/ziANTXL.png) ](https://i.imgur.com/ziANTXL.png)  
[![](https://i.imgur.com/gxBWNsa.png) ](https://i.imgur.com/gxBWNsa.png)  
[![](https://i.imgur.com/nZOP7WF.png) ](https://i.imgur.com/nZOP7WF.png)  
[![](https://i.imgur.com/8NaelYE.png) ](https://i.imgur.com/8NaelYE.png)

(Also, results are from a single run using A1111. I believe that Comfy might do it better since I wasn’t doing the second Refiner pass and for various other boring technical reasons)

---

<div class="post-metadata">

**Author:** ![Darren\_Garrison](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/darren_garrison/32/92_2.png) [@Darren\_Garrison](https://boards.straightdope.com/u/Darren_Garrison)\
**Post date:** [November 29, 2023, 12:16am UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1814 "2023-11-29T00:16:42Z")

</div>

On clipdrop:

> **[Clipdrop - SDXL Turbo](https://clipdrop.co/stable-diffusion-turbo)**
>
> Clipdrop - SDXL Turbo

---

<div class="post-metadata">

**Author:** ![Chronos](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/chronos/32/134_2.png) [@Chronos](https://boards.straightdope.com/u/Chronos)\
**Post date:** [November 29, 2023, 12:55am UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1815 "2023-11-29T00:55:18Z")

</div>

OK, those chesscrapers are just plain cool. A human artist probably wouldn’t have had too much difficulty making that, either… given the concept. But it’s a great concept.

@Darren_Garrison , I think the possum is probably happier in the RunwayML version. It’s still getting a creepyhand stuck to its nape, but its head isn’t dematerializing.

---

<div class="post-metadata">

**Author:** ![Darren\_Garrison](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/darren_garrison/32/92_2.png) [@Darren\_Garrison](https://boards.straightdope.com/u/Darren_Garrison)\
**Post date:** [November 29, 2023, 1:27am UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1816 "2023-11-29T01:27:04Z")

</div>

Since I’ve set up a new Youtube account, here are the more coherent clips that I generated with my limited free Runway Gen2 seconds a while back.

[![](https://img.youtube.com/vi/Oe-VM-bgI1U/maxresdefault.jpg "RunwayML Gen2 test clip compilation") ](https://www.youtube.com/watch?v=Oe-VM-bgI1U)

---

<div class="post-metadata">

**Author:** ![Jophiel](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/jophiel/32/66_2.png) [@Jophiel](https://boards.straightdope.com/u/Jophiel)\
**Post date:** [November 29, 2023, 1:40am UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1817 "2023-11-29T01:40:37Z")

</div>

Being able to run 800 images in just under 5min (4:58) is addicting 😃

Was worried for my SSD but those 800 images only take up 280MB of space at 512x512

---

<div class="post-metadata">

**Author:** ![Pleonast](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/pleonast/32/1183_2.png) [@Pleonast](https://boards.straightdope.com/u/Pleonast)\
**Post date:** [November 29, 2023, 3:10am UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1818 "2023-11-29T03:10:19Z")

</div>

> [@Jophiel](#):
>
> Was worried for my SSD but those 800 images only take up 280MB of space at 512x512

I’ve set my output to a RAM drive, and then only move the good ones to the permanent disk.

---

<div class="post-metadata">

**Author:** ![Darren\_Garrison](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/darren_garrison/32/92_2.png) [@Darren\_Garrison](https://boards.straightdope.com/u/Darren_Garrison)\
**Post date:** [November 29, 2023, 10:08pm UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1819 "2023-11-29T22:08:55Z")

</div>

Another day, another new toy.

> **[This new AI animation tool is blowing people's minds | Digital Trends](https://www.digitaltrends.com/computing/runway-motion-brush-animation-tool/)**
>
> The AI research company, Runway lets you beef up your generative AI images with its new Motion Brush that's a part of its latest update.

---

<div class="post-metadata">

**Author:** ![Darren\_Garrison](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/darren_garrison/32/92_2.png) [@Darren\_Garrison](https://boards.straightdope.com/u/Darren_Garrison)\
**Post date:** [December 2, 2023, 6:03pm UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1820 "2023-12-02T18:03:35Z")

</div>

Wanted a Mandalorian on a DeLorean. For some reason Bing/DE3 keeps thinking that concept should include a cat or small dog and some brown leather bags.

> **[Bing](https://www.bing.com/images/create/photo-of-boba-fett-sitting-on-a-delorean/1-656b6b3ebb3f4fa59eba5ff1564018f7?id=ywEs252uMIYkYJDfJ7UTLA%3d%3d&view=detailv2&idpp=genimg&FORM=GCRIDP&mode=overlay)**
>
> Intelligent search from Bing makes it easier to quickly find what you’re looking for and rewards you.

> **[Bing](https://www.bing.com/images/create/photo-of-boba-fett-sitting-on-a-delorean-3d-model/1-656b6b64dfb341eca0bbc7b0e77d8d0e?id=s6kNZbl3KZ8bCGmWgHKpKQ%3d%3d&view=detailv2&idpp=genimg&FORM=GCRIDP&mode=overlay)**
>
> Intelligent search from Bing makes it easier to quickly find what you’re looking for and rewards you.

> **[Bing](https://www.bing.com/images/create/photo-of-boba-fett-sitting-on-a-delorean-3d-model/1-656b6bf60cba469e925d6804a1ebd053?id=mJgPUvDtrNNrmO8X7HZE0Q%3d%3d&view=detailv2&idpp=genimg&FORM=GCRIDP&mode=overlay)**
>
> Intelligent search from Bing makes it easier to quickly find what you’re looking for and rewards you.

> **[Bing](https://www.bing.com/images/create/photo-of-boba-fett-sitting-on-a-delorean-from-back/1-656b6c5c3e294070bc3ba8ce34e5e4a6?id=x7C%2bVszT8tQdQbSAKoXgvw%3d%3d&view=detailv2&idpp=genimg&FORM=GCRIDP&mode=overlay)**
>
> Intelligent search from Bing makes it easier to quickly find what you’re looking for and rewards you.

> **[Bing](https://www.bing.com/images/create/photo-of-boba-fett-sitting-on-a-delorean-by-junji-/1-656b70ec96c44423a98ed9b554a37e1a?id=z4OdNC7XdPlv5dX%2b8hXxdA%3d%3d&view=detailv2&idpp=genimg&FORM=GCRIDP&mode=overlay)**
>
> Intelligent search from Bing makes it easier to quickly find what you’re looking for and rewards you.

---

<div class="post-metadata">

**Author:** ![Darren\_Garrison](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/darren_garrison/32/92_2.png) [@Darren\_Garrison](https://boards.straightdope.com/u/Darren_Garrison)\
**Post date:** [December 3, 2023, 3:29am UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1821 "2023-12-03T03:29:30Z")

</div>

I’ve been playing around a lot with Stable Video Diffusion. It is a long way from perfected, but what is already there is impressive.

It creates video clips from images without a prompt, and needs to try to recognize the objects in the image but also what type of motion makes sense. A few videos are failures, with little or even no motion. A few are very simple pans across a static image (essentially the “Ken Burns effect”) but with new data created for the edges as the “camera” moves.

The zooms in or out are more sophisticated, moving the elements of the images around relative to each other in an awareness of parallax and that the image contains distinct objects layered in a three-dimensional environment, with the software very accurately determining the edges of individual objects. Some videos involve rotating “inside” the image, understanding the three dimensional nature of the objects and creating new “structure” on the objects as they rotate.

With some images the software recognizes specific types of objects and tries to give them the appropriate type of movement. It can be environmental like billowing clouds and crashing waves or mechanical like turning wheels. It also can often recognize humans and animals and attempt to move their limbs and faces in appropriate ways.

Stable Video Diffusion does not do any of these things perfectly, and makes major mistakes. But this is an experimental early release of the first version of the software, and could (like still image generating) improve rapidly.

SVD creates a set of 24 images, allocated as 6 frames per second for 4 seconds. It seems to me like they could make it an indefinite duration and not RAM-limited, basing each new frame off a number of previous frames, but currently it seems to keep all frames in memory all once and reference all of them–some of the video clips form near-seamless loops, and some loose detail mid-video only to regain them before the end.

This is a compilation of 75 if the 4 second clips, all generated from images I made using Dall-E or Stable Diffusion.

[![](https://img.youtube.com/vi/FAZUe4LfmUU/maxresdefault.jpg "Stable Video Diffusion sample clips") ](https://www.youtube.com/watch?v=FAZUe4LfmUU)

---

<div class="post-metadata">

**Author:** ![Dr.Strangelove](https://sea3.discourse-cdn.com/straightdope/user_avatar/boards.straightdope.com/dr.strangelove/32/6613_2.png) [@Dr.Strangelove](https://boards.straightdope.com/u/Dr.Strangelove)\
**Post date:** [December 3, 2023, 3:34am UTC](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081/1822 "2023-12-03T03:34:04Z")

</div>

> [@Darren\_Garrison](#):
>
> For some reason Bing/DE3 keeps thinking that concept should include a cat or small dog and some brown leather bags.

I think one of those was a small Wookiee. Also, it really likes making a Back to the Future DeLorean with the extra wires and stuff on the outside.

[Previous page](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081.md?page=90)

[Next page](https://boards.straightdope.com/t/digital-art-creator-algorithm-website/962081.md?page=92)
