Hi all. I have been experimenting with GAN algorithms for a while, trying different configuration combinations. Some approaches produce astonishing results.
There are tons of parameters that can affect the result when working with Text-to-Image systems. Apart from the algorithmic structure, there is no limit to what can be done even in end-user-friendly compiled notebooks.
I used the CLIP Guided Diffusion notebook and Gigapixel AI software to create the above work that I tokenized in the NFT Showroom. The concept of prompt is important in such experiments, but I think another important issue is the initial image. Configuration settings made in the initial image significantly increase the approach to the desired result. I want to share init_image that I used in that work process.
Yes, this is a normal human skull. So why did I use it?
If there is balanced complexity in a given area, it is not usually as confusing as unbalanced complexity. Actually, what I meant was that my AI supported portrait experiments were constantly unbalanced. This is technically directly related to the perceptual loss situation in GAN algorithms. To put it simply, in order for the portrait image to be close to the desired image in shape, the more the noise it will use at the beginning is suitable to give the final image, the more satisfactory the output will be in the right proportion.
If you are experimenting AI supported generative art and want to make a portrait, I recommend using a skull as an init_image. The results are generally satisfactory after image enhancement.
Have a nice day!
TR
Merhaba, bir süredir GAN algoritmalarını kullanarak generative art denemeleri yapıyorum. Bazı yaklaşımlar şaşırtıcı sonuçlar verebiliyor.
Text-to-Image sistemlerle çalışırken sonucu etkileyebilecek birçok parametre var. Algoritmik yapının dışında, son kullanıcı dostu derlenmiş notebooklarla bile yapılabileceklerin sınırı yok.
Yukardaki robot portresini gerekli kütüphaneleri derlenmiş CLIP Guided Diffusion notebook ve Gigapixel AI programını kullanarak ortaya çıkardım. NFT Showroom'da da ilk koleksiyon parçamı oluşturmuş oldum.
Bu tarz denemelerde input olarak vereceğiniz açıklayıcı text modelin oluşmasını sağladığı için önemli ve bence bir diğer önemli konu başlangıç görüntüsü...
Robot portresi oluşturulurken yapay zeka modeline insan kafatası görüntüsü başlangıçta tanımlandı...
Yapay zeka destekli generative art denemeleri yapıyorsanız ve bir portre oluşturmak istiyorsanız, init_image olarak bir kafatası kullanmanızı öneririm. Görüntü iyileştirmeden sonra sonuçlar genellikle tatmin edici oluyor.
Herkese iyi günler