En la carrera por las inteligencias artificiales hay una lucha constante por ver quien tiene la IA generadora de imágenes más rápida, precisa y económica. Hemos visto ya varios ejemplos de esto y ha habido un boom increíble de modelos de difusión capaces de generar casi cualquier imagen que se les pida. Entre estas opciones teníamos al ya conocido Dall-e 2 de Openai que es a la fecha uno de los mejores modelos y el más capaz. Sin embargo el problema con Dall-e es la limitación a la hora de pedir inputs con caras de personas famosas y otros temas como el gore y temas sexuales. Después vimos como salió Dall-e mini y eliminaba estas restricciones aunque los resultados eran más bien pobres. Poco después vimos a la luz a Midjourney que eliminaba también algunas restricciones de Dall-e 2 y tenia unos resultados interesantes.
Pues hoy he tenido acceso a una nueva inteligencia artificial llamada Stable Difussion, y es muy interesante. He podido acceder a la beta privada a la que pudieron acceder cerca de 10 mil personas, esto supongo que debido a la poca capacidad de los servidores de la empresa para poder generar tantas imágenes a la vez. Todo se hace por discord al mas puro estilo de Midjourney. Lo más impactante del proyecto es que esta IA se presenta como una buena alternativa a Dall-e 2 pues por ahora es gratis probarla para los que accedimos a la beta y en un futuro sea de código abierto, algo que openai no tiene.
Ahora les quiero mostrar algunas de las cosas que es capaz de hacer y las imagenes que he podido hacer.
Esta fue la primer imagen que hice, le he pedido un gato haciendo trading en wall street y me ha generado esto que me ha parecido muy lindo e interesante.
Después le he pedido que me haga una pintura de un hombre búho fumando una pipa.
Aquí quería probar que tan bien hacía imágenes de gente famosa y bueno le pedí a Will Smith en un sauna y bueno, el resultado es muy impactante.
Después le pedí que me hiciera a Joe Biden siendo golpeado en la cara por la Reina Isabel y bueno jaja.
A continuación le pedí a una chica flotando en un cuarto lleno de agua y me salió esto.
La última prueba era para ver como manejaba otros inputs y le pedí a una chica mirando desde una colina a un asteroide en llamas cayendo a la tierra y me ha gustado.
Bueno espero que les haya gustado lo que he podido generar con esta IA, espero seguir trayendo más material en estos días.
In the race for artificial intelligences there is a constant battle to see who has the fastest, most accurate and cheapest image generating AI. We have already seen several examples of this and there has been an incredible boom of diffusion models capable of generating almost any image that is asked of them. Among these options we had the well-known Dall-e 2 from Openai which is to date one of the best and most capable models. However, the problem with Dall-e is the limitation when requesting inputs with faces of famous people and other topics such as gore and sexual themes. Later we saw how Dall-e mini came out and eliminated these restrictions although the results were rather poor. Shortly after we saw the release of Midjourney which also eliminated some restrictions of Dall-e 2 and had some interesting results.
Today I have had access to a new artificial intelligence called Stable Difussion, and it is very interesting. I have been able to access the private beta to which about 10 thousand people could access, this I suppose due to the limited capacity of the company's servers to generate so many images at once. Everything is done by discord in the purest Midjourney style. The most striking thing about the project is that this AI is presented as a good alternative to Dall-e 2 because for now it is free to try it for those who accessed the beta and in the future it will be open source, something that openai does not have.
Now I want to show you some of the things that it is able to do and the images that I have been able to do.
This was the first image I made, I asked for a cat trading on wall street and it generated this image which I found very nice and interesting.
Then I asked him to make me a painting of an owl man smoking a pipe.
Here I wanted to test how well he did images of famous people and well I asked Will Smith in a sauna and well, the result is very striking.
Then I asked him to make me Joe Biden getting punched in the face by Queen Elizabeth and well haha.
Next I asked him for a girl floating in a room full of water and I got this.
The last test was to see how I handled other inputs and I asked a girl looking down a hill at a flaming asteroid falling to earth and I liked it.
Well I hope you liked what I have been able to generate with this AI, I hope to continue bringing more material in these days.
Translated with www.DeepL.com/Translator (free version)
Imagen hecha por @fclore22