Chapter 8 Inteligencia Artificial o los artificios de la inteligencia Artificial Intelligence or the artifices of intelligence
Mi trabajo artístico lo he desarrollado como Lucian de Silenttio. Me considero un explorador de las artes y en ese afán es que he recurrido a la literatura, la fotografía, la pintura, los textiles y actualmente trabajo con la Inteligencia Artificial. He nacido en una comunidad del altiplano paceño, en Bolivia. Mi pueblo se llama Pizacaviña y somos Aymaras. Tengo formación en Ciencias de la Educación y Filosofía por la Universidad Mayor de San Andrés, áreas que sigo por pasión; para vivir trabajo en la construcción, soy Maestro Albañil.
Como ocurre hoy en día, conocí acerca del lanzamiento de ‘programas’ que se han llamado ‘Inteligencia Artificial’. El año 2023, aruma, amiga artista con quien trabajé en el pasado, me invitó a formar parte de un equipo de artistas de varias partes de Sudamérica y todos, o la mayoría, de algún pueblo indígena. La finalidad era utilizar el programa Midjourney de creación de imágenes para probar, inicialmente, las limitaciones que podía tener para recrear imágenes de personas, accesorios, vivencias y otros elementos propios de los pueblos indígenas. Acepté con mucho gusto.
Tras los primeros acercamientos y conversaciones con el equipo, comencé a buscar videos y todo cuanto pudiese encontrar acerca de las ‘Inteligencias Artificiales’: usos, cuestiones polémicas, elaboración de prompts y más. Comencé a experimentar y crear imágenes con IAs de acceso libre y si bien los resultados eran interesantes, comencé a ver que no eran capaces de recrear personas ni vestimenta de pueblos del altiplano. Cuando tuvimos acceso (como parte del proyecto) a la modalidad de pago, pude ver que la calidad en cuanto a imagen era mejor, pero tampoco en esa modalidad era posible recrear la cultura Aymara.
Este hallazgo no solo fue mío, sino que en todo el equipo estábamos de acuerdo en esto. Los resultados mostraban a personas con rasgos de la India, de Asia o la visión más cliché (y hollywoodense), de los indios con plumas. Siendo honesto, las imágenes realmente tenían una calidad de fotografía y eso obviamente que deslumbra, pero no reflejaba la realidad. En cierto modo estaba creando una imagen errada de la realidad, y esto porque en su momento hubo un auge de imágenes creadas con IA, que se compartían en redes sociales, y mucha gente las estaba tomando como reales. Incluso gente Aymara podía decir que una imagen (con incongruencias evidentes), sí reflejaba cómo somos los Aymaras. Pienso que ahí hay un peligro de falsar la realidad. Por ejemplo, generé muchas imágenes que intentaban mostrar a personas andinas y escenas de su historia, pero a menudo contenían errores en detalles clave (véase Fig. 8.1, imagen de arriba). No obstante, es posible que también haya contribuido a esto con mis persistentes intentos de crear imágenes de samuráis andinos (véase Fig. 8.1, imagen de abajo).
Durante todo ese tiempo y en la actualidad, siempre he trabajado desde mi teléfono celular por varios factores, pero sobre todo por la comodidad y la portabilidad (a veces creaba imágenes mientras iba en transporte público). Tenía mucha curiosidad y ganas de experimentar con las IAs, al punto que comencé a enseñarle Aymara al ChatGPT. Pero eso es para otra reflexión.
Los diálogos y las propias reflexiones a partir de los resultados que comenzamos a ver nos llevaron a varias preguntas, incluyendo esta: ¿La IA ‘entiende’ lo que uno quiere hacer? Y esto surgió porque a la IA le pedí la creación del ‘Khari Khari’ o ‘Kharisiri’, usando la palabra llana. El Khari Khari es un ser entre mítico y real. Se lo describe como una persona de rostro macilento, con los ojos amarillentos e inyectados en sangre, la piel morena, casi tirando a negro, lo que le da un aspecto tétrico. Se dice que se puede convertir en perro negro. Este ser ataca a la persona y le extrae la grasa del cuerpo. La persona suele presentar dos orificios simétricos en la cadera, como la mordida de algún animal. La persona se enferma repentinamente y el mal no se puede detectar a través de análisis de laboratorio u otro método de la medicina occidental. Si la persona no recibe la medicina puede llegar a morir en cuestión de días. Y aquí viene lo interesante: la medicina para este mal funciona como un ‘antiplacebo’; es decir, el ‘enfermo’ no debe saber para qué es la medicina, pues se sabe que, si la persona se entera que es medicina para el Khari Khari, muere (véase Spedding 2005).
Entonces, cuando a la IA le ‘di’ la descripción (el prompt), al principio en español (véase Fig. 8.2, el grupo de imágenes de arriba), los resultados no fueron los esperados (creo que ahora responde mejor a prompts en varios idiomas) y así, recurriendo al traductor, pusimos el prompt en inglés y los resultados fueron más interesantes (véase Fig. 8.2, el grupo de imágenes de abajo). No hicieron una recreación como lo pensaba, pero había algunos rasgos que habían sido captados y se mostraban en algunas de las imágenes con más fuerza que en otras. Volviendo a la pregunta: ¿La IA ‘entiende’ lo que le pedimos? En un primer momento pensé que podía ser así, que quizás alguna energía cósmica actuaba y me daba respuestas puesto que las imágenes que comencé a crear realmente tenían un ‘algo’. Tenían ese ‘algo’ que tiene el arte. Lo mismo pensaban tanto los colegas del equipo como otras personas a las que les mostraba los resultados, ya que a algunas de ellas o ellos, la IA no les respondía del mismo modo.
Sin embargo, al tratarse de programas computacionales, de códigos y algoritmos, pensé que no respondían a una energía cósmica (animismo), sino que dependían del tipo de palabras, de la forma de redacción de las descripciones y también de la selección de las imágenes que se le pedía que volviera a utilizar para hacer variaciones (el Midjourney ofrece cuatro opciones de imágenes, a partir de las cuales se pueden generar variaciones). Creo que esa elección de imágenes le hacía ‘entender’ el tipo de estética, calidad, etcétera que quería crear. Esta idea la reforcé sabiendo que algunas IAs, al guardar la información de los usuarios, siempre ‘saben’ nuestras preferencias y, en este caso, el tipo de estética que uno busca.
El otro tema importante aquí fue el idioma. Como he dicho, palabras del Aymara como jaqi, awayu, etcétera, no las entendía, ofreciendo resultados aleatorios. Con prompts en español, a veces respondía bien, otras veces no. Con el inglés lo hacía mejor, y creo se debe a que es el idioma usado para crear sus códigos y entrenarlos, además la información usada en su entrenamiento es el inglés y ese idioma le ha proporcionado su ‘bagaje cultural’. Obviamente eso hace que trabaje mejor en ese idioma. Me llevó a la necesidad de aprender el idioma para trabajar con estas herramientas. Lo interesante de las IAs generativas (texto o imagen) es que en algunos casos responden a lo solicitado, dando respuestas inexactas o aleatorias, todo con la finalidad de no quedarse sin responder. Esto podría llevar a error de información o representación de una cultura, con el riesgo de que quien no la conoce pueda asumir que lo hecho por la IA es verdadero. Ahí hay un problema y un peligro.
Hay que entender que la IA es una herramienta, igual que una computadora o un martillo. Son herramientas, son medios, no fines en sí mismos, y pienso que como pueblos indígenas podemos apropiarnos de las mismas en ese sentido. Esto conlleva y plantea un tema de ética, una responsabilidad al hacer uso de la herramienta y los productos generados; y digo esto porque con las redes sociales se ha plantado en la cabeza (sobre todo) de la juventud esa búsqueda de ‘viralidad’. En ese afán se pueden crear y compartir imágenes que falsean la realidad (y ojo que no estoy en contra de elaborar imágenes con miradas artísticas o experimentales, no ya solo sobre los pueblos indígenas).
Uno de los potenciales que le he visto a la IA, más allá de generar imágenes a partir de prompts, es la fusión de imágenes.1 El Midjourney permite subir imágenes como base para crear nuevas. En ese ejercicio utilicé fotos de mi rostro y comencé a fusionarlas o combinarlas con imágenes de textiles, de superhéroes, de cualquier temática. Los resultados varían y en algunos casos son bastante impresionantes. En el mismo camino, se me ocurrió utilizar fotos de mi padre, de mi madre, fotos suyas de adultos, para pedirle a la IA que las recreara, pero como mi padre o mi madre se verían de niños ya que no tenemos fotos de este período en la historia de mi familia (véase Fig. 8.3, donde la imagen de IA de la derecha se basa en la fotografía real de la izquierda). Ese fue un ejercicio interesante, porque en algunos casos pude ver cómo se conservaban algunos rasgos faciales, por ejemplo, los ojos, dando como resultado algo bastante personal e incluso emotivo, aun cuando otros detalles, como la ropa y otras características, distaban mucho de ser precisos.
A partir de esa experiencia es que he creado imágenes haciendo fusiones, y realmente los resultados han sido impresionantes (véanse Figs. 8.4 y 8.5). Estas experiencias, estas formas de ‘relacionarme’ con las IAs, así como los diálogos en los grupos focales, nos han llevado a preguntarnos lo que muchos se cuestionan hoy en día: ¿Las imágenes creadas con IA son arte? ¿Son ‘mi’ (tu) arte? ¿Las ‘firmarías’ como propias? Aunque hay diferentes posiciones y argumentos, considero que son arte, pero este producto artístico ya no es producto solo del humano, sino de esa interacción de la maquina con el ser humano. Esa interacción es fundamental, y podemos decir que es un trabajo de colaboración.
Algo que queda claro es que la imagen creada con la IA, no se puede considerar un producto solo o de ‘autoría’ de la IA. Para que la podamos considerar de ese modo, la IA tendría que crear sola, ‘a voluntad’ podríamos decir, sin la interacción con ningún ser humano, y vemos que eso aún no es posible (quizás en algunos años hayan IAs que puedan hacerlo, quizás a partir de la interacción con nosotros como usuarios ya se está entrenando una IA con esa característica). Por otro lado, para que consideremos que una imagen creada con IA es de autoría del artista, no tendría que haber la interacción de la manera en que se desarrolla. Si vemos con detenimiento una determinada imagen, la forma en que está elaborada (composición, colores, rasgos, etcétera), no puede existir solo a partir de la imaginación del artista. La IA crea una imagen a partir de un prompt, que es una descripción en texto de algo que el artista piensa o que desea recrear; pero el artista no necesariamente ha visualizado todos los detalles de esa descripción. Veamos, puedo escribir: ‘un dragón de fuego volando sobre una montaña’. A partir de ese texto, dos personas pueden imaginar añadiendo detalles de colores, de características del dragón, del fuego, de la montaña, etcétera, así como cada quien puede tener una imagen mental de la composición de la imagen. La imagen creada con la IA a partir de esa descripción tampoco será exactamente como el artista visualizó la escena, y es por esto que digo que la imagen no es cien por ciento el trabajo del artista. También es el caso, particularmente con prompts más detallados, que algunas partes de la descripción pueden ser obviadas, es decir, la IA puede establecer una jerarquía de oraciones y omitir aquellas que considere reiterativas o secundarias, pero que añaden detalles clave que la persona desea aparezcan en la imagen. Esto lo sé por mi propia experiencia de creación de imágenes a partir de texto o a partir de otras imágenes.
Por eso, si bien es una herramienta, como puede ser un pincel, un cincel o un martillo, el pincel, por ejemplo, no ‘decide’ por sí mismo dónde, cómo y con qué color hacer un trazo. En todo momento es el artista quien produce una imagen o recrea un retrato; en cada pincelada el artista está comprometido. Con la IA no pasa esto. Es la IA, la que trabaja, realiza la composición, el uso de colores y otros detalles. No es como trabajar con el Photoshop, donde el artista puede cortar partes de la imagen, manipular sus tonos, subir o bajar el contraste a voluntad; el Photoshop no lo hace solo. En cambio, la IA, sí trabaja sola y ofrece una cantidad de imágenes (cuatro, o más, varía de una IA a otra) como propuesta. Estas propuestas pueden ser completamente aleatorias o guardar algunos elementos comunes, pero sin ser las mismas. Esto me lleva a cuestionar si la firmaría como un producto solo mío. Creo que no me atribuiría todo el crédito, pues hay que reconocer que es la interacción de un artista que tiene toda su cultura, sus vivencias personales, sus búsquedas, su lenguaje, lo que puede dar pie a generar imágenes que sean artísticas, pero es la IA la que propone las imágenes y el artista quien selecciona aquella o aquellas que responden mejor a lo que ha descrito, a sus búsquedas, a su estética, etcétera.
Creo que al presente, cual si fuese una moda, el uso de las IAs ha decaído o realmente se ha ‘acomodado’ y, tras la efervescencia, ahora es utilizada como herramienta por los profesionales de diversas áreas. Esto plantea obviamente continuar y profundizar los debates en torno a la legislación sobre el uso y los alcances de las IAs. Y como anteriormente dije: eso es para dialogar más profunda y largamente.
Fig. 8.1. Lucian de Silenttio, arriba: Bernal Capchiri; abajo: Samurái andino (IA) / top: Bernal Capchiri; bottom: Andean Samurai (AI). Details on page 59.
Fig. 8.1 details. Lucian de Silenttio, arriba: Bernal Capchiri, basada en una fotografía real (nótese que, aunque la imagen muestra un entorno de tierras altas, los textiles que aparecen en la imagen son de las tierras bajas de Bolivia); abajo: Samurái andino, prompt original en inglés: ‘Joven, rostro andino, vestido de samurái, vestido con textiles andinos, una katana en sus manos, en la montaña, fotografía’. Imágenes generadas con Midjourney v.5.2, julio de 2023.
Lucian de Silenttio, top: Bernal Capchiri, based on a real photograph (note that, although the image shows a highland setting, the textiles featured in the image are from lowland Bolivia); bottom: Andean Samurai, original prompt in English: ‘Young man, Andean face, dressed as a samurai, clothed with Andean textiles a katana in his hands, in the mountains, photograph’. Images generated with Midjourney v.5.2, July 2023.
Fig. 8.2. Lucian de Silenttio, El Khari Khari (IA) / The Khari Khari (AI). Details on page 59.
Fig. 8.2 details. Lucian de Silenttio, El Khari Khari, grupo de imágenes de arriba, prompt original en español: ‘Hombre de aspecto extraño, a veces es muy flaco, con ojos hundidos y piel pálida, a veces es grueso y con ojos amarillos, lleva una especie de jeringas, Khari khari, fotografía’; grupo de imágenes de abajo, prompt original en inglés: ‘Hombre de aspecto extraño, rostro andino, a veces es muy flaco, con ojos hundidos y piel pálida, a veces es grueso y con ojos amarillos, lleva una especie de jeringas, Khari khari, foto’. Imágenes generadas con Midjourney v.5.1, junio de 2023.
Lucian de Silenttio, The Khari Khari, upper set of images, original prompt in Spanish: ‘Strange-looking man, sometimes he’s very thin, with sunken eyes and pale skin, sometimes he’s thickset with yellow eyes, he carries some kind of syringes, Khari khari, photograph’; lower set of images, original prompt in English: ‘Strange-looking man, Andean face, sometimes he is very skinny, with deep-set eyes and pale skin, sometimes he is thick and with yellow eyes, he carries some kind of syringes, Khari khari, photo’. Images generated with Midjourney v.5.1, June 2023.
Fig. 8.3. Lucian de Silenttio, Sandalio Torrez Quispe, izquierda: fotografía del padre del artista, editada en Photoshop 2009 para agregar un filtro sepia; derecha: retrato del padre del artista como niño, basado en la fotografía anterior. Imagen generada con Midjourney v.5.2, noviembre de 2023.
Lucian de Silenttio, Sandalio Torrez Quispe, left: photograph of the artist’s father, edited in Photoshop 2009 to add a sepia filter; right: portrait of the artist’s father as a child, based on the previous photograph. Image generated with Midjourney v.5.2, November 2023.
Fig. 8.4. Lucian de Silenttio, izquierda: fotografía de un textil andino Jalq’a con aguja y dedal; derecha: Indígena del futuro, basada en un prompt que combina la fotografía anterior y prompts textuales. Imagen generada con Midjourney v.5.2, noviembre de 2023.
Lucian de Silenttio, left: photograph of Andean Jalq’a textile with needle and thimble; right: Futuristic Indigenous Person, based on a prompt blending the previous photograph and textual prompts. Image generated with Midjourney v.5.2, November 2023.
Fig. 8.5. Lucian de Silenttio, arriba: Ovillos de lana; abajo: Wiphala accidental (IA) / top: Balls of Wool; bottom: Accidental Wiphala (AI). Details on page 59.
Fig. 8.5 details. Lucian de Silenttio, arriba: Ovillos de lana, basada en una fotografía tomada por el artista en octubre de 2020. Imagen generada con Midjourney v.5.2, noviembre de 2023; abajo: Wiphala accidental, un remix de fotografías tomadas por el artista en agosto de 2021. Imagen generada con Midjourney v.6.0, julio de 2024.
Lucian de Silenttio, top: Balls of Wool, based on a photograph taken by the artist in October 2020. Image generated with Midjourney v.5.2, November 2023; bottom: Accidental Wiphala, a remix of photographs taken by the artist in August 2021. Image generated with Midjourney v.6.0, July 2024.
My artist’s name is Lucian de Silenttio. I consider myself an explorer of the arts and work across literature, photography, painting, textiles and now also with Artificial Intelligence. I am from Pizacaviña, a community in the Altiplano near La Paz, Bolivia, and we are Aymara. I have a degree in education and philosophy from the Universidad Mayor de San Andrés; fields that I pursue out of passion. To earn my living, I work in construction; I am a master stonemason.
As is the case for a great many people, I had heard about the launch of programmes that were referred to as ‘Artificial Intelligence’, but I hadn’t used any of them when, in 2023, aruma, an artist friend with whom I had worked in the past, invited me to join a team of artists from various parts of South America, all, or most of whom, were Indigenous. The goal was to use the Midjourney image creation programme to test the limitations it might have in recreating images of people, objects, experiences and other elements characteristic of Indigenous Peoples. I gladly accepted.
After the initial meetings with the team, I began searching for videos and anything I could find about Artificial Intelligence – its uses, controversial issues, prompt creation and more. I began experimenting and creating images with freely available AI tools, and while the results were interesting, I could see that the tools weren’t capable of representing the people and clothing of the Altiplano communities. When we gained access (as part of the project) to a paid account, I saw that the image quality was better, but even so, it wasn’t possible to recreate Aymara culture.
This finding wasn’t just mine; the entire team agreed on it. The results showed people with features from India, Asia or the more clichéd ‘Hollywood’ vision of Native Americans with feather headdresses. To be honest, the images were of really good quality in photographic terms, and that’s obviously impressive, but they didn’t reflect reality. In a way, it was creating a false image of reality, and this was because, at the time, there was a boom in images created with AI, which were shared on social media, and many people thought they were real. Even Aymara people would say that an image (with obvious inconsistencies) reflected what we Aymara people are like. I think there’s a real danger of falsifying reality with these tools. For example, I generated many images that attempted to show Andean people and scenes from their history, but key details were often wrong (see Fig. 8.1, top image). However, I may have also contributed to this with my persistent attempts to create images of Andean samurais (see Fig. 8.1, bottom image).
Throughout that time and still today, I’ve done everything on my mobile phone. This is for several reasons but above all for its convenience and portability: sometimes I created images while riding on public transport. I was so curious and eager to experiment with AI that I even started teaching ChatGPT Aymara (but this is a discussion for another time).
The conversations we had as part of the project and our own reflections based on the results we were seeing led us to several questions, including this one: Does the AI ‘understand’ what you’re trying to do? And this arose because I asked the AI to create the ‘Khari Khari’ or ‘Kharisiri’, to use the more common term. The Khari Khari is a being that is somewhere between mythical and real. It is described as a person with a gaunt face; yellowish, bloodshot eyes; and brown (almost black) skin, all giving it a sinister appearance. It is said to be able to turn into a black dog. This being attacks a person and drains the body of its fat. The person usually has two symmetrical holes in their hip, like an animal bite. The person falls ill suddenly, and the illness cannot be detected through laboratory tests or other Western medical methods. If the person does not receive the right medicine, they can die within days. And here comes the interesting part: the medicine for this illness functions as an ‘antiplacebo’; that is, the patient must not know what the medicine is for, since it is known that if the person finds out it is medicine for the Khari Khari, they will die (see Spedding 2005).
So, when I gave the AI the description (the prompt), initially in Spanish (Fig. 8.2, top set of images), the results weren’t what I expected (I think it now responds better to prompts in various languages), and so, using translation tools, I put the prompt in English, and the results were more interesting (Fig. 8.2, bottom set of images). They didn’t recreate the Khari Khari as I had hoped, but there were some features that had been captured and were more strongly evident in some of the images than in others. To return to the question: Does the AI ‘understand’ what we’re asking it to do? At first, I thought it might be that some cosmic energy was acting and giving me answers, since the images I began to create really had a ‘something’ – that something that art has. My colleagues on the team and other people to whom I showed the results thought the same, since the AI did not respond to some of them in the same way.
Since we’re working with computer programmes, codes and algorithms, I thought they couldn’t really respond to an (animistic) cosmic energy but rather depended on the type of words, the way the descriptions were written and also the selection of images the AI tool was asked to reuse to make variations (Midjourney offers four image options each time, from which variations can be generated); I think that this process of image selection made it ‘understand’ the type of aesthetic, quality and so forth that I wanted to create. In some cases the AI tool also saves a user’s past work, so it always ‘knows’ our preferences and, in this case, the type of aesthetic we are looking for.
The other important issue here was the question of language. As I mentioned, the AI tool couldn’t understand Aymara words like jaqi, awayu and so on, so it offered really random results. This means that AI doesn’t completely grasp my culture, and I say ‘completely’ because language (words, grammatical constructions, meanings and so on) carries culture within it. With prompts in Spanish, sometimes it responded well and other times not so well. It performed better in English. I believe this is because English is the language used to create the algorithms and train the models, plus the data in the datasets are in English, and this language has provided the AI tool with its ‘cultural baggage’. Obviously, this makes these tools work better in English, and I needed to start to learn the language to work with them. One of the interesting things about generative AI, for text or images, is that when it responds to requests with inaccurate or random answers, this is because it’s programmed to always give an answer, even where it has no data. This can lead to misinformation or misrepresentation of a culture, with the risk that someone unfamiliar with a culture might assume that what the AI has produced is true. This is problematic and even dangerous.
It’s important to remember that AI is a tool, just like a computer or a hammer: they are means, not ends in themselves, and I think that as Indigenous Peoples we can take ownership of them in that sense. This raises an ethical issue around the responsibility we assume in using the tool and sharing the outputs generated by it. I say this because social media have instilled a quest for ‘viral’ fame in the minds of young people, especially. In this pursuit, images can be created and shared that distort reality (although I should emphasise that I’m not against creating images that take an artistic or experimental approach in relation to Indigenous Peoples or anyone else).
One of the potentially interesting features that I’ve seen with these AI tools, beyond simply generating images from textual prompts, is image fusion.2 Midjourney allows you to upload images as a basis for creating new ones. In that way, I used photos of my face and began combining them with images of textiles, superheroes or any other theme. The results varied and in some cases were quite impressive. Along the same lines, I came up with the idea of using photos of my mother and father as adults to ask the AI to recreate the same image but as my father or mother would have looked as a child, because we have no photos of this period in our family history (see Fig. 8.3, where the AI image on the right is based on the real photograph on the left). This was an interesting exercise because in some cases, I could see how some facial features, such as the eyes, were preserved, resulting in an image that was quite personal and even moving, even if other details of clothing, for example, were far from accurate.
Since then, I’ve carried on creating images by manipulating and merging existing images, and the results have truly been impressive (see Figs 8.4 and 8.5). These experiences, these ways of developing a ‘relationship’ with AI, as well as the debates we had in the focus groups, have led us to ask a question that many other people are asking today: Are images created with AI art? Are they ‘my’ (your) art? Would you ‘sign’ them as your own work? Although there are different positions and arguments, I do consider them to be art, but this artistic product is no longer solely the work of a human artist but of the interaction between a machine and a human being. This interaction is fundamental, and we can say these are collaborative artworks.
One thing that’s clear to me is that an image created with AI cannot be considered a product solely ‘authored’ by AI; for us to consider it that way, AI would have to create it on its own, ‘at will’ we could say, without interaction with any human being, and we can see that this isn’t yet possible (although perhaps in a few years there will be AIs that can do this; perhaps based on the interactions that we users are having with it now, an AI of that sort is already being trained). On the other hand, it’s also not possible for us to say that an image created with AI is solely the work of the artist. If we look closely at a given image, the way it’s put together (composition, colours, facial features and so on), this is not only the result of the artist’s imagination. AI creates an image from a prompt, which is a text description of something the artist is thinking of or wants to recreate, but the artist hasn’t necessarily visualised all the details of that description. For example, I might write: ‘a fiery dragon flying over a mountain’. From that text, two people might imagine their own versions, adding details of colours and characteristics of the dragon, the fire, the mountain and so on, reflecting each person’s mental image of the composition of the image. The image created by AI from that description will also not be exactly how the artist visualised it, and that is why I say that this image is not entirely the artist’s own work. It is also the case that, particularly with more detailed prompts, the AI may omit some parts of the description that it considers repetitive or secondary, according to the way it weights elements of the prompt, but those same elements might add key details that the artist wanted to appear in the image. I know all this from my own experience of creating images from textual prompts and from other images.
So, although AI is a tool, like a paintbrush, a chisel or a hammer, a paintbrush, for example, does not ‘decide’ for itself where, how or with what colour to make a brushstroke. At all times, it is the artist who produces the image or paints the portrait; the artist is engaged in every brushstroke. This isn’t the same with AI; it’s the AI that does this work, that puts together the composition, that selects the use of colours and other details. It’s not like working with Photoshop, where the artist can cut out parts of the image, manipulate the tones and raise or lower the contrast at will; Photoshop doesn’t make these choices alone. AI, on the other hand, works alone and then offers the artist a number of suggested images to choose from. These suggestions can be completely random or retain some common elements, with smaller differences between them. This leads me to question whether I would sign an image as solely my own work. I don’t think I would take all the credit for it, because you have to recognise that it is the artist who has all their culture, their personal experiences, their intellectual curiosity and their language, all of which can lead to the generation of images that are artistic in nature, but it is the AI that suggests the images and the artist who selects the one or those that best respond to what they have described, to what they’re looking for, to their aesthetics and so on.
I think that AI image generation has been a bit of a fad. At present, its use has declined or has actually become ‘conventional’, and, after all the excitement, it is now simply used as a tool by professionals in various fields. This obviously calls for continuing and deepening the debates around legislation on the use and scope of AI. This is a topic that, as I said before, we should discuss more fully.
Notes
1 Con el tiempo, Midjourney ha evolucionado para permitir la creación de imágenes que se combinan, con o sin texto.
Nota de los editores: En el caso de las imágenes generadas por Lucian a partir de otras imágenes, combinaciones de imágenes y texto, o combinaciones de imágenes, no hemos incluido los prompts, ya que carecen de sentido fuera de contexto. Algunos incluyen información sobre si la imagen es el resultado de un proceso de aumento de resolución o de reinterpretación, así como detalles sobre el estilo o la versión de Midjourney utilizada. Muchos evidencian la gran cantidad de iteraciones que Lucian realizó desde la imagen inicial hasta obtener una imagen que le satisfizo y que quiso compartir en su Instagram.
2 Over time, Midjourney has developed to allow image prompts that blend two or more images, either with or without a textual prompt.
Editors’ note: For images generated by Lucian that are based on images, combinations of images and textual prompts or combinations of images, we have not included the prompts because they are meaningless out of context. Some of them include information about whether the image is the result of ‘upscaling’ or ‘reinterpreting’, with further information about the style or version of Midjourney used. Many of them evidence the significant number of iterations that Lucian went through from the initial prompt until he reached an image that he was pleased with and happy to share on Instagram.