Sign In

Meditations on LoRas training I - Meditaciones sobre el entrenamiento de LoRas I

2

Meditations on LoRas training I - Meditaciones sobre el entrenamiento de LoRas I

I have learned three things:

  1. That the faces do not reproduce does not mean that their features do not 'bleed'. And this gives problems in the style LoRas, since you want them to be as general and versatile as possible. It remains pending for future experiments to see how to enhance the variability of faces and, above all, that they do not begin to look like cousins and brothers of the faces that were in the dataset.

  2. When the LoRa gets the concept he usually gets it very well and when he doesn't, either it's a disaster or you get a LoRa that brings you something fun that can be useful. Beyond this, once you get the concept, the hard work and quality comes from debugging, filtering and improving the dataset. (And already the thousand technical variables that can be touched, little by little xD)

  3. You have to be a little careful with the words to prevent the Civitai moderation system from going crazy and the investment in time (and GPU) does not fall on deaf ears seeing you at the end 4 cats (e.g. jgfanalogc_V1_SD1 and jgfanalogb_V1_SD1; two LoRas to which I have dedicated many hours and they have a great personal load since it is the first time that I work with my own photographic style seriously with my best photographs over more than 15 years... and... there they are dead laughing below a quick experiment with 4 silly photos of the salads with a face that my mother makes xD)

Footnote: Civitai's LoRas training is not as bad as I had read (or, at least, beyond the lack of privacy and extra moderation; for my current standards that are not sensitive to these two points does the work.)

And to finish, I'm going to take advantage of this first article to pay a small tribute to Nekodificador and LDWorksDavid; two great professionals who have helped me put the solid pillars necessary for my knowledge about AI to flourish (and hopefully bear fruit xD). In particular, I want to thank Neko for everything he has taught me about ComfyUI and David for everything that has to do with the subject of training. That and to all the others, who if they read this they will know how to recognize themselves, who have supplied me with the material means to be here making the dramatic transformation from being an 'intellectual man' to a 'man of action'; that much to my regret, it will never be complete, but I do hope that every day it will be more functional. And, as this is already looking like the thanks of a doctoral thesis, let's stop here xD Thank you all!

(And yes, I'm very dyslexic, and worse in English, which is why I like to include both versions. If you read something backwards or misspelled, I've tried to avoid it, but it wasn't a good idea to go over this text 100,000 times so that I still miss something. Is it better to spend the time continuing to research AI, creating interesting works, and making more and better LoRas, don't you think?)

- - -

Tres cosas he aprendido:

  1. Que las caras no se reproduzcan no quiere decir que sus rasgos no 'sangren'. Y esto da problemas en los LoRas de estilo, dado que pretendes que sean lo más generalistas y versátiles posibles. Queda pendiente para próximos experimentos ver cómo potenciar la variabilidad de caras y, sobre todo, que no empiecen a parecer todos primos y hermanos de las caras que estaban en el dataset.

  2. Cuando el LoRa pilla el concepto lo suele pillar muy bien y cuando no, o es el desastre o consigues un LoRa que te aporta algo divertido que puede ser de utilidad. Más allá de esto, una vez pillado el concepto el trabajo duro y la calidad viene de depurar, filtrar y mejorar el dataset. (Y ya las mil variables técnicas que se pueden tocar, poco a poco xD)

  3. Hay que tener un poquito de cuidado con las palabras para evitar que el sistema de moderación de Civitai se vuelva loco y la inversión en tiempo (y GPU) no caiga en saco roto viéndote al final 4 gatos (ej. jgfanalogc_V1_SD1 y jgfanalogb_V1_SD1 ; dos LoRas a los que he dedicado muchísimas horas y tienen una gran carga personal dado que es la primera vez que trabajo con mi propio estilo fotográfico en serio con mis mejores fotografías a lo largo de más de 15 años... y... ahí están muertos de risa por debajo de un experimento rápido con 4 fotos tontas a las ensaladas con cara que hace mi madre xD)

Nota al pie: El entrenador de LoRas de Civitai no es tan malo como había leído (o, por lo menos, más allá de la falta de privacidad y la gran moderación; para mis estándares actuales que no sean sensibles a estos dos puntos hace el trabajo.)

Y para terminar, voy a aprovechar este primer post para hacer un pequeño homenaje a Nekodificador y LDWorksDavid ; dos grandes profesionales que me han ayudado a poner los pilares sólidos necesarios para que mi conocimiento sobre IA florezca (y esperemos que de fruto xD). En particular quiero agradecer a Neko todo lo que me ha enseñado de ComfyUI y a David todo lo que tiene que ver con el tema del entrenamiento. Eso y a todos los demás, que si leen esto sabrán reconocerse, que me han surtido de los medios materiales para estar aquí haciendo la transformación dramática de ser un 'hombre intelectual' a un 'hombre de acción'; que muy a mi pesar, nunca será completa, pero sí espero que cada día sea más funcional. Y, como esto ya va pareciendo los agradecimientos de una tesis doctoral vamos a parar aquí xD ¡Gracias a todos!

(Y sí, soy muy disléxico y, en inglés, peor, por eso me gusta poner ambas versiones. Si leéis algo al revés o mal escrito, he intentado evitarlo, pero tampoco era plan de repasar este texto 100.000 veces para que se me siga escapando algo, ¿mejor dedicar el tiempo a seguir con la IA investigando, haciendo obras interesantes y hacer más y mejores LoRas, ¿no creéis?)

2