girl_e000010_02_20250520032052.png

Process

Tools

Techniques

  • txt2img

    Generates images directly from text prompts using diffusion models. This foundational technique is essential for testing LoRA outputs, fine-tuning character anatomy, and exploring style or pose variations purely through descriptive input.
  • img2img

    Transforms an existing image into a refined or altered version based on a new prompt. Excellent for enhancing LoRA results by maintaining structure while improving details like fingers, facial alignment, or artistic consistency.
  • inpainting

    Allows selective editing of specific regions in an image (e.g., redrawing hands or fixing eyes) while preserving the rest. Extremely powerful for correcting flawed anatomy in LoRA outputs without starting from scratch.
  • workflow

    Refers to a structured pipeline combining multiple steps—txt2img, ControlNet, LoRA merging, upscaling—into a single process. Critical for advanced LoRA generation, ensuring consistency, precision, and high-quality anatomy in every output.
  • vid2vid

    Transforms an existing video into a new animated sequence guided by prompts or reference styles. Useful for applying LoRA styles to motion footage, preserving body proportions and consistency over time.
  • txt2vid

    Creates videos directly from text prompts. Combines the creativity of txt2img with temporal coherence. Ideal for exploring how LoRA-trained characters perform across animated timelines, including hand gestures and body movement.
  • img2vid

    Converts static images into short animated clips, often using motion interpolation or AI animation tools. Excellent for testing if LoRA-generated character designs retain anatomical clarity in motion.
  • controlnet

    Provides fine-grained control over composition and anatomy using reference maps (e.g., pose, depth, line art). A game-changer for LoRA-based generation, especially when aiming for precise hands, limbs, or symmetrical proportions.

Generation data

COPY ALL

Resources used

Prompt

External Generator
txt2img
masterpiece, catgirl, short hair, black ears, blue tail, wearing oversized hoodie, sitting on a windowsill with city lights, holding a steaming mug, fluffy hair, relaxed pose, intricate background bokeh

Negative prompt

bad hands, bad anatomy, extra fingers, missing fingers, blurry, fused limbs, worst quality, lowres, watermark, signature, text, jpeg artifacts

Other metadata

cfgScale:7.5
steps:24
sampler:DPM++ SDE Karras

Discussion