An AI artist explains his workflow

The creation of high-quality artificial intelligence (AI) art often transcends simple prompt entry, involving a sophisticated blend of digital tools and traditional artistic skill. As the artist in the video above vividly illustrates, bringing a complex vision to life, such as a boxing match between Stelphi and Muhammad Ali, can demand an intensive investment of time—sometimes up to

17 hours for a single piece. This remarkable dedication highlights a crucial aspect of modern art creation: the evolving **AI art workflow** is a dynamic collaboration between human ingenuity and machine capability.

This article delves into the intricate process, offering insights into how artists navigate the complexities of generative AI to achieve their creative goals. We will explore the specific techniques, parameters, and artistic decisions that transform an initial concept into a finished masterpiece, acknowledging the blend of various software tools. Understanding this sophisticated approach is essential for anyone looking to master the **AI art creation process** and push the boundaries of digital expression.

From Concept to Canvas: Initiating Your AI Art Project

Every compelling artwork begins with a clear vision, and AI art is no exception. The artist’s journey with Stelphi, his “very funny and very clumsy” alter ego, commenced with a specific scene in mind: a boxing match against Muhammad Ali. Such a detailed concept provides a strong anchor, preventing the creative process from drifting aimlessly.

Initially, a traditional sketch forms the bedrock of this **generative AI art** process. While AI models like Stable Diffusion are powerful, they are also “extremely cheeky,” as the artist notes, often diverging from the original idea if left undirected. A foundational sketch serves as a visual blueprint, ensuring the artist maintains control over the narrative and composition. This crucial first step emphasizes that even in a technology-driven art form, traditional artistic practices remain invaluable for guiding the machine toward a coherent outcome.

Precise Posing: Harnessing ControlNet and Manual Refinement

Achieving a specific pose for characters within AI-generated imagery can be particularly challenging. The initial phase often involves trying a variety of random prompts to generate an initial pose that aligns with the artist’s vision. When the desired stance for Stelphi proved elusive through prompting alone, a shift to manual intervention became necessary. The artist meticulously recreated the pose in Photoshop, showcasing the essential role of external editing software in this **AI art workflow**.

A significant advancement in **AI image generation** is the use of extensions like ControlNet. This tool allows artists to exert precise control over the AI’s output by providing structural guidance through input images such as depth maps, Canny edges, or skeletal poses. For instance, reproducing a complex pose that once required extensive manual effort can now take mere minutes, potentially as little as 15 minutes, thanks to ControlNet’s integration. This drastically improves efficiency and accuracy, enabling artists to guide the AI with unprecedented precision and consistency.

Decoding Stable Diffusion: Key Parameters for Realism

Mastering the technical parameters within Stable Diffusion is paramount for achieving realistic and detailed **digital art**. The artist emphasizes that several options significantly influence the final output, allowing for nuanced adjustments that elevate an image from generic to professional.

Samplers: Sculpting Details and Realism

Samplers are crucial algorithms that determine how Stable Diffusion converts noise into a coherent image. Different samplers produce distinct visual characteristics, impacting aspects like realism and fine details. For example, the Euler sampler often yields a “very synthetic” or “fake” appearance, which can be useful for stylized or abstract works. Conversely, the DPM family of samplers, such as DPM++ 2M Karras, often excels at rendering realistic textures, performing “great” when replicating intricate details like human skin. Artists must experiment with various samplers to discover which one best suits their specific artistic intentions and desired aesthetic outcome.

Steps: Guiding the Generative Process

“Steps” refers to the number of iterations Stable Diffusion performs to refine an image from the initial noise. A lower number of steps results in quicker generations but typically less detail and fidelity. Conversely, a higher step count allows the AI more time to work on the prompt, leading to a more refined and detailed image. While a balance is key, understanding the relationship between steps and image quality enables artists to optimize their workflow for both speed and precision in **AI art creation**.

Inpaint and Outpaint: Targeted Editing and Creative Expansion

Two foundational techniques in **AI image editing** are inpaint and outpaint, which empower artists to modify or expand their creations with great flexibility. Inpaint allows artists to select a specific area of an image and instruct the AI to change only that section. This targeted modification is invaluable for correcting imperfections, altering specific elements like clothing or facial features, or refining small details without affecting the entire composition.

On the other hand, outpaint prompts the machine to imagine and generate content beyond the existing canvas. Based on the surrounding pixels, the AI extrapolates what lies “outside the box,” creating seamless extensions to backgrounds or adding new elements that blend organically with the original image. These capabilities are indispensable for iterating on compositions and expanding creative possibilities within the **Stable Diffusion workflow**.

Crafting Consistent Characters: Training Custom AI Models

Achieving consistency for a custom character like Stelphi across multiple images requires advanced techniques beyond standard prompting. The artist effectively employed a method of training a specialized model on Stelphi’s face. This involved creating a 3D model of the character, capturing numerous snapshots of his face from various angles, and then using these images to train a custom AI model. This process generates a unique “keyword” that, when used in prompts, ensures a consistent facial structure and appearance for Stelphi in subsequent generations.

Denoising Strength: Controlling AI’s Creative Intervention

Denoising strength is another critical parameter that offers artists control over the AI’s influence during the generation process. This setting dictates how much the AI can alter the input image or seed. A higher denoising strength grants the AI more freedom to introduce new elements and significantly change the image, which is useful for transforming concepts dramatically. Conversely, a lower denoising strength retains more of the original image’s integrity, making subtle refinements without drastic alterations. This parameter is particularly crucial when trying to replicate specific faces or maintain precise details, allowing the artist to balance AI-driven creativity with their own artistic direction in **AI art production**.

The Hybrid Approach: Merging AI Output with Digital Artistry

The successful creation of complex AI art rarely relies solely on generative models; it often involves a sophisticated dance between AI and traditional digital editing tools. The artist’s workflow clearly demonstrates this hybrid approach, allocating approximately 50% of the effort to Stable Diffusion, 40% to Photoshop, and 10% to Procreate. This significant reliance on manual editing underscores the irreplaceable value of human artistry.

Replicating specific individuals, such as the legendary Muhammad Ali, presents unique challenges, especially regarding facial features and physique. The artist initially requested Stable Diffusion to generate a face resembling Ali, but then performed extensive manual adjustments in Photoshop. This involved warping traits like the nose, jawline, and eyes, and meticulously reharmonizing exposure and skin tone to achieve an authentic portrayal. This back-and-forth process, where artists “look back a lot” between AI outputs and manual refinement, is characteristic of advanced **AI art creation**. It ensures that artistic intent is paramount, guiding the AI to produce results that resonate with the human creator’s vision, even when rectifying peculiar AI outputs, such as unexpected body details.

Mastering Difficult Details: Realistic Hands and Physique

Certain elements consistently pose a significant challenge for AI generative models, with human hands being a prime example. Their complex anatomy, subtle gestures, and numerous small details often lead to distorted or unrealistic renditions from AI alone. To circumvent this persistent hurdle, the artist ingeniously incorporated a personal solution: photographing his own hands in various positions and then integrating these images into the artwork, with about 50% of Stelphi’s hands originating from this method. This pragmatic approach highlights the creative problem-solving essential in a modern **AI art workflow**.

Another specific challenge involved accurately depicting Muhammad Ali’s physique. AI models often default to generating overly “buffed” figures, reflecting contemporary ideals of athleticism rather than historical accuracy. The artist explicitly aimed for a realistic portrayal of Ali’s body—”big and healthy” but “not super fit, without a six pack.” Achieving this required constant manual refinement and iterative adjustments, demonstrating that guiding the AI to specific, nuanced physical characteristics demands significant human intervention. The iterative process ensures historical accuracy and artistic intent override generic AI outputs.

The Future of Art: AI as an Opportunity for Artists

The artist’s perspective, rooted in two decades of traditional painting and five years of digital art, offers a compelling vision for the future of creativity. Rather than feeling “threatened” by AI, he views it as a profound “opportunity.” This sentiment is shared by many who believe that **generative AI art** is not replacing artists but rather expanding their toolkit and opening up entirely “new branches of art.”

AI tools facilitate experimentation and accelerate certain aspects of the creative process, allowing artists to focus more on conceptualization and refinement. This technology can democratize art creation, enabling new talented individuals to explore digital artistry without years of traditional training in certain techniques. The fusion of **AI art workflow** with human creativity represents an exciting frontier, pushing the boundaries of what is possible and redefining the very nature of artistic expression.

Prompting the AI Artist: Your Workflow Queries

What does creating AI art involve?

Creating high-quality AI art often requires more than just typing prompts; it involves blending digital tools like AI models with traditional artistic skills. This combination helps artists bring complex visions to life.

How do artists usually begin an AI art project?

Artists typically start with a clear vision for their artwork. They often use a traditional sketch as a visual blueprint to guide the AI model and maintain control over the composition.

What is ControlNet and why is it useful?

ControlNet is a tool that allows artists to precisely control the poses and structure of characters in AI-generated images. It uses input images, like sketches or depth maps, to guide the AI’s output more accurately and efficiently.

Why do AI artists often use programs like Photoshop in their workflow?

Many AI artists use a hybrid approach, combining AI tools with traditional digital editing software like Photoshop. This allows them to refine details, correct imperfections, and ensure the final artwork aligns perfectly with their artistic vision.

Leave a Reply

Your email address will not be published. Required fields are marked *