Image mixer is a Stable Diffusion Image Variations model that accepts multiple CLIP image embeddings as inputs. Image Mixer is a Stable Diffusion variant that blends multiple input images via CLIP embeddings, a niche tool for advanced users exploring concept blending and image variation.
Image mixer is a Stable Diffusion Image Variations model that has been optimized to accept multiple CLIP image embeddings as inputs.
How do you use Image Mixer?
1Enter your inputs
Enter images or text/urls to mix together. Choose their strengths.
2Tweak settings
The CFG scale means how close the image looks to the prompt description and/or image. The more steps, the better quality. Seed is a number used to initialize the random number generator, ensuring the same or different results are obtained each time it is used. Steps is the number of iterations or mixing steps to perform during the mixing process.
3Hit Generate
Generate your image and wait.
Pros and cons
Pros
Mixes multiple image embeddings to produce blended variations — unusual capabilityAI
Built on Stable Diffusion so output quality benefits from the broader ecosystemAI
Useful for prompt engineers exploring composition and concept blendingAI
Cons
Niche use case — not for general image generation needsAI
Requires understanding of CLIP embeddings to use effectivelyAI
Output quality varies wildly depending on the input image combinationsAI
The walkthrough on this page covers 3 steps: 1. Enter your inputs 2. Tweak settings 3. Hit Generate.
What are the limitations of Image Mixer?
Niche use case — not for general image generation needs. Requires understanding of CLIP embeddings to use effectively. Output quality varies wildly depending on the input image combinations.