This is my "Ultimate" SDXL workflow. After about a year of messing around with SDXL checkpoints. This is what i've sort of landed with that has a good balance of quality and speed. 3 minutes from 0 to 2721x6167 that...
Laufzeitprofil
Quellenbeschreibung
Intro
This is my "Ultimate" SDXL workflow. After about a year of messing around with SDXL checkpoints. This is what i've sort of landed with that has a good balance of quality and speed. 3 minutes from 0 to 2721x6167 that blows the doors off Flux. It tries to push SDXL to the max in terms of detail and resolution trying to match FLUX. Tons of great SDXL models that I wish had the clean Flux look. You can make stuff that looks VERY close to Flux by using alot of these great extentions that in my opinion put SDXL on Meth. Some of this stuff is very obscure and most workflows are overly bloated with nonsense. Hopefully this is something that you can just expand on as a base.
QuickStart
Everything is initially setup for pure text-to-image. Enter a prompt and hit queue. Example prompt is provided.
KI-generierter Kommentar
KI-generierte Erklärung auf Grundlage von Quellen- und Konfigurationsdetails. Vorschläge sind klar gekennzeichnet.
SDXL Ultimate 16 MegaPixel is an SDXL workflow for generating images from text prompts, refining source images through image-to-image or depth-controlnet guidance, and upscaling results through tiled stages toward 8MP and 16MP outputs.
The default path is pure text-to-image: enter a prompt and queue it. Enable the image-to-image or depth-controlnet group for those paths.
The notes describe composition and cleanup around 2MP before tiled refinement at 8MP and 16MP.
Inputs are text prompts, source images for image-to-image, and optional depth-controlnet guidance.
The workflow covers text-to-image, image-to-image, and upscaling, with stages described at 2MP, 8MP, and 16MP.
It uses SDXL 1.0 as its base model and is published by Smeks.
Named model files are 4x_foolhardy_Remacri.pth, longclip-L.pt, sdxl_lustify_endgame.safetensors, and sdxl_tile_realistic_v2_fp16.safetensors.
Named workflow components include ACN_AdvancedControlNetApply_v2, ACN_ScaledSoftControlNetWeights, ACN_TimestepKeyframeInterpolation, Any Switch (rgthree), Anything Everywhere, Anything Everywhere3, ApplyFBCacheOnModel, BasicGuider, BasicScheduler, CheckpointLoaderSimple, CLIPTÉ
Other named workflow components include CLIPTextEncodeBREAK, ControlNetLoader, DetailDaemonSamplerNode, DynSamplerSelect, easy imageBatchToImageList, easy imageListToImageBatch, Empty Latent by Pixels (WLSH), Fast Groups Muter (rgthree), GetNode, Image Comparer (rgthree), Image
SeaArtLongXLClipMerge is described as extending CLIP handling from 77 to 248 tokens.
Conditional merges are entered in one prompt with BREAK separators; the notes describe PerturbedAttention settings of 1.5–2.
For the 8MP and 16MP stages, the notes specify 30% denoise for defogging and sharpening.
Before use, verify that 4x_foolhardy_Remacri.pth, longclip-L.pt, sdxl_lustify_endgame.safetensors, and sdxl_tile_realistic_v2_fp16.safetensors are available; no source URL is provided for them.
Vorschlag · nicht geprüft
Verify that the listed workflow components are available in your setup before use; no source URLs are provided for them.
Vorschlag · nicht geprüft
Muss das bearbeitet werden?
Melden Sie sich an, um eine Änderungsanfrage zu senden.
Quellen
1 QuelleQuellenauszüge
1 AuszugSource context: 703 downloads · Type Workflows · Base model SDXL 1.0
Optional Stuff
Everything is setup to do depth-controlnet or image-to-image without messing with anything. Just enable the following group and go.
image-to-image
depth-controlnet
Advanced
There are many knobs to tweak. Honestly, you'll have to experiment and see what does what, but breifely I think its worth at least noting from my experience most important bits.
SeaArtLongXLClipMerge All SD checkpoints have a clip limited to 77 tokens and so this patches any clip to 248 tokens allowing alot more detailes. Actual human readable prompts work with it much better similar to flux but it's still very limited. Each tile can hold 248 tokens generated from Florence + base prompt which gives unbelevable results at high tile counts.
Clip Text Encode with BREAK This in conjuction with the long clip merge node gives alot of new possibilities and simplifies a ton of noodles around conditioning. Conditional merges are all in one prompt seperated by the BREAK.
CLIP NegPip + PerturbedAttentionGuidance This definetly helps cleanup the image alot. Generations are much more coherent and make sense. Alot of artifacting goes away by simply having these enabled. A setting of 1.5-2 is about all you really want.
Dynamic Sampler + Detail Daemon I have all this setup how i want but you may want to experiment, but these dynamic samplers push alot more refined + smooth details out of the image almost like flux. Honestly you dont even need to pass the image through Flux because it looks that clean.
Conclusion
Basically steps 1 and 2 are the primers for the actual tiling steps 3 and 4. The image has to be nearly perfect in composition and mostly free from any major artifacting before going into tiling. Step 1 tries to just get a good composition and you may have to change around conditioning and CFG ramping to get a desired result depending on model used etc. Step 2 does the heavy lifting to clean up everything and 2 Mega Pixels seems to be the sweet spot with the TTPlanet Realistic Tile Controlnet before requiring full breakup into tiles. Step 3 and 4 require a nearly perfect image for them to work properly. Going to 8MP and 16MP only requires 30% denoise for "defogging" and "sharpening" everything to bring it all into focus. After 2MP the need for CFG and denoise goes away as you already have the image. It's really just a matter of massaging defects out along the way to 16MP. Thats just my perspective on it all.
Curious to see what people think and would wanna see anything that could improve on this because ive tried so many workflows in the past and nothing ever seemed to really go as far as this.
Geschätzter VRAM Bedarf
Schätzung nicht verfügbar
63,9 MB über 1 von 8 Modell-Dateien. Gesamtmodell-Dateien + 25% Ladeaufwand + 2 GB Ausführungs-Puffer, aufgerundet.
Anforderungen
50 AnforderungenUpscaler · 63.9 MB · PTH · Unknown
ImageUpscaleWithModel
Nicht aufgelöstUpscaler · Unknown
longclip-L.pt
Nicht aufgelöstText encoder · Unknown
sdxl_lustify_endgame.safetensors
Nicht aufgelöstCheckpoint · Unknown
sdxl_tile_realistic_v2_fp16.safetensors
Nicht aufgelöstControlNet · Unknown
UpscaleModelLoader
Nicht aufgelöstUpscaler · Unknown
VAEDecode
Nicht aufgelöstVAE · Unknown
VAEEncode
Nicht aufgelöstVAE · Unknown
ACN_AdvancedControlNetApply_v2
Nicht aufgelöstKnotenpaket · Unknown
ACN_ScaledSoftControlNetWeights
Nicht aufgelöstKnotenpaket · Unknown
ACN_TimestepKeyframeInterpolation
Nicht aufgelöstKnotenpaket · Unknown
Any Switch (rgthree)
Knotenpaket · Unknown
Anything Everywhere
Nicht aufgelöstKnotenpaket · Unknown
Anything Everywhere3
Nicht aufgelöstKnotenpaket · Unknown
ApplyFBCacheOnModel
Nicht aufgelöstKnotenpaket · Unknown
BasicGuider
Nicht aufgelöstKnotenpaket · Unknown
BasicScheduler
Nicht aufgelöstKnotenpaket · Unknown
CheckpointLoaderSimple
Nicht aufgelöstKnotenpaket · Unknown
CLIPTextEncodeBREAK
Nicht aufgelöstKnotenpaket · Unknown
ControlNetLoader
Nicht aufgelöstKnotenpaket · Unknown
DetailDaemonSamplerNode
Nicht aufgelöstKnotenpaket · Unknown
DynSamplerSelect
Nicht aufgelöstKnotenpaket · Unknown
easy imageBatchToImageList
Nicht aufgelöstKnotenpaket · Unknown
easy imageListToImageBatch
Nicht aufgelöstKnotenpaket · Unknown
Empty Latent by Pixels (WLSH)
Nicht aufgelöstKnotenpaket · Unknown
Fast Groups Muter (rgthree)
Nicht aufgelöstKnotenpaket · Unknown
GetNode
Nicht aufgelöstKnotenpaket · Unknown
Image Comparer (rgthree)
Nicht aufgelöstKnotenpaket · Unknown
Image Save
Nicht aufgelöstKnotenpaket · Unknown
ImageScaleToTotalPixels
Nicht aufgelöstKnotenpaket · Unknown
JoinStrings
Nicht aufgelöstKnotenpaket · Unknown
KSamplerSelect
Nicht aufgelöstKnotenpaket · Unknown
LayerMask: LoadFlorence2Model
Nicht aufgelöstKnotenpaket · Unknown
LayerUtility: Florence2Image2Prompt
Nicht aufgelöstKnotenpaket · Unknown
MarkdownNote
Nicht aufgelöstKnotenpaket · Unknown
PerturbedAttention
Nicht aufgelöstKnotenpaket · Unknown
Power Lora Loader (rgthree)
Nicht aufgelöstKnotenpaket · Unknown
PreviewImage
Nicht aufgelöstKnotenpaket · Unknown
RandomNoise
Nicht aufgelöstKnotenpaket · Unknown
SamplerCustomAdvanced
Nicht aufgelöstKnotenpaket · Unknown
ScheduledPerpNegCFGGuider //Inspire
Nicht aufgelöstKnotenpaket · Unknown
SeaArtLongXLClipMerge
Nicht aufgelöstKnotenpaket · Unknown
Seed Everywhere
Nicht aufgelöstKnotenpaket · Unknown
SetNode
Nicht aufgelöstKnotenpaket · Unknown
ShowText|pysssss
Nicht aufgelöstKnotenpaket · Unknown
Text Multiline
Nicht aufgelöstKnotenpaket · Unknown
TiledDiffusion
Nicht aufgelöstKnotenpaket · Unknown
TTP_Image_Assy
Nicht aufgelöstKnotenpaket · Unknown
TTP_Image_Tile_Batch
Nicht aufgelöstKnotenpaket · Unknown
TTP_Tile_image_size
Nicht aufgelöstKnotenpaket · Unknown