Depth → Image (ControlNet)

Upload a photo - we estimate its depth map (MiDaS) and use that 3D structure to control a generated image. Keeps the spatial layout while changing everything else.

We estimate a MiDaS depth map from this. Images with clear foreground/background separation work best.
وازهێنان 0.7 زۆرتر
~1,200 tokens (SDXL × 1.2 ControlNet)
ئەنجام

خۆشت دەوێت Free.ai؟ بڵێ بە هاوڕێکانت!

ئەم ماڵپەڕە بایەخ بدە