Depth → Image (ControlNet)

Upload a photo - we estimate its depth map (MiDaS) and use that 3D structure to control a generated image. Keeps the spatial layout while changing everything else.

We estimate a MiDaS depth map from this. Images with clear foreground/background separation work best.
Ko te kaipāpāho 0.7 He uaua
~1,200 tokens (SDXL × 1.2 ControlNet)
Whakamutunga

E hiahia ana Free.ai? Whakapāpāho ki ōna hoa!

Whakawa tēnei pātaka