Depth → Image (ControlNet)

Upload a photo - we estimate its depth map (MiDaS) and use that 3D structure to control a generated image. Keeps the spatial layout while changing everything else.

We estimate a MiDaS depth map from this. Images with clear foreground/background separation work best.
Ónyénwē 0.7 N'ihi
~1,200 tokens (SDXL × 1.2 ControlNet)
Ihenhọrọ ahụ

Ịhụnanya Free.ai? Kpọtụrụ enyi gị!

Nhazi ihuakwụkwọ a