Private by design
Your source video is not uploaded to Z-Image. Model inference and video encoding happen on this device.
Turn a short video into a grayscale depth video. Processing stays in your browser and uses your device GPU when available.
Drop a video here or click to choose
MP4, WebM or MOV · up to 50 MB · first 10 seconds
The first run downloads an approximately 50 MB model and caches it in your browser.
Depth Anything V2 Small · Apache-2.0Drop a video here or click to choose
Your source video is not uploaded to Z-Image. Model inference and video encoding happen on this device.
Depth Anything V2 Small estimates each frame, while temporal smoothing reduces brightness flicker and resets automatically at scene cuts.
Chrome or Edge with WebGPU is fastest. Other modern browsers use a slower WebAssembly fallback when supported.
A depth sequence separates scene geometry from color and texture. It gives AI video workflows a cleaner signal for relative distance, occlusion and motion while prompts or reference images control appearance.
Keep the foreground, background and object boundaries easier to distinguish when rerendering or changing style.
Consistent depth frames can help downstream models follow scene layout, parallax and camera movement with less drift.
Use the exported sequence as a depth condition, mask or starting point for depth of field, fog, compositing and new-view workflows.
Choose an MP4, WebM or MOV file. Only the first 10 seconds are processed.
360p and 8 FPS are faster; 480p and 12 FPS preserve more motion detail.
Keep this tab open while the browser estimates, smooths and encodes the depth frames.
It is a sequence in which pixel brightness represents relative distance from the camera. A common convention is brighter for nearer areas and darker for farther areas, and this tool lets you invert that range.
Compatible workflows can use it as a geometry condition alongside a text prompt or reference image. The depth sequence guides layout and motion, while the other inputs define subject appearance and style.
Typical uses include video rerendering, style transfer, ControlNet-style depth guidance, masks, depth of field, fog, compositing and experimental new-view generation.
Yes. The free mode runs locally on your device and does not use Z-Image credits.
The browser downloads and prepares the depth model once. Later runs can reuse the cached files.
No. This release produces relative depth for visual effects and masks, not real-world distance measurements.
No. A depth map is an analysis pass, so the exported video is silent.
Entdecken Sie unsere beliebtesten Kreativ-Tools
Bild hochladen, mit einem Satz transformieren
Ein Satz, KI bietet unendliche Prompt-Kreativität.
Bild hochladen, Prompt sofort erhalten.
Entdecken Sie tausende hochwertiger KI-Prompts.
Kombinieren Sie mehrere LoRA-Modelle, um einzigartige KI-Kunstwerke zu erstellen
Generate creative videos from text or images with AI.
Verwandeln Sie Ihren Text sofort in atemberaubende Bilder.
Entdecken Sie kuratierte Kunststile für Ihre Kreationen.
Entfernen Sie Hintergründe sofort mit KI-Präzision.
Verbessern Sie die Bildauflösung auf bis zu 4K/8K.
Erweitern Sie Bilder mit Outpainting auf jedes Seitenverhältnis.
Wandeln Sie Bilder lokal im Browser in anpassbare ASCII-Kunst um.
Verwandeln Sie animierte GIFs in herunterladbare ASCII-Animationen.
Konvertiere MP4-, WebM- oder MOV-Clips in herunterladbare ASCII-Videos.