Private by design
Your source video is not uploaded to Z-Image. Model inference and video encoding happen on this device.
Turn a short video into a grayscale depth video. Processing stays in your browser and uses your device GPU when available.
Drop a video here or click to choose
MP4, WebM or MOV · up to 50 MB · first 10 seconds
The first run downloads an approximately 50 MB model and caches it in your browser.
Depth Anything V2 Small · Apache-2.0Drop a video here or click to choose
Your source video is not uploaded to Z-Image. Model inference and video encoding happen on this device.
Depth Anything V2 Small estimates each frame, while temporal smoothing reduces brightness flicker and resets automatically at scene cuts.
Chrome or Edge with WebGPU is fastest. Other modern browsers use a slower WebAssembly fallback when supported.
A depth sequence separates scene geometry from color and texture. It gives AI video workflows a cleaner signal for relative distance, occlusion and motion while prompts or reference images control appearance.
Keep the foreground, background and object boundaries easier to distinguish when rerendering or changing style.
Consistent depth frames can help downstream models follow scene layout, parallax and camera movement with less drift.
Use the exported sequence as a depth condition, mask or starting point for depth of field, fog, compositing and new-view workflows.
Choose an MP4, WebM or MOV file. Only the first 10 seconds are processed.
360p and 8 FPS are faster; 480p and 12 FPS preserve more motion detail.
Keep this tab open while the browser estimates, smooths and encodes the depth frames.
It is a sequence in which pixel brightness represents relative distance from the camera. A common convention is brighter for nearer areas and darker for farther areas, and this tool lets you invert that range.
Compatible workflows can use it as a geometry condition alongside a text prompt or reference image. The depth sequence guides layout and motion, while the other inputs define subject appearance and style.
Typical uses include video rerendering, style transfer, ControlNet-style depth guidance, masks, depth of field, fog, compositing and experimental new-view generation.
Yes. The free mode runs locally on your device and does not use Z-Image credits.
The browser downloads and prepares the depth model once. Later runs can reuse the cached files.
No. This release produces relative depth for visual effects and masks, not real-world distance measurements.
No. A depth map is an analysis pass, so the exported video is silent.
Explore our most popular creative tools
Upload image, transform with one sentence
One sentence, AI provides infinite prompt creativity.
Upload image, retrieve prompt instantly.
Discover thousands of high-quality AI prompts.
Combine multiple LoRA models to create unique AI artwork
Generate creative videos from text or images with AI.
Turn your text into stunning images instantly.
Explore curated artistic styles for your creations.
Instantly remove backgrounds from images with precision.
Enhance image resolution up to 4K/8K.
Expand images to any aspect ratio with outpainting.
Convert images into customizable ASCII art locally in your browser.
Turn animated GIFs into downloadable ASCII animations.
Convert MP4, WebM or MOV clips into downloadable ASCII videos.