Private by design
Your source video is not uploaded to Z-Image. Model inference and video encoding happen on this device.
Turn a short video into a grayscale depth video. Processing stays in your browser and uses your device GPU when available.
Drop a video here or click to choose
MP4, WebM or MOV · up to 50 MB · first 10 seconds
The first run downloads an approximately 50 MB model and caches it in your browser.
Depth Anything V2 Small · Apache-2.0Drop a video here or click to choose
Your source video is not uploaded to Z-Image. Model inference and video encoding happen on this device.
Depth Anything V2 Small estimates each frame, while temporal smoothing reduces brightness flicker and resets automatically at scene cuts.
Chrome or Edge with WebGPU is fastest. Other modern browsers use a slower WebAssembly fallback when supported.
A depth sequence separates scene geometry from color and texture. It gives AI video workflows a cleaner signal for relative distance, occlusion and motion while prompts or reference images control appearance.
Keep the foreground, background and object boundaries easier to distinguish when rerendering or changing style.
Consistent depth frames can help downstream models follow scene layout, parallax and camera movement with less drift.
Use the exported sequence as a depth condition, mask or starting point for depth of field, fog, compositing and new-view workflows.
Choose an MP4, WebM or MOV file. Only the first 10 seconds are processed.
360p and 8 FPS are faster; 480p and 12 FPS preserve more motion detail.
Keep this tab open while the browser estimates, smooths and encodes the depth frames.
It is a sequence in which pixel brightness represents relative distance from the camera. A common convention is brighter for nearer areas and darker for farther areas, and this tool lets you invert that range.
Compatible workflows can use it as a geometry condition alongside a text prompt or reference image. The depth sequence guides layout and motion, while the other inputs define subject appearance and style.
Typical uses include video rerendering, style transfer, ControlNet-style depth guidance, masks, depth of field, fog, compositing and experimental new-view generation.
Yes. The free mode runs locally on your device and does not use Z-Image credits.
The browser downloads and prepares the depth model once. Later runs can reuse the cached files.
No. This release produces relative depth for visual effects and masks, not real-world distance measurements.
No. A depth map is an analysis pass, so the exported video is silent.
Explore nuestras herramientas creativas más populares
Cargue una imagen, transfórmela con una frase
Una frase, la IA proporciona creatividad infinita para sus mensajes.
Cargue una imagen, obtenga el mensaje al instante.
Descubra miles de mensajes de IA de alta calidad.
Combina múltiples modelos LoRA para crear obras de arte de IA únicas
Generate creative videos from text or images with AI.
Convierta su texto en imágenes impresionantes al instante.
Explora estilos artísticos seleccionados para tus creaciones.
Elimina fondos de imágenes al instante con precisión de IA.
Mejora la resolución de la imagen hasta 4K/8K.
Expande imágenes a cualquier relación de aspecto con outpainting.
Convierte imágenes en arte ASCII personalizable directamente en el navegador.
Convierte GIF animados en animaciones ASCII descargables.
Convierte vídeos MP4, WebM o MOV en vídeos ASCII descargables.