Private by design
Your source video is not uploaded to Z-Image. Model inference and video encoding happen on this device.
Turn a short video into a grayscale depth video. Processing stays in your browser and uses your device GPU when available.
Drop a video here or click to choose
MP4, WebM or MOV · up to 50 MB · first 10 seconds
The first run downloads an approximately 50 MB model and caches it in your browser.
Depth Anything V2 Small · Apache-2.0Drop a video here or click to choose
Your source video is not uploaded to Z-Image. Model inference and video encoding happen on this device.
Depth Anything V2 Small estimates each frame, while temporal smoothing reduces brightness flicker and resets automatically at scene cuts.
Chrome or Edge with WebGPU is fastest. Other modern browsers use a slower WebAssembly fallback when supported.
A depth sequence separates scene geometry from color and texture. It gives AI video workflows a cleaner signal for relative distance, occlusion and motion while prompts or reference images control appearance.
Keep the foreground, background and object boundaries easier to distinguish when rerendering or changing style.
Consistent depth frames can help downstream models follow scene layout, parallax and camera movement with less drift.
Use the exported sequence as a depth condition, mask or starting point for depth of field, fog, compositing and new-view workflows.
Choose an MP4, WebM or MOV file. Only the first 10 seconds are processed.
360p and 8 FPS are faster; 480p and 12 FPS preserve more motion detail.
Keep this tab open while the browser estimates, smooths and encodes the depth frames.
It is a sequence in which pixel brightness represents relative distance from the camera. A common convention is brighter for nearer areas and darker for farther areas, and this tool lets you invert that range.
Compatible workflows can use it as a geometry condition alongside a text prompt or reference image. The depth sequence guides layout and motion, while the other inputs define subject appearance and style.
Typical uses include video rerendering, style transfer, ControlNet-style depth guidance, masks, depth of field, fog, compositing and experimental new-view generation.
Yes. The free mode runs locally on your device and does not use Z-Image credits.
The browser downloads and prepares the depth model once. Later runs can reuse the cached files.
No. This release produces relative depth for visual effects and masks, not real-world distance measurements.
No. A depth map is an analysis pass, so the exported video is silent.
最も人気のあるクリエイティブツールを探索
画像をアップロードし、一文で変化させます。
一言で、AIが無限のプロンプトの創造性を提供します。
画像をアップロードして、すぐにプロンプトを取得。
数千の高品質なAIプロンプトを発見。
複数のLoRAモデルを組み合わせて、独自のAIアートワークを作成します
Generate creative videos from text or images with AI.
テキストを即座に素晴らしい画像に変換します。
あなたの作品のために厳選されたアーティスティックスタイルを探索しましょう。
AIで背景を瞬時に削除し、高精度な切り抜きを実現。
画像を4K/8Kにアップスケールし、細部を鮮明に復元。
AIアウトペインティングであらゆる比率に背景を拡張。
ブラウザー内で画像をカスタマイズ可能なASCIIアートに変換します。
アニメーションGIFをダウンロード可能なASCIIアニメに変換します。
MP4、WebM、MOV 動画をダウンロード可能な ASCII 動画に変換します。