Grayscale depth map
Official Video-Depth-Anything (Small, Apache-2.0) with temporal consistency, so brightness stays stable across frames. It is widely used as the depth control signal that downstream creative tools expect.
Drop in a clip, pick a mode, hit convert. XIANGYU Deep Video Converter generates grayscale depth maps, human pose skeletons and 478-point face point clouds — all on your own machine. The original audio is preserved automatically, and nothing is ever uploaded.
One clip, six looks. Intermediate depth and pose results are cached, so switching modes does not re-run inference.
Official Video-Depth-Anything (Small, Apache-2.0) with temporal consistency, so brightness stays stable across frames. It is widely used as the depth control signal that downstream creative tools expect.
PyTorch KeypointRCNN (COCO 17 keypoints) draws the skeleton over the original footage, making motion trajectories readable at a glance.
Depth map as the base layer with the pose skeleton stacked on top — spatial layering plus visible motion.
MediaPipe FaceLandmarker extracts 478 facial landmarks rendered as a colored point cloud, preserving every expression detail.
Depth base + pose skeleton + face point cloud in a single overlay, for maximum information density.
KMeans analyzes dominant colors and maps them to pseudo-depth (black is always 0), or supply your own color-map JSON to control each color precisely.
This is not a cloud service. There is no server to upload to, because there is no upload code.
The interface is a local web service bound to 127.0.0.1. Your footage never crosses the network boundary, which also means no queue, no rate limit, no credit system.
Inference runs locally with CUDA when a compatible GPU is present, and falls back to CPU automatically otherwise. Each completed job tells you which one it used.
The original soundtrack is remuxed into the output at 96k-256k (192k default), with channels configurable. Start time and duration trims are applied to audio in sync.
Every model ships inside the installer. After installation you can disconnect from the internet entirely and everything still works.
No sign-up, no cloud project, no telemetry dashboard. You install it, paste your license key once, and use it.
A single lifetime license, no subscription and no seat fee. Verification is offline Ed25519, so licensing never depends on our servers staying up.
The trial has every feature enabled — only clip length is capped at 10 seconds. Run your own footage first, then decide.
One payment · yours permanently
Windows x64 installer, about 2.6 GB including all model weights, with resume support. Everything works offline after installation.
This is an independent, fully local tool with no ads, no telemetry and no data monetisation. Installation, licensing and conversion results — any question is welcome by email, and replies usually arrive within 24 hours.
Email: luyingzhi2@gmail.com
Buy license · $9.99 lifetime