Depth Anything V2

Monocular depth estimation model producing detailed depth maps from single images.

Open SourceSelf HostedOffline CapableGPU Required (4GB+ VRAM)
0.0 (0)

About

Depth Anything V2 refines the original monocular depth estimation model with markedly better fine-grained detail and robustness, while running faster with fewer parameters than diffusion-based depth methods. From a single image or video frame it predicts a relative depth map, and fine-tuned checkpoints deliver metric depth. Four encoder scales are offered: small at 24.8M parameters under Apache 2.0, plus base at 97.5M, large at 335.3M, and a planned 1.3B giant, with the larger variants licensed CC-BY-NC-4.0 for non-commercial use. Larger models also hold up better on video sequences, and variable input resolution can be used to extract extra detail. The model loads directly through Hugging Face Transformers and has been ported to Apple Core ML, TensorRT, and ONNX, with community integrations for ComfyUI and Android. Computer vision researchers and developers of robotics, 3D reconstruction, and image generation pipelines that need depth conditioning are the primary users.

Reviews (0)

Leave a Review

No reviews yet. Be the first to review!

Details

Price
Free
Platform
Local/Desktop
Difficulty
Easy (2/5)
License
Apache-2.0
Minimum VRAM
4 GB
Added
Apr 3, 2026

Related Tools

Featured

Contrastive language-image pre-training model by OpenAI for zero-shot visual classification.

Open SourceSelf HostedOfflineGPU 4GB+
Intermediate
0.0 (0)

Lightweight face recognition and analysis framework wrapping multiple models.

Open SourceSelf HostedOffline
Easy
0.0 (0)

Foundation model for monocular depth estimation by TikTok.

Open SourceSelf HostedOfflineGPU 4GB+
Easy
0.0 (0)

Meta AI research platform for object detection, segmentation, and pose estimation.

Open SourceSelf HostedOfflineGPU 8GB+
Advanced
0.0 (0)

Self-supervised vision transformer by Meta producing universal visual features.

Open SourceSelf HostedOfflineGPU 6GB+
Intermediate
0.0 (0)

Simple and effective multi-object tracking using every detection box.

Open SourceSelf HostedOfflineGPU 4GB+
Intermediate
0.0 (0)
Browse all Computer Vision & Object Detection tools