Depth Anything v2
depth-anything-v2Monocular depth-estimation model trained on 595k labeled and 62M unlabeled images. Strong zero-shot generalization in indoor and outdoor scenes.
- Price
- β $0.0050/run
- Input β output
- Text + Image β Text
- Developer
- Community
- Updated
- September 23, 2026
Playground
Try Depth Anything v2
No input form
No input form for this model yet
Its inputs are not documented yet. So that no run fails on a wrong input, we don't offer a form here. Pick a comparable model instead.
Examples
Open full size
InputNo text prompt: the model only takes the input shown.
Open full size
InputNo text prompt: the model only takes the input shown.
Open full size
InputNo text prompt: the model only takes the input shown.
About Depth Anything v2
Depth Anything v2 is a model by Community in the Multimodal category. On Railwail, Depth Anything v2 costs β $0.0050 per run.
Pricing
| Typical run (β 4 s on A100 (40GB)) | $0.0050 per run |
|---|---|
| GPU time (A100 (40GB)) | $0.00138 per GPU second |
- Billed by the GPU time the run actually takes. When the run starts, 3Γ the typical price is reserved from your balance and settled afterwards.
- 1 credit = $0.01
Cost calculator
Price calculator
Typical according to the provider: about 3.6 s
Total
$0.50
50 credits
Per run
$0.005 Β· 0.5 credits
Billed by the actual GPU time; this is an estimate.
API
No verified API example
The public API passes a different input format than this model needs. Use the playground above.
Specifications
- Model ID
depth-anything-v2- Developer
- Community
- Category
- Multimodal
- Input
- Text, Image
- Output
- Text
- Billing
- By usage (tokens or GPU time)
- Catalog entry updated
- September 23, 2026
Tags
- replicate
- depth
- vision-understanding
- open-weights
Frequently asked questions
What is Depth Anything v2?
Depth Anything v2 is a model by Community in the Multimodal category.
How much does Depth Anything v2 cost on Railwail?
On Railwail, Depth Anything v2 costs β $0.0050 per run. You are charged for what each request actually uses. Usage is paid from prepaid credits; 1 credit equals $0.01.
How fast is Depth Anything v2?
There are not enough measured runs of Depth Anything v2 on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.
Is Depth Anything v2 better than BLIP?
That depends on the task. Depth Anything v2 (Community) and BLIP (Salesforce) are both models in the Multimodal category. The comparison page shows their prices and specifications side by side.
Compare Depth Anything v2 and BLIPCan Depth Anything v2 process images?
Yes. Depth Anything v2 accepts images as input in addition to text.
Comparable models
All in this category- BLIPSalesforce
Salesforce BLIP. Vision-language model for image captioning and visual question answering. Given an image it writes a short natural-language caption, or answers a question about the image when one is supplied. A widely used baseline for automatic captioning.
- CLIP InterrogatorCommunity
pharmapsychotic's CLIP Interrogator. Takes an image and produces a Stable-Diffusion-style text prompt by combining BLIP captioning with CLIP to rank likely subjects, artists, mediums and styles. Commonly used to reverse-engineer a prompt from an existing picture.
Meta Segment Anything 2. Promptable segmentation across images and video with temporal memory. Zero-shot, point/box/mask prompts, fast on a single H100.
All models through one API
One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = $0.01.