Annotator-DepthAnything — Annotation AI Model

Depth Anything V1 Large estimates depth from one photo, no stereo rig — near reads white, far reads black, and fine structure like hair or foliage survives. Feeds ControlNet Union depth mode. 1.4GB VRAM, sub-second.

Details

  • CategoryAnnotation
  • Year2024
  • LicenseApache-2.0

Compliance & Provenance

  • ProviderOpen-source · Specialized
  • EU AI Act RiskMinimal Risk
  • Art. 50 TransparencyNot applicable

Inputs & Outputs

  • ImageInput · image

    Source image for depth estimation. Works on any natural scene — indoors, outdoors, portraits, objects. RGB accepted; alpha channel is ignored.

  • Depth MapOutput · image

    Grayscale depth map at the same resolution as the input. White pixels are closest to the camera, black pixels are farthest. Feed into ControlNet-XL-Union (depth mode) for depth-guided image generation.

Tags

  • annotator
  • depth-estimation
  • monocular-depth
  • preprocessing
  • controlnet
  • depth-anything
  • apache-2.0

Alternatives in Image Control

  • ControlNet XL Canny

    Guided SDXL image generation from a Canny edge map + text prompt. Sharper structural adherence than Union. Outputs up to 1024×1024.

  • ControlNet XL Union

    Guided SDXL image generation from a control image + text prompt. One model, 6 modes: canny, depth, openpose, hed, normal, segment. Outputs up to 1024×1024.

Resources