DriftCapture · const

MINIATURE_DEPTH_ANYTHING_3

Four blocks of 32 channels in four heads of 8 — the smallest a two-dimensional rotary embedding divides — with the camera token, query-key norms and rotary embedding from the third block, so one within-view and one across-view block each carry all three; a trained grid of 2, so any other grid is resized; and a head of 16 features.

Explained in DriftCapture.

const MINIATURE_DEPTH_ANYTHING_3: DepthAnything3Config
import { MINIATURE_DEPTH_ANYTHING_3 } from '@driftengine/capture';