DriftCapture · const
MINIATURE_DEPTH_ANYTHING_3
Four blocks of 32 channels in four heads of 8 — the smallest a two-dimensional rotary embedding divides — with the camera token, query-key norms and rotary embedding from the third block, so one within-view and one across-view block each carry all three; a trained grid of 2, so any other grid is resized; and a head of 16 features.
Explained in DriftCapture.
const MINIATURE_DEPTH_ANYTHING_3: DepthAnything3Configimport { MINIATURE_DEPTH_ANYTHING_3 } from '@driftengine/capture';