Detector inference plus tracking and rule evaluation runs at 30 ms median inference per camera on the node, 28 to 34 ms inference across cameras, measured over 380 seconds of live multi-camera load. Camera to alert is a different number: detectors sample each camera once every 1.5 seconds by default, and that sampling interval, not inference, is the dominant term. End to end lands around 0.8 s typically and near 1.5 s at p90, and both are floors that exclude camera and network time. Customer pilots should validate camera-to-alert latency on the actual network, VMS workflow, and alert destination.