A
licenseNot graded
qualityB
maintenanceEnables real-time fusion of 2D/3D visual bounding boxes, head-pose gaze vectors, and acoustic beamforming azimuth angles into cross-modal focal salience scores to resolve deictic spatial references for multimodal agents.
7
MIT