Live demo
Point your camera. Listen for obstacles.
Detection runs entirely in your browser (onnxruntime-web, WASM) — no video ever leaves this device. Distance, direction, and the announcement text stream live over a WebSocket from the Python distance service, computed by a real monocular depth model, not a sensor (browsers have none anyway). Speech is produced natively via the free Web Speech API. Performance metrics for each run are logged to the backend.
Downloading the detection model (~10MB) and compiling it in your browser…
Model (validated, test set)
Precision 0.51Recall 0.40F1 0.45
This session (live)
Inference —FPS —