InSighter is a real-time assistive-navigation prototype. It detects 59 common indoor obstacles — from people and furniture to appliances — from a live camera feed, estimates how far away each one is, and speaks a clear, spoken cue — entirely on-device, in real time.
Fine-tuned from a COCO-pretrained checkpoint onto the objects that matter most for everyday indoor navigation — every indoor-plausible COCO class, plus door, stairs, and window sourced and verified outside COCO.
Detects people ahead and announces how far away and which direction they are.
Flags seating obstacles before you walk into them, with a proximity warning up close.
Picks out tables and desks — common waist-height hazards that canes can miss.
Locates doorways so a room transition is never a surprise.
Detection, distance estimation, and speech are each handled by the component best suited for the job — nothing is reimplemented twice across platforms.
A YOLO26n model runs entirely in the browser (onnxruntime-web/WASM) or on-device on iOS (CoreML) — no video ever leaves the phone.
A monocular depth model runs server-side and streams back real distance and direction over a live WebSocket, validated against ground-truth depth data.
The native speech engine (Web Speech API / AVSpeechSynthesizer) reads the closest obstacle aloud — free, offline-capable, no cloud TTS.
Grant camera access in your browser and hear real-time obstacle announcements.