$ cat /etc/cookies.conf
We use cookies to understand how people use this site.
Analytics cookies help us improve your experience.
They are off by default. Nothing tracks you until you say so.
$ select cookie_preferences
Members-Only
Recent Talks & Demos are for members only
You must be an AI Tinkerers active member to view these talks and demos.
Demonstrates an edge‑device system that combines multimodal LLMs with segmentation, depth, and recognition models to provide real‑time navigation assistance for low‑vision users.
“Spatial Sense” is a navigational aid project for individuals with low vision, utilizing the synergy of Multimodal Large Language Models (LLMs) with “Segment Anything,” “Depth Anything,” and “Recognize Anything” models for real-time environmental object detection and obstacle negotiation. Central to the project is an LLM agent, engineered to operate on edge devices like Qualcomm Snapdragon boards, ensuring quick, on-device processing for immediate user feedback.
SpatialSense implements recognition, segmentation, depth estimation using ZoeDepth and SAM.
Loading recent emails...