This invention describes a system for automatically providing shoppers with product information and navigation guidance inside stores. It works by using a smartphone, AR device, or smart glasses to capture images of products on shelves. These images are then fed into an Artificial Intelligence model called a Vision and Language Model (VLM), which analyzes the visual content to recognize products, answer user questions about them, and generate real-time shopping assistance or directions, all without needing barcodes.
Why it matters: The rapid development of smaller, more efficient VLMs and their deployment on edge devices has made real-time, in-store analysis more feasible since 2024. This improves the practicality of VLM-driven AR experiences for shoppers.
AI gives you a few directions you could take this. Pick one, and we check whether your version is different enough to patent, then write the filing.
Reinvent this with AI