AI virtual assistant — object detection and answer generation
The context
A multimodal assistant project, able to handle together what it sees and what it is told.
The problem
Most assistants handle either text or image, rarely both in the same exchange. Yet a user who shows an object and asks a question about it expects a single answer.
What we delivered
Creation of an AI virtual assistant that detects objects, understands inputs and generates answers.
The results
VisionNova detects the objects present in a scene, interprets the inputs received and generates an answer in natural language.
The project is at prototype stage.
Technologies
AI deployed
VisionNova
The assistant that sees what it is shown and answers what it is told.
Updated Aug. 11, 2026