AI virtual assistant — object detection and answer generation

NovaAI Tech internal project · Services and SMEs · Software project

The context

A multimodal assistant project, able to handle together what it sees and what it is told.

The problem

Most assistants handle either text or image, rarely both in the same exchange. Yet a user who shows an object and asks a question about it expects a single answer.

What we delivered

Creation of an AI virtual assistant that detects objects, understands inputs and generates answers.

The results

VisionNova detects the objects present in a scene, interprets the inputs received and generates an answer in natural language.

The project is at prototype stage.

Technologies

  • Python
  • LLM
  • Vision par ordinateur

AI deployed

  • VisionNova

    The assistant that sees what it is shown and answers what it is told.

Updated Aug. 11, 2026

Sovereign zone

A project to scope ?

Describe your requirement in a few lines. We come back to you within 48 hours with a costed proposal.

Request a quote Free assessment — 30 min

The assessment is a thirty-minute conversation, with no commitment : we look at your processes and tell you frankly whether software is justified — including when the answer is no.

WhatsApp