New AI module

AI Observer in Xeoma

Video surveillance that finds exactly what you need using artificial intelligence. Describe the object you’re looking for in natural language, and the system will be looking for it in the camera’s field of view. A single module replaces dozens of dedicated detectors. Everything runs locally, on your own server.

Xeoma AI Observer in action
Match found
What the module does

A camera that understands what’s happening

AI Observer module icon Xeoma’s AI Observer, powered by visual language models — the next level of language models (LLMs) — takes video analytics beyond individual detectors: it analyzes the scene as a whole, describes it in natural language, and reacts to what you’ve defined in plain words. This module is also called the detector of everything or the universal detector — it replaces dozens of dedicated detectors. Three capabilities in one module:

Detection by free-form description

Finds by description — instead of a detector for every task

Describe in free form what you need to detect, and the module will trigger on it. For example: “a weapon”, “a crack in the wall”;, or a specific type of cargo. There’s no longer any need to configure a separate detector for each object — just describe what you’re looking for, or pick one of the ready-made scenarios.

Responds to rare and non-standard situations that used to require custom development.
Description of what's happening in the frame

Describes what’s happening

The module continuously analyzes the frame and produces a coherent text description of what the camera sees: what objects are in the frame and what they are doing.

Works in real time, across the whole frame, with many objects at once.
Description on camera previews and in the archive

Overlays the description onto the video stream

The text description of what’s happening can be shown on top of the camera image and overlaid onto recordings and their exports. The operator grasps the essence of the scene at once, even if there are small details in the frame they didn’t notice.

Coming soon: searching the archive for events by these text descriptions — a feature in development. Once released, it will let you find the moment you need in recordings simply by describing what you’re looking for.
Requirements of the AI Observer module Requirements: The “AI Observer” module requires a discrete (dedicated) graphics card with at least 8 GB of VRAM.
Use cases

One module, dozens of tasks

Xeoma’s AI Observer replaces a host of specialized detectors and removes the need to build new ones: no commissioning paid development and waiting for it to be implemented. AI Observer is equally useful wherever it matters to understand the context of a scene and to find non-standard objects and events — from a construction site to a retail floor.

Cargo control at a construction site
Construction & logistics
Tracking by type of cargo
Describe the dry bulk cargo in natural language — for example, “coal” — and log its arrival and dispatch without manual oversight.
Analytics on the retail floor
Retail
From theft to product display
A wide range of tasks: from preventing theft and monitoring staff presence to keeping an eye on how products are displayed on the shelves.
Searching for vehicles by description
Transport & parking
Finding a vehicle by description
Alerts and search by vehicle attributes (brand, color, model type, license plate) – all in one detector.
Detecting cracks on underwater structures
Industry & mining
Defect detection
Inspection: cracks, corrosion, weld damage, material quality analysis.
Monitoring in manufacturing and warehousing
Manufacturing & warehouses
Non-standard situations
Works even with rare events for which no ready-made detector exists.
Monitoring crop ripening by description
Agriculture
Monitoring crop ripening
Ripening stage, signs of plant disease, empty beds in the field and in the greenhouse.
Smart city
Cities & public areas
Flexible monitoring
Adapts to changing surveillance tasks — you only need to rephrase the description.

Local, flexible, and predictable in price

AI Observer follows Xeoma’s general logic: processing on your own server, flexible module chains, and clear licensing.

1
Everything on your server

Scene analysis and description are performed locally. The video isn’t sent to a third-party cloud — the data stays inside your network.

2
A license without overpaying

The module is purchased per number of cameras, with no charge per frame or per event. There’s no need to use several detectors, commission paid customization, and wait for it to be implemented.

3
Flexible setup with chains

Connect AI Observer with any of Xeoma’s reactions — recording, notifications, alarm — into a single chain tailored to your task.

Quick start

Test it in 5 minutes

1
Launch Xeoma in the Trial edition

Download and launch Xeoma. The program starts in the Trial editionwith 99% of all features available. For convenient testing you can get a demo license.

2
Add a camera

Let Xeoma find cameras automatically on the network, or add a camera manually by its IP address. Any type is supported, and there’s a built-in demo camera.

3
Add the “AI Observer” module

Add the module to the camera chain.

4
Describe what to detect

Enter a query — for example, “a person in a yellow baseball cap” or “a KIA sedan” — or use one of the ready-made queries offered.

5
Set up reactions

Recording, notifications, alarm, your own reactions — choose the result you need. More about notifications in Xeoma.

Setting up the AI Observer module in Xeoma
The Trial edition has a number of limitations. To get properly acquainted with AI Observer, request a free demo license — it unlocks all features, no credit card and with no obligations. All you need is an email address.
Get a free demo license

Put AI Observer to the test

Get a demo license and test it in real conditions at your own pace, or buy a license at an affordable price.