Computer Vision
Recognize objects in an image and generate a text description

About the service
Our computer vision service allows automatically analyzing supported uploaded images and generating their textual description. This becomes possible thanks to leveraging artificial intelligence technologies and deep learning of neural networks.
The recognition process goes as follows:
- The user uploads up to 20 images in JPG, PNG, WebP, AVIF, HEIC or HEIF format. Each image can be up to 15 MB, with width + height from 300 to 18,000 px.
- Our application analyzes the image, detects objects, people or other elements present in it.
- Using the trained neural network, the textual description of the detected objects is generated.
- The user receives a detailed text listing everything determined by computer vision.
The main advantages of our service:
- High image recognition speed
- Support for popular web and mobile image formats
- Automatic detection of multiple objects and details
- Generation of comprehensive descriptions in natural language
This technology can be used to automate various business processes, improve user experience, in security systems, and many other areas.
What's new
New Quality of Textual Image Descriptions
Thanks to the new large model, the descriptions are more detailed and in a natural style.