MODELS
Unlimited Basic subscription: unlimited Qwen 3.8 27B for €14.95/month, with no per-token billing

Use cases / Image and video

Image analysis

Analyze images with AI and turn visual content into information.

Ask questions about images, classify content and extract useful details for your application. Multimodal models combine vision and language to interpret what an image contains.

With QDivZero, you can deploy image analysis models and integrate them through an API. Your product supplies the image and instructions; the model returns descriptions, categories or data to continue your process.

Turn images into information

Use what appears in an image in your processes.

Add visual understanding to your application. Query, classify and connect images to the context each task needs.

Questions about images

Get descriptions or query specific image features. Provide the purpose of the question to guide the response.

Visual content classification

Organise images using your product’s labels. Validate categories and retain originals to review ambiguous results.

Element inspection

Query visible differences and potential anomalies. Evaluate images from your environment to decide what information your inspection workflow can use.

Image and text analysis

Connect images to instructions and questions. Combine visual information with other task data within your application.

Build it with QDivZero

From an image to the information your application needs.

Send an image and a question or task to a compatible vision model. Validate outputs with examples from your environment and connect them to your application’s workflow.

Your application flow

  1. Image

    Prepare visual input compatible with the model.

  2. Instruction

    Specify elements, categories, or information to analyze.

  3. Vision

    The multimodal model interprets the image and request.

  4. Result

    Use the description or data in your application.

Compute

Open-weight / Hugging Face

You can start with…

Choose a multimodal model for your image types, supported resolution and required information. Evaluate answers with your categories and examples; a visual LLM and a specialized model can solve different tasks. Use consistent questions and criteria when comparing alternatives. Observe visible information recognition and details the image cannot support, and define how your application handles those outputs.

Multimodal reasoning

DeepSeek V4.1 Flash

Explore reasoning and understanding of text and images in multi-step tasks.

View model on Hugging Face

Text and vision

Qwen3.8-27B

For conversation, code, and tasks combining text, images, and your own context.

View model on Hugging Face

Image analysis

Image analysis: frequently asked questions

What is AI image analysis?

AI image analysis uses vision models to interpret visual content. Generate descriptions, assign categories, or answer questions about an image. Deploy compatible multimodal models on QDivZero to add these capabilities to your application.

How do I analyze an image with a multimodal model?

Prepare an image in a compatible format and send it with a question or text instructions. Specify the information you need and the output format. The multimodal model interprets both inputs and returns a result for your application.

Can I classify images automatically with AI?

Yes. Define required categories and use a vision model suited to your content. Classify product images, documents, or other collections. Compare generated labels with real examples before incorporating results into your process.

Which open-weight models can I use for computer vision?

Start with multimodal models such as GLM-5.3 Flash, Qwen3.8, or DeepSeek V4.1 Flash for visual understanding. Capabilities vary between models and versions. Check inputs, runtime compatibility, and results on your application images.

Can AI inspect images and detect anomalies?

Analyze visible characteristics or differences with a model suited to the task. For anomalies specific to a product or environment, evaluate specialist models and representative examples. Quality depends on the data, visual signal, and model capabilities.

How do I integrate image analysis through an API?

Deploy a compatible vision model on QDivZero and connect your application to its endpoint. Prepare the image and instructions in a supported format. Your application receives generated descriptions or fields and uses them in the required process.

How does a multimodal model differ from an image classifier?

A multimodal model can relate an image to questions and text instructions. A specialized classifier usually returns categories defined during training. Choose based on whether you need flexible answers about content or a specific visual task with stable labels.

How do I evaluate AI image analysis quality?

Use images representative of your application and define the expected answer for each task. Compare accuracy, missed details and behavior with unclear images. Also measure response time and cost at the resolution and volume you need to process.

Ready to help your application understand images?

Integrate visual analysis, classification and image questions into your product. Choose a compatible multimodal model, deploy it in QDivZero and connect its information with your processes through an API.