Skip to main content

Overview

The Vision API enables you to bring human-like reasoning capabilities to your video and imagery data. Upload any format of video from any camera, environment, or industry and leverage spatial and temporal reasoning beyond metadata-based tagging and traditional vision approaches.

Authentication

All API requests require a an API token. Tokens are secrets and should be kept securely and privately. Do not share it with others or expose it in any client-side code. Requests should include your API token in an Authorization HTTP header as follows:
Manage your API token through the access management options (Users).

Basic Workflow

Upload and query an Input.

1. Get a pre-signed URL for video or image upload

Go to endpoint documentation

2. Upload your video or image file

Use the signed_url from step 1 to upload your file. Once uploaded, inputs are updated with a valid status if it’s format is supported.

3. Query

Query against your video or image Input with free-form natural language.
Go to endpoint documentation

Advanced Workflows

Core Concepts

Understand inputs, collections, inference jobs, queries, and predictions.

Management Features

Set up organizations, users, and projects for granular access control.

Collection Querying

Aggregate your entire dataset in a collection and search at scale.

Configure custom background queries

Define a query that runs automatically and returns exactly what you need with structured responses.