Overview
The Vision API enables you to bring human-like reasoning capabilities to your video and imagery data. Upload any format of video from any camera, environment, or industry and leverage spatial and temporal reasoning beyond metadata-based tagging and traditional vision approaches.Authentication
All API requests require a an API token. Tokens are secrets and should be kept securely and privately. Do not share it with others or expose it in any client-side code. Requests should include your API token in an Authorization HTTP header as follows:Basic Workflow
Upload and query an Input.1. Get a pre-signed URL for video or image upload
2. Upload your video or image file
Use the
signed_url from step 1 to upload your file. Once uploaded, inputs are updated with a valid status if it’s format is supported.3. Query
Query against your video or image Input with free-form natural language.Go to endpoint documentation
Response Example
Response Example
Advanced Workflows
Core Concepts
Understand inputs, collections, inference jobs, queries, and predictions.
Management Features
Set up organizations, users, and projects for granular access control.
Collection Querying
Aggregate your entire dataset in a collection and search at scale.
Configure custom background queries
Define a query that runs automatically and returns exactly what you need with structured responses.