Retrieve an indexed asset
This method retrieves information about an indexed asset, including its status, metadata, and optional embeddings or transcription.
Use this method to:
-
Monitor the indexing progress:
- Call this endpoint after creating an indexed asset
- Check the
statusfield until it showsready - Once ready, your content is available for search and analysis
-
Retrieve the asset metadata:
- Retrieve system metadata (duration, resolution, filename)
- Access user-defined metadata
-
Retrieve the embeddings:
- Include the
embeddingOptionparameter to retrieve video embeddings - Requires the Marengo video understanding model to be enabled in your index
- Include the
-
Retrieve transcriptions:
- Set the
transcriptionparameter totrueto retrieve spoken words from your video
- Set the
Authentication
Your API key.
You can find your API key on the API Keys page.
Path parameters
Query parameters
Specifies which types of embeddings to retrieve. Values: visual, audio, transcription. For details, see the Embedding options section.
To retrieve embeddings for a video, it must be indexed using the Marengo video understanding model. For details on enabling this model for an index, see the Create an index page.
Response
A string indicating the date and time, in the RFC 3339 format (“YYYY-MM-DDTHH:mm:ssZ”), that the indexing task was created.
A string indicating the date and time, in the RFC 3339 format (“YYYY-MM-DDTHH:mm:ssZ”), that the indexing task was last updated. The platform updates this field every time the indexing task transitions to a different state.
A string indicating the date and time, in the RFC 3339 format (“YYYY-MM-DDTHH:mm:ssZ”), that the indexing task has been completed.
System-generated metadata about the indexed asset.
User-defined metadata for this indexed asset.
The platform returns this object only for the videos that you uploaded with the enable_video_stream parameter set to true.
Contains the embedding and the associated information. The platform returns this field when the embedding_option parameter is specified in the request.
An array of objects that contains the transcription. For each time range for which the platform finds spoken words, it returns an object that contains the fields below. If the platform doesn’t find any spoken words, the data field is set to null.