Marengo 3.0

Twelve Labs ๐Ÿ“ Embedding

Twelve Labs multimodal video embedding model. Converts video, audio, and text into a shared vector space for semantic search across 36 languages.

Specifications

ModalitiesInput: video, audio, text, image  ยท  Output: embedding
FeaturesTools: No Streaming: No
Capabilitiesembedding, search

Pricing

$0.001 / generation

Platform credit pricing โ€” live rates on the Models page.

Use Marengo 3.0 on Zubnet →