Twelve Labs builds a video understanding platform that allows enterprises to search, analyse, and extract insights from video content at scale. The company's technology is grounded in multimodal foundation models, combining visual and language understanding to process petabytes of video data. Its platform is used by over 30,000 developers and companies, including organisations in sports and media.
The company offers two core AI models: Marengo, a multimodal encoder, and Pegasus, a video-language foundation model. These underpin a platform that can be deployed across cloud, private cloud, and on-premise environments. Twelve Labs holds SOC 2 Type 2 certification, reflecting a focus on enterprise-grade security requirements.
Headquartered in San Francisco with an office in Seoul, Twelve Labs has raised $107 million in funding. Its technology serves industry verticals ranging from professional sports leagues to media companies, enabling them to derive structured information from large video libraries without manual review.






