Companies

Twelve Labs

twelvelabs.io

TwelveLabs delivers enterprise video AI that searches, analyzes, and understands video across vision, audio, and language.

HQSan Francisco, California, United States
Employees51-200
Funding$110M
Valuation$300M
17 active roles
Profile 6mo agoJobs checked 16h ago
AI / MLFoundation Model ProviderB2B SaaSSeries D+$50M-$200M

About

TwelveLabs builds enterprise video-native multimodal AI models and APIs that search, analyze, and understand video across vision, audio, and language. It sells to teams in media, sports, advertising, government, security, and other video-intensive industries, differentiating through foundation models designed specifically for video understanding.

Market

Twelve Labs competes in the video-native multimodal AI and video intelligence market, positioning itself as a provider of foundation models and infrastructure that enable machines to interpret visual, audio, and spoken information in video. Its differentiation is a video-specific focus spanning semantic search, summarization, event and object identification, and insight extraction through APIs and multimodal models, rather than general-purpose AI alone.

Target Customers

Twelve Labs primarily targets video-intensive organizations in media and entertainment, security, enterprise knowledge management, and related industries that need programmatic video search, summarization, event detection, and insight extraction. The likely buyers are engineering, AI/platform, and product teams; the available evidence does not specify a particular company-size range.

At a Glance

Problem

Video is a high-value but poorly indexed information source: organizations store it in archives, camera systems, broadcasts, meetings, factories, stadiums, and other repositories, yet typically access it through filenames, folders, captions, transcripts, or human memory. That makes finding a precise moment expensive and slow, while leaving valuable footage underused for analysis, reuse, personalization, and monetization. The pain is especially acute when teams must search across large libraries at scene level rather than merely locate a file.

The clearest killer use case is sports and media production. MLSE reports that Twelve Labs reduced video search and retrieval from 16 hours to 9 minutes, enabling faster highlight creation and more personalized fan content. Similarly, SBS used the technology to search for specific scenes and reuse archived visual-effects footage, replacing manual deep-dives into large volumes of video data with semantic retrieval.

Product / Service

Twelve Labs is an enterprise video-intelligence platform delivered through REST APIs and Python and Node.js SDKs. Customers upload videos and use the platform to search, analyze, generate embeddings, or reason across a knowledge store. Its multimodal models combine visual content, audio, spoken words, and on-screen text, allowing users to search with natural-language or image queries, identify moments and interactions, summarize or analyze a video, answer questions, and feed video embeddings into their own machine-learning systems.

The product is offered as usage-based APIs and platform capabilities including indexing and search, embedding, and analysis. Marengo provides video representations for retrieval and machine-learning workflows, while Pegasus supports video analysis and generation-oriented tasks. The benefit is that customers can turn an archive into a machine-readable, time-addressable knowledge base without building and maintaining separate models for text, image, and audio analysis; Twelve Labs also exposes a pricing calculator and a free-start path for API usage.

Market

Twelve Labs competes in enterprise video intelligence, multimodal video foundation models, and AI infrastructure for making video searchable and usable by applications and agents. Its differentiation is native semantic and temporal understanding across vision, audio, language, and their relationships, rather than relying only on keywords, metadata, or transcripts. The competitive set includes hyperscaler services such as Google Cloud Video Intelligence, Microsoft Azure AI Video Indexer, and Amazon Rekognition Video, which provide automated video recognition, indexing, or stored and streaming video analysis.

The company has meaningful commercial and financing traction rather than appearing to be a pre-revenue concept. Public customer evidence includes MLSE and SBS deployments, and Twelve Labs says its technology is being used across media, entertainment, sports, and advertising workflows. By July 2026 it had announced a $100 million round co-led by NEA and NAVER Ventures, followed by $30 million in strategic investments from Databricks, Snowflake, SK Telecom, HubSpot Ventures, and In-Q-Tel; the available evidence does not disclose revenue, so adoption, customer outcomes, partnerships, and funding are the clearest public traction signals.

Founders & Leadership

Jae LeeFounder
CEO and Co-founder
Aiden LeeFounder
Co-founder and CTO
Soyoung LeeFounder
Co-founder and Head of Go-to-Market
Dave ChungFounder
Co-founder
SJ KimFounder
Co-founder
Yoon KimPresident and Chief Strategy Officer

Funding History

2022-03
Seed$5M

Index Ventures

2022-12
Seed extension$12M

Radical Ventures

2023-07
Seed / strategic investment$10M

NVentures, Intel Capital, Samsung NEXT

2024-04
Series A$50M

New Enterprise Associates (NEA), NVentures

2024-12
Strategic investment (Tracxn labels it Series A)$30M

Databricks, Snowflake Ventures, SK Telecom, HubSpot Ventures, IQT

2026-07
Series B$100M

New Enterprise Associates (NEA), NAVER Ventures

Recent News

2026-07-01funding
We raised $100M to build Video Superintelligence

TwelveLabs raised $100 million in Series B funding, co-led by NEA and NAVER Ventures. The company will use the funding to scale its Video Cognition System and advance its Video Superintelligence vision.

2026-06-01product
TwelveLabs Bring Its Video Understanding Technology Directly to Creators

TwelveLabs announced Rodeo, its first application-layer product, an AI-powered creative copilot that lets creators find, edit, and assemble video footage using natural language.

2026-02-25partnership
VAST Data and TwelveLabs Partner to Expand Video Intelligence for the World’s Largest and Most Secure Video Archives

VAST Data and TwelveLabs announced a partnership enabling a customer-managed deployment path for TwelveLabs’ video foundation models on the VAST AI Operating System. The collaboration supports video search, analytics, and understanding across on-premises, cloud, and other controlled environments.

2025-11-30product
Marengo 3.0: Real-World Multimodal Embedding AI

TwelveLabs announced Marengo 3.0, a multimodal embedding model for video retrieval. It supports composed queries, multilingual search, and long-form video understanding.

2025-11-20partnership
Automating Video Intelligence in Frame.io Workflows

TwelveLabs integrated its Marengo and Pegasus models into Frame.io V4 through Custom Actions. The integration enables creative teams to semantically search video libraries within their workflows.

2025-09-05partnership
Become a TwelveLabs Partner: Enterprise Video AI Program

TwelveLabs introduced its enterprise video AI Partner Program, offering technical enablement, commercial incentives, and go-to-market support for solution partners.

2025-08-29partnership
Cross-Modal Search with TwelveLabs Marengo & S3 Vectors

TwelveLabs described an integration of Marengo with Amazon S3 Vectors for cross-modal search, semantic retrieval, and scalable AWS video workflows.

2025-08-22partnership
From MP4 to API: Video AI with Marengo & Pegasus on AWS

TwelveLabs presented an end-to-end video AI pipeline using Marengo and Pegasus on Amazon Bedrock. The workflow supports video search, analysis, and metadata generation.

Active Roles

17
Seoul, South Korea/Engineering/2d ago
Seoul, South Korea/Engineering/4d ago
Remote US/Engineering/10d ago
Seoul, South Korea/Engineering/11d ago
San Francisco/Operations/14d ago
San Francisco/Data & Analytics/16d ago
Seoul, South Korea/Engineering/21d ago
Seoul, South Korea/Engineering/30d ago
San Francisco/Engineering/34d ago
Senior GTM Systems EngineerRemote$126k – $140k
Remote US/Engineering/34d ago
San Francisco/Sales/44d ago
San Francisco/Engineering/45d ago
Seoul, South Korea/HR & Recruiting/63d ago
Seoul, South Korea/Data & Analytics/146d ago
Seoul, South Korea/Engineering/196d ago

Business Model

TwelveLabs offers free access, usage-based Developer pricing tied to the amount of video processed, and enterprise contracts. Revenue comes from APIs and platform capabilities for video indexing, search, embedding, and analysis.

Products

Video understanding and analysis platformMultimodal AI models for videoVideo search and retrieval APIsVideo summarization and insight-extraction capabilities

Customers

MLSE (Maple Leaf Sports & Entertainment)NFL MediaOracleMindsDBVoxel51MindproberSource DigitalGS SHOPUNICEF KoreaSBSDYN SportAffiliateNetworkProtegeMantis SolutionsQencode

Tech Stack

Video-native multimodal AIFoundation modelsVideo understanding and machine learningSemantic video searchVisual, audio, and spoken-language processingAPIs for video analysis

Competitors

Deepgram
Tavus
Amazon Web Services

Key Investors

Databricks, Databricks Ventures, Firstman Studios, HubSpot Ventures, InnoWhale Ventures