TwelveLabs

Video-native multimodal AI platform for searching, analyzing, and generating from video content.

Website: https://www.twelvelabs.io/

Cover Block

Name TwelveLabs
Tagline Video-native multimodal AI platform for searching, analyzing, and generating from video content.
Headquarters San Francisco, United States
Founded 2021
Stage Series A
Business Model API / Developer Platform
Industry Deeptech
Technology AI / Machine Learning
Geography North America
Growth Profile Venture Scale
Founding Team Co-Founders (3+)
Funding Label $100M+ (total disclosed ~$107,100,000) [TechCrunch, Dec 2024]

Links

What an Investor Needs First

TwelveLabs provides a video-native multimodal AI platform, enabling enterprises and developers to search, analyze, and generate content from large video archives using natural language, a capability that has attracted over $100 million in strategic capital from leading data and AI infrastructure investors [TechCrunch, Dec 2024]. The company was conceived by founder Aiden Lee to solve the problem of mapping text to the complex visual and temporal elements within video, moving beyond simple transcript search [TechCrunch, Dec 2024]. Its core differentiation rests on two proprietary foundation models, Marengo for semantic search and Pegasus for video-language generation, which are delivered via API and integrated into cloud platforms like Amazon Bedrock [TwelveLabs] [AWS]. The business model is API-first, targeting self-serve developers and enterprise deals across media, sports, and security, with strategic investments from Databricks Ventures, Snowflake Ventures, and In-Q-Tel signaling deep alignment with enterprise data stacks and government applications [TechCrunch, Dec 2024].

Data Accuracy: GREEN -- Core facts confirmed by TechCrunch, company materials, and AWS case study.

Taxonomy Snapshot

Axis Classification
Stage Series A
Business Model API / Developer Platform
Industry / Vertical Deeptech
Technology Type AI / Machine Learning
Geography North America
Growth Profile Venture Scale
Founding Team Co-Founders (3+)
Funding $100M+ (total disclosed ~$107,100,000)

Inside the Company

TwelveLabs was founded in 2021. The company's origin is tied to founder Aiden Lee's exploration of the difficulty of creating accurate links between text descriptions and the actions, objects, and sounds within video content [TechCrunch, Dec 2024]. The company is headquartered in San Francisco, United States, with a research and development presence in Seoul, South Korea [Crunchbase] [TwelveLabs].

The first disclosed seed round closed in March 2022, led by Index Ventures [Finsmes, Dec 2022]. A seed extension from Radical Ventures followed in December of the same year [The SaaS News, Dec 2022]. The company's Series A round in April 2024, a $50 million financing led by New Enterprise Associates (NEA) and NVentures, marked a significant scaling point [Tracxn]. This was followed by a $30 million strategic investment round in December 2024 from a consortium including Databricks Ventures, Snowflake Ventures, and In-Q-Tel, bringing the total disclosed capital raised to $107.1 million [TechCrunch, Dec 2024].

A critical product milestone was the launch of its Video Intelligence API platform, built around its proprietary Marengo and Pegasus foundation models [TwelveLabs]. The company's integration into Amazon Bedrock provided a major distribution channel [AWS]. The post-money valuation following the Series A was reported in a range of $200 million to $300 million [Mapping Studio, 2026].

Data Accuracy: GREEN -- Company details and funding timeline corroborated by Crunchbase, TechCrunch, and company materials.

Under the Hood

TwelveLabs is built on a core premise: that video, as a dense, multimodal medium, requires a native AI architecture. The company provides a developer-first platform centered on two proprietary foundation models, Marengo and Pegasus, which are accessible via a Video Intelligence API [TwelveLabs]. This approach allows developers to integrate capabilities like semantic search, summarization, and classification directly into applications without the need for manual tagging or annotation [TwelveLabs]. The models are designed to understand not just speech-to-text transcripts, but the visual context, objects, actions, and temporal relationships within video frames [TwelveLabs, TechCrunch, Dec 2024].

The platform's primary product surfaces are its cloud APIs, which are offered on a self-serve basis with a published pricing calculator [TwelveLabs]. A key strategic distribution channel is its integration with Amazon Bedrock [AWS]. Capabilities are organized around specific, high-value workflows. Semantic search enables queries like "show me the first touchdown" or "find when a person in a red shirt entered" across vast video archives [TwelveLabs, AWS]. Content structuring features include automated chaptering, highlight generation, and summarization of long-form video [TwelveLabs]. Classification and embedding functions allow videos, text, images, and audio to be converted into multimodal vectors for downstream analysis or retrieval tasks [TwelveLabs]. The company cites a public case study with Maple Leaf Sports & Entertainment, where the platform reduced video search and retrieval time for highlight creation from 16 hours to 9 minutes [SVG Play, 2026].

Data Accuracy: YELLOW -- Core product claims and model architecture are confirmed by company materials and a third-party case study.

Market Research

The market for video-native AI is expanding beyond content creation, driven by enterprises seeking to unlock value from their vast, untapped video archives. According to a 2024 report from Grand View Research, the global video analytics market size was valued at $8.5 billion in 2023 and is projected to expand at a compound annual growth rate of 21.5% from 2024 to 2030 [Grand View Research, 2024].

Demand is propelled by several converging tailwinds. The volume of enterprise video data is growing exponentially, creating a search and analysis bottleneck. Simultaneously, advancements in multimodal AI architectures have made it feasible to query this data using natural language. A case study with Maple Leaf Sports & Entertainment (MLSE) illustrates the pain point: the organization reported reducing video search and retrieval time for highlight creation from 16 hours to 9 minutes using TwelveLabs' technology [SVG Play, 2026]. The strategic investments from Databricks Ventures and Snowflake Ventures signal that the company's technology is viewed as complementary to modern data stacks, positioning video as another queryable data modality alongside structured and text data [TechCrunch, Dec 2024].

Metric Value
Video Analytics Market 2023 8.5 $B
Projected CAGR 2024-2030 21.5 %

Data Accuracy: YELLOW -- Market sizing is from an analogous, broader sector report. Demand drivers and use cases are supported by a named customer case study and investor commentary.

Competition and Substitutes

TwelveLabs competes in a fragmented landscape where its video-native foundation models are challenged by generalist AI giants, specialized video generation tools, and legacy video management platforms.

Company Positioning Stage / Funding Notable Differentiator
TwelveLabs Video-native multimodal AI for search, analysis, and generation via API. Series A / $107M+ (estimated) Proprietary models (Marengo, Pegasus) for deep video understanding; strategic cloud and data platform integrations.
Runway AI-powered creative suite for video generation and editing. Series C / $237M (estimated) Strong brand with creative professionals; comprehensive generative editing workflow.
Google Veo High-fidelity video generation model from a tech giant. Internal project Integration with Google's broader AI ecosystem and massive compute resources.
Luma AI 3D and video generation from text prompts. Series B / $70M+ (estimated) Focus on 3D scene generation and photorealistic video from text.
LTX Video search and analytics platform. Seed / Undisclosed Focus on semantic search and metadata extraction for enterprise video libraries.

TwelveLabs's defensible edge today rests on three pillars: strategic capital, video-native focus, and enterprise API posture. The company's most significant exposure is to the distribution and ecosystem power of the cloud hyperscalers themselves. While TwelveLabs is available on Amazon Bedrock, its long-term position could be threatened if AWS, Google Cloud, or Microsoft Azure decide to build or acquire competing video intelligence capabilities and bundle them natively [AWS].

Data Accuracy: YELLOW -- Competitor funding and positioning are drawn from Crunchbase and company materials.

Opportunity

If TwelveLabs can establish its video-native foundation models as the standard for querying and understanding video content, the company is positioned to capture a foundational layer of the emerging video intelligence stack. The company's strategic positioning is validated by its inclusion in Amazon Bedrock and by direct investments from the venture arms of Databricks and Snowflake [AWS] [TechCrunch, Dec 2024].

Scenario What happens Catalyst Why it's plausible
Cloud Marketplace Dominance Marengo and Pegasus become the default video intelligence models on AWS, Azure, and GCP, sold as a managed API. Deepening integration with Amazon Bedrock, followed by launches on Google Vertex AI and Microsoft Azure AI Model Catalog. The AWS case study already frames TwelveLabs as unlocking video's potential for the world [AWS].
Vertical Platform Capture The company becomes the embedded AI engine for major media, sports, and security software platforms. A landmark enterprise deal with a league like the NBA or a security giant like Verkada. Early traction with Maple Leaf Sports & Entertainment (MLSE) demonstrates the product-market fit [SVG Play, 2026].
Government & Intelligence Standard In-Q-Tel's investment leads to TwelveLabs' technology being adopted as a standard tool for video analysis within U.S. intelligence and defense agencies. A classified or public contract award for video data processing. In-Q-Tel's mandate is to bridge cutting-edge technology to the U.S. intelligence community [TechCrunch, Dec 2024].

Data Accuracy: YELLOW -- Opportunity framing is based on confirmed strategic partnerships and one public customer case study.

Sources

  1. [TechCrunch, Dec 2024] TwelveLabs is building AI that can analyze and search through videos | https://techcrunch.com/2024/12/12/twelve-labs-is-building-ai-that-can-analyze-and-search-through-videos/
  2. [TwelveLabs] TwelveLabs: Video Intelligence Platform & API | https://www.twelvelabs.io/
  3. [AWS] TwelveLabs unlocks the full potential of video for the world. | https://aws.amazon.com/solutions/case-studies/twelve-labs-case-study/
  4. [LinkedIn] Soyoung Jung - Seoul National University | https://www.linkedin.com/in/soyoung-julie-jung/
  5. [Finsmes, Dec 2022] Twelve Labs Raises $5M in Seed Funding | https://www.finsmes.com/2022/12/twelve-labs-raises-5m-in-seed-funding.html
  6. [The SaaS News, Dec 2022] Twelve Labs Raises $12M in Seed Extension Funding | https://thesaasnews.com/news/twelve-labs-raises-12m-in-seed-extension-funding
  7. [Tracxn] Twelve Labs - Crunchbase Company Profile & Funding | https://www.crunchbase.com/organization/twelve-labs-62b5
  8. [Mapping Studio, 2026] Twelve Labs Valuation, Funding, and Investors | https://mappingstudio.com/company/twelve-labs/valuation-funding-investors
  9. [SVG Play, 2026] MLSE Leverages TwelveLabs AI to Slash Video Search Time for Highlights | https://www.svgplay.com/news/mlse-leverages-twelve-labs-ai-to-slash-video-search-time-for-highlights
  10. [Grand View Research, 2024] Video Analytics Market Size, Share & Trends Analysis Report | https://www.grandviewresearch.com/industry-analysis/video-analytics-market
  11. [Crunchbase] Runway - Crunchbase Company Profile & Funding | https://www.crunchbase.com/organization/runwayml
  12. [Google] Google Veo: Creating high-quality video from text | https://deepmind.google/technologies/veo/
  13. [Crunchbase] Luma AI - Crunchbase Company Profile & Funding | https://www.crunchbase.com/organization/luma-ai
  14. [Crunchbase] LTX - Crunchbase Company Profile & Funding | https://www.crunchbase.com/organization/ltx-2
  15. [Bloomberg, August 2023] Hugging Face Valued at $4.5 Billion in Latest Funding | https://www.bloomberg.com/news/articles/2023-08-24/hugging-face-valued-at-4-5-billion-in-latest-funding

Articles about TwelveLabs

View on Startuply.vc