Skip to content

TwelveLabsのPegasus 1.5 レビュー

TwelveLabsは、マルチモーダルインテリジェンスを搭載したエンタープライズビデオAIを提供し、視覚、音声、言語にわたるビデオの検索、分析、理解を可能にします。

shipped 2026年4月21日videofreemium
Domain rating69Monthly visits3.9K/mo
videocoderesearch
Pegasus 1.5 by TwelveLabs - AI tool for pegasus twelvelabs. Professional illustration showing core functionality and features.

注目ポイント

1Pegasus 1.5は2026年4月20日に正式にリリースされ、Time Based Metadata Extraction (TBM)を導入しました。
2このモデルは、単一のAPI呼び出しで最大2時間のビデオを処理し、構造化されたタイムスタンプ付きメタデータを抽出します。
3社内ベンチマークによると、Pegasus 1.5は集計セグメンテーション品質においてGemini 3 Proおよび3.1 Proを13.1%上回りました。
4TwelveLabsのプラットフォームは、約1分で1時間のビデオをインデックス化でき、リアルタイム処理速度の約60倍を達成します。

Stork’s verdict on Pegasus 1.5 by TwelveLabs

Pegasus 1.5は、カスタムJSONスキーマによる正確なTime Based Metadata Extractionを提供しますが、そのAPI駆動の性質は開発者リソースを必要とします。

Pegasus 1.5 by TwelveLabs reviewed by Stork AI · stork.ai/ja/pegasus-1-5-by-twelvelabs

Stork Quadrant

Becomes the API· 27/100

Replaceable as a UI, but kept alive as the API the agents call.

TwelveLabs built a capable multimodal video understanding API before the frontier labs caught up. That window is closing. GPT-4o, Gemini 1.5 Pro, and Claude already handle video natively, and they're getting faster and cheaper. There's no proprietary data, no network, no regulatory gate — just a specialized model that bigger players will commoditize.

Claude Sonnet 4.6, scored 2026-05-30

Defensibility · 0/100

  • Physical-world coupling
  • Regulatory moat
  • Network liquidity
  • Proprietary refreshing data
  • High-trust catastrophic workflows
  • Multi-party coordination
  • Brand / community / taste

An LLM alone could replace

  • Summarize what happens in a video by describing its content
  • Transcribe audio and extract key topics or themes from spoken content
  • Answer questions about a video's subject matter given a transcript or description
  • Generate metadata tags or chapter markers for video content

Agent-Readiness · 60/100

  • Verified MCPStork MCP listing: io-twelvelabs-twelvelabs-mcp-server (untested)
  • Listed on agent surfacesStork:io-twelvelabs-twelvelabs-mcp-server
  • Usage-based pricingpricing page heuristic match: https://www.twelvelabs.io/pricing
  • Headless agent auth
  • Public OpenAPIhttps://docs.twelvelabs.io/v1.3/docs/resources/platform-overview
  • Active changeloghttps://www.twelvelabs.io/blog/introducing-pegasus-1-5 (2026-04-19)
  • llms.txthttps://www.twelvelabs.io/llms.txt

Score history · +5 pts over 3 re-scores

How to defend

Go vertical and own the liability: pick one industry where wrong video analysis has real consequences — insurance claims, legal evidence, broadcast compliance — and become the vendor that signs the contract and bears the risk. That's the only move that creates a moat here.

  • Ship an MCP server and list it on Stork — biggest single point gain (+25).
  • Expose API-key auth with a self-serve sandbox tier; remove sales-call gates (+15).

Pegasus 1.5 by TwelveLabs について

本社
San Francisco, USA
設立
2020
チーム規模
51-100
資金調達
Series A

仕様

APIドキュメント

API提供状況

はい、公開API

overview

TwelveLabsのPegasus 1.5とは?

TwelveLabsのPegasus 1.5は、Twelve Labsによって開発されたAIビデオ推論モデルであり、開発者や企業が未加工のビデオを大規模に構造化されたクエリ可能なデータに変換することを可能にします。これはビデオを多次元ボリュームとして処理し、構造化されたタイムスタンプ付きメタデータを抽出し、単一のAPI呼び出しで最大2時間のビデオをサポートします。2026年4月20日にリリースされたPegasus 1.5は、その中核機能であるTime Based Metadata Extraction (TBM)を導入しました。これにより、ユーザーは事前のインデックス作成や前処理なしに、特定のタイムスタンプ付きデータを受信するためのカスタムJSONスキーマを定義できます。このモデルは、時間的境界を自律的に発見し、ビデオ全体の期間にわたって構造化されたメタデータを抽出することで、Pegasus 1.2などの以前のバージョンからの大幅な進歩を示しています。TwelveLabsは、NAB Show 2026でこのリリースとその他の進歩を発表しました。

features

TwelveLabsのPegasus 1.5の主な機能

TwelveLabsのPegasus 1.5は、高度なビデオインテリジェンスとビデオコンテンツからの構造化データ抽出のために設計された包括的な機能スイートを提供します。

  • 視覚、音声、言語にわたるマルチモーダルインテリジェンスを搭載したエンタープライズビデオAI。
  • 自然言語クエリを使用したセマンティックビデオ検索と取得。
  • 長尺コンテンツからの自動ビデオ要約とインサイト生成。
  • タイムスタンプ付きデータのためのカスタムJSONスキーマ定義を可能にするTime Based Metadata Extraction (TBM)。
  • ビデオアセットのコンテンツモデレーション、コンプライアンス、ブランドセーフティ機能。
  • 単一のパイプラインによるマルチモーダルデータ取り込みで、リアルタイム処理速度の約60倍を達成。
  • 既存の開発者ワークフローやアプリケーションへのシームレスな統合のためのAPIおよびSDKアクセス。
  • 特定のイベントのタイムスタンプを正確に特定し、ビデオを意味のある部分に分割するためのTemporal GroundingとSegmentation。
  • コンテキストを維持しながら、最大2時間のビデオをシングルパスで処理する長尺ビデオサポート。
  • 参照画像を使用して、エンティティ(人物、製品、ロゴ)のすべてのインスタンスをタイムスタンプ付きで検出できるMultimodal Prompting。

use cases

TwelveLabsのPegasus 1.5は誰が使うべきか?

TwelveLabsのPegasus 1.5は、高度なビデオ理解と構造化データ抽出機能を必要とする幅広いユーザー向けに設計されています。

  • 開発者と企業: AIを活用したビデオアプリケーションの構築、既存システムへの高度なビデオインテリジェンスの統合、および未加工のビデオを大規模に検索可能でAI対応のデータに変換するために。
  • メディア企業とクリエイター: コンテンツの要約、詳細な説明、時間的グラウンディング、長尺コンテンツを物語の単位にセグメンテーションするため、および自然言語を使用して映像の検索、編集、組み立てを支援するために。
  • スポーツ団体: 正確な時間的境界を持つプレー(例:ゴール、ファウル、ダンク)の自動検出、リアルタイムのハイライト作成、およびパフォーマンス分析を可能にするために。
  • ブランドマーケターと広告代理店: ビデオコンテンツ内のブランドの登場、シーンの切り替わり、および文脈的瞬間を特定し、ターゲット広告とコンテンツの収益化を可能にするために。
  • 政府機関とセキュリティオペレーター: コンテンツ分析、コンプライアンス、機密または規則違反コンテンツの検出、およびビデオ映像からの大規模な自動文書化とレポート生成のために。

pricing

TwelveLabsのPegasus 1.5の料金とプラン

TwelveLabsは、PegasusモデルとMarengoモデルの両方へのアクセスを提供するAPI向けに、3つの主要な料金プランを持つFreemiumモデルを提供しています。すべてのプランでレート制限が適用され、使用タイプ(ビデオ/オーディオは期間ベース、テキストはトークンベース、エンドポイントはリクエストベース)によって異なります。適用されるいずれかの制限を超過すると、エラーが発生します。Pegasusモデルの特定のトークンごとの料金が利用可能です。

  • Free Plan: 基本的な制限を無料で提供し、ユーザーがプラットフォームの機能を試すことができます。
  • Developer Plan: 月額利用料に基づいて料金が異なり、Duration per day (DPD)、Duration per hour (DPH)、Requests per day (RPD)、Requests per minute (RPM)、Tokens per day (TPD)、Tokens per minute (TPM)などの次元で段階的に使用制限が増加する3つのティアを提供します。
  • Enterprise Plan: 大規模なデプロイメントと特定の組織要件に対応するカスタム制限とオーダーメイドの料金ソリューションを提供します。
  • 1k Input Text Tokensあたり: $0.001(サードパーティの料金概要に記載されているように、Pegasusモデルの「Input text」に特化)。
  • 1k Output Text Tokensあたり: $0.007(公式Twelve Labs料金計算ツールに記載されているように、Pegasus Analyze APIの使用における「Output text tokens」の場合)。

Pros

  • +Achieves high accuracy in video segmentation and time-based metadata extraction.
  • +Provides reliable, schema-compliant JSON outputs for structured data.
  • +Supports processing of long-form videos up to two hours in a single request.
  • +Offers flexible analysis options including synchronous and efficient batch processing.
  • +Multimodal prompting with image references enhances the precision of video queries.
  • +Demonstrates strong competitive performance against leading general-purpose AI models in video understanding tasks.

Cons

  • Not aligned with HIPAA compliance standards, limiting use in certain healthcare contexts.
  • Detailed pricing for all usage dimensions (e.g., video duration processing) requires deeper inquiry beyond publicly listed token costs.
  • Primarily API-driven, which may necessitate developer resources for full implementation and integration.
  • The platform's focus on enterprise and developer use cases might mean a less intuitive out-of-the-box user interface for non-technical users.

類似ツール

TwelveLabsのPegasus 1.5と競合他社との比較

TwelveLabsはPegasus 1.5を主要なビデオ推論モデルとして位置づけており、汎用モデルと比較して、セグメンテーション品質や構造化JSON出力の信頼性などの主要分野で優れたパフォーマンスを示しています。そのビデオネイティブな推論とTime Based Metadata Extractionが主要な差別化要因です。

1

Mixpeek is designed for teams building video intelligence applications, offering a comprehensive platform that handles ingestion, extraction, indexing, and retrieval of video data.

While both offer video AI, Mixpeek provides an end-to-end solution for building custom video intelligence applications with deep content analysis. TwelveLabs focuses on quick cloud-based video understanding with natural language queries and generative text outputs from its foundation models like Pegasus.

2
Google Video Intelligence API

This API provides robust video annotation and content categorization, deeply integrated with the Google Cloud ecosystem for large-scale analytics and developer-focused applications.

Google Video Intelligence API is a developer-centric service for annotating and categorizing video content within the Google Cloud environment. TwelveLabs offers a more integrated platform for natural language video search and generative text outputs, built on its own multimodal foundation models.

3
Clarifai Video

Clarifai Video is a visual AI platform that provides dedicated video analysis models and a visual workflow builder, enabling non-ML engineers to train and chain custom concept detection models.

Clarifai Video emphasizes customizable concept detection and a user-friendly workflow builder for tailored AI solutions. TwelveLabs, with Pegasus 1.5, focuses on advanced multimodal video understanding, summarization, and the generation of structured, time-based metadata through natural language queries.

4
Memories.ai

Memories.ai is an AI video intelligence platform focused on large-scale search, summarization, and multimodal understanding, with an emphasis on contextual memory and timeline insights for streamlined workflows.

Memories.ai provides a platform for broad video intelligence tasks including search and summarization, leveraging contextual memory. TwelveLabs' Pegasus 1.5 specifically advances video understanding by generating structured, time-based metadata across entire videos, moving beyond clip-based answers to enable schema-first interaction for precise temporal boundaries.

Storkでもっと

関連AIツール

同じカテゴリの他のツール(共通タグで関連付け)