Skip to content

Pegasus 1.5 by TwelveLabs 리뷰

Pegasus 1.5 by TwelveLabs는 시각, 오디오, 음성 정보를 통합하여 비디오 콘텐츠를 깊이 이해하고 정확한 텍스트 설명 및 분석을 생성하는 비디오 우선 언어 모델입니다.

shipped 2026년 4월 21일videofreemium
Domain rating69Monthly visits3.9K/mo
videocoderesearch
Pegasus 1.5 by TwelveLabs - AI tool for pegasus twelvelabs. Professional illustration showing core functionality and features.

핵심 포인트

12026년 4월 20일 정식 출시되었으며, NAB Show 2026에서 시연되었습니다.
2시간 기반 메타데이터 추출(TBM)을 위해 단일 API 호출로 최대 2시간 길이의 비디오를 처리합니다.
3집계된 세분화 품질 벤치마크에서 Gemini 3.1 Pro보다 13.1% 뛰어납니다.
4약 1분 만에 1시간 분량의 비디오를 인덱싱하며, 하루 10,000시간 이상을 지원합니다.

Stork’s verdict on Pegasus 1.5 by TwelveLabs

Pegasus 1.5는 사용자 지정 JSON 스키마를 통해 정밀한 Time Based Metadata Extraction을 제공하지만, API 기반 특성상 개발자 리소스가 필요합니다.

Pegasus 1.5 by TwelveLabs reviewed by Stork AI · stork.ai/ko/pegasus-1-5-by-twelvelabs

Stork Quadrant

Becomes the API· 27/100

Replaceable as a UI, but kept alive as the API the agents call.

TwelveLabs built a capable multimodal video understanding API before the frontier labs caught up. That window is closing. GPT-4o, Gemini 1.5 Pro, and Claude already handle video natively, and they're getting faster and cheaper. There's no proprietary data, no network, no regulatory gate — just a specialized model that bigger players will commoditize.

Claude Sonnet 4.6, scored 2026-05-30

Defensibility · 0/100

  • Physical-world coupling
  • Regulatory moat
  • Network liquidity
  • Proprietary refreshing data
  • High-trust catastrophic workflows
  • Multi-party coordination
  • Brand / community / taste

An LLM alone could replace

  • Summarize what happens in a video by describing its content
  • Transcribe audio and extract key topics or themes from spoken content
  • Answer questions about a video's subject matter given a transcript or description
  • Generate metadata tags or chapter markers for video content

Agent-Readiness · 60/100

  • Verified MCPStork MCP listing: io-twelvelabs-twelvelabs-mcp-server (untested)
  • Listed on agent surfacesStork:io-twelvelabs-twelvelabs-mcp-server
  • Usage-based pricingpricing page heuristic match: https://www.twelvelabs.io/pricing
  • Headless agent auth
  • Public OpenAPIhttps://docs.twelvelabs.io/v1.3/docs/resources/platform-overview
  • Active changeloghttps://www.twelvelabs.io/blog/introducing-pegasus-1-5 (2026-04-19)
  • llms.txthttps://www.twelvelabs.io/llms.txt

Score history · +5 pts over 3 re-scores

How to defend

Go vertical and own the liability: pick one industry where wrong video analysis has real consequences — insurance claims, legal evidence, broadcast compliance — and become the vendor that signs the contract and bears the risk. That's the only move that creates a moat here.

  • Ship an MCP server and list it on Stork — biggest single point gain (+25).
  • Expose API-key auth with a self-serve sandbox tier; remove sales-call gates (+15).

Pegasus 1.5 by TwelveLabs 소개

본사
San Francisco, USA
설립
2020
팀 규모
51-100
투자
Series A

사양

API 제공 여부

예, 공개 API

overview

Pegasus 1.5 by TwelveLabs란 무엇입니까?

Pegasus 1.5 by TwelveLabs는 Twelve Labs가 개발한 멀티모달 비디오 AI 플랫폼으로, 개발자, 기업 및 크리에이터가 원본 비디오 콘텐츠를 대규모로 구조화되고 쿼리 가능한 데이터로 변환할 수 있도록 합니다. 이 플랫폼은 시각, 오디오, 음성 정보를 통합하여 비디오 콘텐츠를 깊이 이해하고, 특히 시간 기반 메타데이터 추출(TBM)을 통해 정확한 텍스트 설명 및 분석을 생성합니다. 이 고급 AI 비디오 추론 모델은 2026년 4월 20일에 정식 출시되었으며 NAB Show 2026에서 선보였습니다. 핵심 기능은 사용자가 사용자 정의 JSON 스키마를 정의하고 단일 API 호출로 최대 2시간 길이의 비디오 콘텐츠에서 타임스탬프가 지정된 구조화된 메타데이터를 받을 수 있도록 하여, 수집 파이프라인, 전처리 또는 인덱싱의 필요성을 없앱니다. TwelveLabs는 Pegasus 1.5의 '비디오 우선' 아키텍처를 강조하며, 이는 비디오를 다차원 볼륨으로 처리하여 자산의 전체 시간적 흐름에 걸쳐 지속적인 추론을 가능하게 합니다.

features

Pegasus 1.5 by TwelveLabs의 주요 기능

Pegasus 1.5 by TwelveLabs는 멀티모달 인텔리전스를 활용하여 포괄적인 비디오 이해 및 분석을 위해 설계된 다양한 기능을 제공합니다.

  • 시각, 오디오, 언어 전반의 멀티모달 인텔리전스로 구동되는 엔터프라이즈 비디오 AI.
  • 자연어 쿼리를 사용한 의미론적 비디오 검색 및 추출.
  • 장편 콘텐츠에서 자동 비디오 요약 및 인사이트 생성.
  • 타임스탬프가 지정된 구조화된 데이터를 위한 사용자 정의 JSON 스키마 출력을 허용하는 시간 기반 메타데이터 추출(TBM).
  • 상황별 설명을 통한 콘텐츠 조정, 규정 준수 및 브랜드 안전 감지.
  • 단일 파이프라인을 통해 약 60배 실시간 속도로 멀티모달 데이터 수집.
  • 약 1분 만에 1시간 분량의 비디오를 인덱싱하며, 하루 10,000시간 이상을 지원합니다.
  • 단일 API 호출로 최대 2시간 길이의 장편 비디오 처리를 지원하며, 컨텍스트를 유지합니다.
  • 타임스탬프가 지정된 인스턴스와 함께 참조 이미지에서 엔티티 식별(사람, 제품, 로고)을 가능하게 하는 멀티모달 프롬프팅.
  • docs.twelvelabs.io에서 포괄적인 문서와 함께 사용자 정의 애플리케이션 및 워크플로 통합을 위한 API 사용 가능.

use cases

누가 Pegasus 1.5 by TwelveLabs를 사용해야 합니까?

Pegasus 1.5 by TwelveLabs는 대량의 비디오 콘텐츠를 처리, 분석하고 인사이트를 도출하기 위한 고급 비디오 인텔리전스를 필요로 하는 조직 및 개인을 위해 설계되었습니다.

  • 개발자 및 기업: 사용자 정의 비디오 인텔리전스 애플리케이션 구축, 비디오 분석 워크플로 자동화, 기존 플랫폼에 멀티모달 AI 통합을 위해.
  • 미디어 및 엔터테인먼트 회사: 수십 년간의 아카이브 영상을 즉시 검색 가능한 구조화된 자산으로 변환하여 의미론적 검색, 자동 하이라이트 생성 및 효율적인 콘텐츠 재사용을 가능하게 합니다.
  • 스포츠 조직: 실시간 하이라이트 생성 및 성과 분석을 위해 특정 선수의 시즌별 행동과 같은 모든 플레이 및 이벤트를 즉시 식별합니다.
  • 광고 대행사 및 브랜드 마케터: 수천 시간의 콘텐츠에서 제품 화면 시간을 측정하고 타겟 광고 및 콘텐츠 수익화를 위한 브랜드 노출을 식별합니다.
  • 보안 운영자 및 규정 준수 팀: 정책 시행 및 브랜드 안전을 위해 무기 또는 폭력과 같은 민감하거나 규칙을 위반하는 콘텐츠를 정확한 타임스탬프 및 상황별 설명과 함께 감지합니다.

pricing

Pegasus 1.5 by TwelveLabs 가격 및 요금제

Twelve Labs는 프리미엄 모델로 운영되며, 개별 개발자부터 대기업에 이르기까지 다양한 사용 규모에 맞춰진 여러 요금제를 제공합니다. 모든 요금제에 걸쳐 속도 제한이 적용되며, 비디오/오디오 처리를 위한 기간 기반, 텍스트 출력을 위한 토큰 기반, 모든 엔드포인트를 위한 요청 기반 등 사용 유형에 따라 다릅니다. 적용 가능한 제한을 초과하면 오류가 발생합니다. Pegasus 모델의 경우, 입력 텍스트는 1,000 토큰당 $0.001이며, Pegasus Analyze API의 출력 텍스트는 1,000 토큰당 $0.007입니다.

  • 무료: 초기 탐색 및 소규모 프로젝트에 적합한 기본 제한을 무료로 제공합니다.
  • 개발자: 제한이 증가하는 세 가지 티어를 제공하며, 가격은 월별 지출에 따라 다릅니다.
  • 엔터프라이즈: 대량 또는 특수 요구 사항이 있는 조직을 위한 맞춤형 제한 및 맞춤형 가격 솔루션을 제공합니다.

Pros

  • +Achieves high accuracy in video segmentation and time-based metadata extraction.
  • +Provides reliable, schema-compliant JSON outputs for structured data.
  • +Supports processing of long-form videos up to two hours in a single request.
  • +Offers flexible analysis options including synchronous and efficient batch processing.
  • +Multimodal prompting with image references enhances the precision of video queries.
  • +Demonstrates strong competitive performance against leading general-purpose AI models in video understanding tasks.

Cons

  • Not aligned with HIPAA compliance standards, limiting use in certain healthcare contexts.
  • Detailed pricing for all usage dimensions (e.g., video duration processing) requires deeper inquiry beyond publicly listed token costs.
  • Primarily API-driven, which may necessitate developer resources for full implementation and integration.
  • The platform's focus on enterprise and developer use cases might mean a less intuitive out-of-the-box user interface for non-technical users.

유사한 도구

Pegasus 1.5 by TwelveLabs 대 경쟁사

TwelveLabs는 Pegasus 1.5를 우수한 비디오 추론 모델로 포지셔닝하며, 특히 비디오를 종종 오디오가 첨부된 이미지 시퀀스로 취급하는 범용 AI 모델과 차별화됩니다. '비디오 우선' 아키텍처와 시간 기반 메타데이터 추출(TBM) 기능은 명확한 이점을 제공합니다.

1

Mixpeek is designed for teams building video intelligence applications, offering a comprehensive platform that handles ingestion, extraction, indexing, and retrieval of video data.

While both offer video AI, Mixpeek provides an end-to-end solution for building custom video intelligence applications with deep content analysis. TwelveLabs focuses on quick cloud-based video understanding with natural language queries and generative text outputs from its foundation models like Pegasus.

2
Google Video Intelligence API

This API provides robust video annotation and content categorization, deeply integrated with the Google Cloud ecosystem for large-scale analytics and developer-focused applications.

Google Video Intelligence API is a developer-centric service for annotating and categorizing video content within the Google Cloud environment. TwelveLabs offers a more integrated platform for natural language video search and generative text outputs, built on its own multimodal foundation models.

3
Clarifai Video

Clarifai Video is a visual AI platform that provides dedicated video analysis models and a visual workflow builder, enabling non-ML engineers to train and chain custom concept detection models.

Clarifai Video emphasizes customizable concept detection and a user-friendly workflow builder for tailored AI solutions. TwelveLabs, with Pegasus 1.5, focuses on advanced multimodal video understanding, summarization, and the generation of structured, time-based metadata through natural language queries.

4
Memories.ai

Memories.ai is an AI video intelligence platform focused on large-scale search, summarization, and multimodal understanding, with an emphasis on contextual memory and timeline insights for streamlined workflows.

Memories.ai provides a platform for broad video intelligence tasks including search and summarization, leveraging contextual memory. TwelveLabs' Pegasus 1.5 specifically advances video understanding by generating structured, time-based metadata across entire videos, moving beyond clip-based answers to enable schema-first interaction for precise temporal boundaries.

Stork에서 더 보기

관련 AI 도구

같은 카테고리의 다른 도구 — 공통 태그로 연결