overview
What is Pegasus 1.5 by TwelveLabs?
Pegasus 1.5 by TwelveLabs is a video language model developed by Twelve Labs that enables developers and enterprises to analyze, search, and understand video content using multimodal AI. It transforms raw video into structured, queryable data, specializing in video-to-text generation and segmentation. Released on April 20, 2026, Pegasus 1.5 represents a significant advancement in video reasoning models, offering capabilities such as Time Based Metadata Extraction (TBM) which allows users to define custom JSON schemas for timestamped, structured metadata from videos up to two hours long. The model processes multiple modalities within videos—visual, audio, and textual information—to produce contextually relevant text and structured data. It supports direct video analysis from URLs, assets, or base64 strings, and features multimodal prompting, allowing the inclusion of reference images for enhanced context. Pegasus 1.5 utilizes a shared context window of 261,120 tokens for input and output, supporting responses up to 98,304 tokens. As of May 28, 2026, it supports synchronous analysis, and batch analysis for up to 1,000 requests was introduced on June 18, 2026.
