Categories: Web and IT News

TwelveLabs Unveils Pegasus-1.2 to Efficiently Process and Understand Videos of Varying Lengths and Complexities, Expanding Possibilities for Video AI

Model’s outstanding performance provides nuanced approach to video understanding

TwelveLabs, the video understanding company, announced the release of its Pegasus 1.2 multimodal foundation model, which represents a significant leap forward in industry-grade video language models. Pegasus 1.2 achieves state-of-the-art performance in long video understanding. The model can support videos that are up to one hour long with best-in-class accuracy while also maintaining low latency and competitive pricing. TwelveLabs’ embeddings storage intelligently caches videos, allowing for repeated queries to the same video to be even faster and cheaper. With its latest advances, Pegasus 1.2 serves as a precision tool that delivers business value through its focused, intelligent system design—excelling exactly where production-grade video processing pipelines need it most.

“We are thrilled to debut Pegasus-1.2, which is designed to address the fundamental limitations of existing video language models by introducing a novel approach to spatio-temporal comprehension.”

“Video understanding represents one of the most complex challenges in artificial intelligence, requiring sophisticated models that can simultaneously interpret spatial details, temporal dynamics, and contextual nuances,” said Aiden Lee CTO of TwelveLabs. “We are thrilled to debut Pegasus-1.2, which is designed to address the fundamental limitations of existing video language models by introducing a novel approach to spatio-temporal comprehension.”

Marketing Technology News: Adobe Forecasts Record $240.8 Billion U.S. Holiday Season Online with Black Friday Growth to Outpace Cyber Monday

TwelveLabs’ Pegasus foundation model was built to generate text descriptions about a video, “understanding” its content through analysis of both visual and audio elements. In doing so, it enables the production of summaries, highlights, titles, detailed reports and more based on prompts. This allows users to extract meaningful information from video content through text generation more efficiently than ever before.

Pegasus works in conjunction with TwelveLabs’ Marengo model, a state-of-the-art multimodal embedding model, to bring human-like understanding to videos.

Marketing Technology News: MarTech Interview with Adam Brotman, Co-Founder and Co-Ceo @ Forum3

Pegasus-1.2 Takes Video Understanding to the Next Level

The core innovation of the new Pegasus-1.2 lies in its ability to dynamically balance computational efficiency with comprehensive understanding of videos across varying lengths and complexities. By implementing an advanced vision-encoding strategy and a sophisticated token reduction method, Pegasus-1.2 can capture fine-grained spatial and temporal features while maintaining computational efficiency. This approach enables Pegasus-1.2 to seamlessly transition between understanding short video clips and analyzing extended sequences up to one hour in length, a capability that significantly expands the practical applications of video AI.

Through rigorous testing, the model not only excels in low-level perceptual tasks but also demonstrates advanced reasoning skills across different video understanding domains. Importantly, Pegasus-1.2 achieves these capabilities with a compact architecture, challenging the prevailing assumption that superior performance necessitates exponentially larger model sizes. This positions Pegasus-1.2 as a significant advancement in the field of multimodal AI, offering a more efficient and nuanced approach to video language understanding.

The post TwelveLabs Unveils Pegasus-1.2 to Efficiently Process and Understand Videos of Varying Lengths and Complexities, Expanding Possibilities for Video AI first appeared on PressReleaseCC.

TwelveLabs Unveils Pegasus-1.2 to Efficiently Process and Understand Videos of Varying Lengths and Complexities, Expanding Possibilities for Video AI first appeared on Web and IT News.

awnewsor

Recent Posts

MiMedia Announces Distribution Agreement with Leading Telco Africell to Deliver Personal Cloud Services to Consumers Across All of Africell’s African Markets

The post MiMedia Announces Distribution Agreement with Leading Telco Africell to Deliver Personal Cloud Services…

30 minutes ago

Sequans to Participate in the H.C. Wainwright 28th Annual Global Investment Conference, September 14-16, 2026

The post Sequans to Participate in the H.C. Wainwright 28th Annual Global Investment Conference, September…

30 minutes ago

WELL Health Announces WELLTRUST(TM) Surpasses 100,000 Patient Consents

The post WELL Health Announces WELLTRUST(TM) Surpasses 100,000 Patient Consents first appeared on PressReleaseCC. WELL…

31 minutes ago

Physicists Catch Gravity Shaping a Quantum Wave for the First Time

Physicists have finally watched gravity leave its mark on a quantum object in free fall.…

31 minutes ago

Google’s Android Security Update Blocks Custom ROMs, Developer Tools and Apps

Google’s latest Android security enhancements have left many users frustrated after the company introduced stricter…

31 minutes ago

Tesla Driver Assistance Ran a Stop Sign and Killed a Driver. Its Own Data Proves the System Was Active

A Tesla Model 3 tore through a stop sign in a rural New Jersey township…

31 minutes ago

This website uses cookies.