Illustration for: TwelveLabs Raises $100M to Build Video Superintelligence

TwelveLabs Raises $100M to Build Video Superintelligence

TwelveLabs raised a $100M Series B co-led by NEA and NAVER Ventures to expand beyond video-understanding models into a full-stack agentic system combining perception, knowledge and reasoning for video.

TC
By the Funding Desk
Edited by Trace Cohen · Early-stage VC & angel · Founder, New York Venture Partners
1 min read
ShareXLinkedInEmail

THE RUNDOWN

1

TwelveLabs closed a $100M Series B co-led by NEA and NAVER Ventures, with Amazon, Radical Ventures, Korea Investment Partners, Index Ventures, Quadrille Capital and Red Bull Ventures also participating

2

The company is expanding beyond core video-understanding models into what it calls a full-stack agentic intelligence system for video, combining perception, knowledge and reasoning in one architecture

3

Amazon's participation as both investor and likely infrastructure partner mirrors a pattern of hyperscalers taking strategic stakes in AI application-layer startups building on top of their cloud

4

Video remains one of the least-solved modalities for foundation models relative to text and images, making a well-capitalized specialist a real bet against multimodal giants like Google's Gemini adding video understanding natively

TC

The VC Read · Trace's Take

Trace Cohen

TwelveLabs' bet only works if video-specific accuracy stays meaningfully ahead of what Gemini and GPT ship natively -- that gap has been closing all year, and $100M buys maybe 18 months of runway to prove the specialist thesis before the giants catch up. I'd want to see enterprise ACV growth, not just funding size, before underwriting this as a durable category leader rather than a feature Google eventually ships for free.

Analysis

TwelveLabs raised $100 million in a Series B round co-led by NEA and NAVER Ventures, with Amazon, Radical Ventures, Korea Investment Partners, Index Ventures, Quadrille Capital and Red Bull Ventures also participating, according to GlobeNewswire.

From Video Search to Full-Stack Agents

The company started by building models that could search and understand video content -- finding specific moments, actions or objects across large video libraries -- and is now expanding into what it describes as a full-stack agentic intelligence system for video, combining perception, knowledge and reasoning into a single architecture rather than a narrower search tool.

Competing With the Multimodal Giants

That's a real bet against a difficult trend: Google's Gemini and OpenAI's latest models are adding native video understanding directly into their general-purpose foundation models, narrowing the gap that specialist video-AI companies like TwelveLabs have relied on. Amazon's participation as an investor, alongside its cloud relationship, suggests TwelveLabs is positioning itself as infrastructure hyperscalers want to build on rather than compete with directly.

What to watch: whether TwelveLabs can keep a meaningful accuracy or latency edge over general-purpose multimodal models on video-specific tasks as the gap continues to close, and whether its next round comes with disclosed enterprise revenue rather than just capability benchmarks.

ShareXLinkedInEmail

Key Sources

2 sources

THE WIRE in your inbox— Tech, startup & VC news with Trace's take. Free, no spam.