Twelve Labs launches Jockey video AI agent
Jockey is a conversational video intelligence agent by Twelve Labs designed to search, analyze, and reason across entire media libraries. Currently in research preview, the agent can plan multi-step video workflows, edit clips, and integrate with LLMs like Claude via the Model Context Protocol.
Jockey represents a significant evolution from passive search to active agentic video manipulation, but its ultimate utility depends heavily on processing cost, query speed, and integration friction for large media libraries.
* Native video-first reasoning bypasses fragile and expensive transcription/tagging middleware.
* Supporting Model Context Protocol (MCP) allows seamless ecosystem integration with popular LLM assistants out-of-the-box.
* Latency and API cost scaling remain the primary hurdles for deploying agentic video search across massive enterprise archives.
* The transition from an open-source demo to a research preview signals a push towards commercializing custom video agent infrastructure.
DISCOVERED
1d ago
2026-07-21
PUBLISHED
1d ago
2026-07-21
RELEVANCE
AUTHOR
[REDACTED]