Versos AI Adds Plain-Language Curation for Licensed Video Training Data
The workflow is designed to turn a detailed footage brief into a dataset while keeping ownership, licensing and provenance information alongside the selected material.
Loading page…
The workflow is designed to turn a detailed footage brief into a dataset while keeping ownership, licensing and provenance information alongside the selected material.
Listen to this story
Versos AI is adding agentic curation to its video-data workflow, combining scene- and frame-level search with technical filtering and rights metadata for controlled AI licensing. Buyers can specify requirements such as subjects, actions, duration, resolution, language, format, production quality, and permitted use, then receive a structured dataset rather than a raw search result.
Versos’ agents translate natural-language briefs into search, grading, and dataset-assembly criteria.
Search results include ownership, licensing, and provenance information—not just visual relevance.
The stack uses NVIDIA CUDA Toolkit, Nemotron Ultra, and LangChain; Versos says it is model-agnostic.
Versos AI has launched a workflow that lets AI teams describe needed video training data in plain language, then turns that request into structured search and curation criteria. The system is designed to find, assess and assemble footage prepared for controlled AI licensing into a dataset.
Video-data requests can combine what must appear on screen with technical and legal conditions. Versos says its workflow is designed to handle subjects, actions and environments alongside duration, resolution, format, language, production quality and permitted use. The company says meeting a detailed brief can otherwise require multiple searches, manual inspection and dataset preparation.
The distinguishing feature is not simply prompt-based video search. Versos says the workflow searches footage prepared for controlled AI licensing and maintains its associated ownership, licensing and provenance information. That places documented permissions alongside relevance and technical fit when the system selects material for a training dataset.
Versos says NVIDIA CUDA Toolkit supports inference in its Video Library Intelligence Platform, NVIDIA Nemotron Ultra runs the agents, and LangChain coordinates the workflow. The company says its architecture is model-agnostic, allowing other open-weight or frontier models when customer performance, cost or deployment requirements differ.
Versos plans to demonstrate the capability at IBC2026 in Amsterdam from September 11 to 14. The announcement describes the intended workflow, but the company has not presented independent evidence here about how well it performs across customer datasets or licensing requirements.
Loading discussion...
Join the conversation
What information would make a training dataset usable for you?
Be the first to share a perspective or an experience.
Reader comments
Newest comments first. Replies stay oldest first.