Multimodal search and data flywheels for physical AI
We have taken VLM-based video search from prototype to production: models caption driving clips, and the captions become embeddings for semantic retrieval over a petabyte-scale video datalake. We build the retrieval foundation to support data flywheel that powers end to end training loops.