Pegasus 1.6 brings video understanding to physical AI, says TwelveLabs
AI-summarised brief · reviewed before publication
TwelveLabs released Pegasus 1.6, a video‑understanding model that adds temporal context, spatial reasoning, and task completion judgment to egocentric footage. The update targets physical AI, enabling robots, drones, and autonomous vehicles to interpret real‑world actions from a first‑person perspective without specialized cameras. Pegasus 1.6 supports five workflows—action segmentation, dense captioning, quality scoring, search, and curation—allowing developers to convert raw video into structured, reviewable training data. The company claims the model speeds up model training and improves safety in physical environments.
💡 Why It Matters
- · By turning everyday egocentric video into actionable knowledge, Pegasus 1.6 lowers the barrier for robotics teams to train on realistic human behavior, accelerating deployment of safer autonomous systems.