TwelveLabs, a video intelligence company, has unveiled Pegasus 1.6, a model that advances the field of physical AI by transforming egocentric video footage into actionable insights. This development empowers machines such as robots and drones to better understand and interact with real-world environments by offering features like action segmentation, dense captioning, and quality scoring. The model addresses key challenges in processing vast amounts of complex video data, enabling improved training for robotics and other AI systems by utilizing structured, real-world human experiences. This marks a significant step in making video data a more useful asset in teaching machines to operate and reason within physical spaces.
Publicznie pokazujemy glowna mysl. Utworz darmowe konto, aby czytac caly artykul, zapisac go, omowic i polaczyc z kontekstem rynkowym oraz OSINT.
