Multimodal Surrogates for Video Browsing
Wei Ding, Gary Marchionini, Dagobert Soergel
Abstract
Three types of video surrogates - visual (keyframes), verbal (keywords/phrases), and combination of the two - were designed and studied in a qualitative investigation of user cognitive processes. The results favor the combined surrogates in which verbal information and images reinforce each other, lead to better comprehension, and may actually require less processing time. The results also highlight image features users found most helpful. These findings will inform the interface design and video representation for video retrieval and browsing.
Create a lesson
Related papers
Gender and the Production of Research Impact
Sanger Wagner, Charles Rahal, Melinda C. Mills
A Comparative Evaluation of Digitization Pipelines for Historiographical Sources
Marina Gómez Rey, Patricia Callejo, Mario Muñoz-Organero et al.
COCI: Conference Organisers and Content Identifier
Angelo Salatino, Francesco Osborne, Alexis Vizcaino et al.
Towards a Definition of the Computational Architecture of Open Scholarly Infrastructures
Ivan Heibi, Mario Petrella, Angelo Di Iorio et al.
Beyond FAIR Data: Instrument Traces for Active and Autonomous Scientific Experimentation
Sergei V. Kalinin, Boris N. Slautin, Yu Liu et al.
Taxonomy-aware distances between scholarly topic profiles via an exact simplex embedding
Dmitry Gubanov, Alexander Chkhartishvili