Skip to main navigation Skip to search Skip to main content

Creating saliency maps for video stimuli and comparing with eye movement data

  • Silesian University of Technology

Research output: Contribution to journalConference articlepeer-review

Abstract

The paper presents a deep-learning model that may be used to calculate saliency maps for video content. Classic saliency algorithms take into account only spatial information obtained from an image to generate a saliency map showing regions of higher importance-hence, regions where people are more likely to turn their gaze and attention. Algorithms for video stimuli add temporal information about the movements of objects on a frame-to-frame basis, resulting in more complex three-dimensional architectures. The paper analyses existing models and proposes a model based on one of them. The model's performance is compared with the literature using four widely accepted measures (AUC, NSS, SIM, and CC). It is comparable and, in many cases, even better than already published models. Additionally, because of some improvements in the architecture, it is significantly faster in terms of the number of frames processed per second.

Original languageEnglish
Pages (from-to)2922-2932
Number of pages11
JournalProcedia Computer Science
Volume246
Issue numberC
DOIs
Publication statusPublished - 2024
Event28th International Conference on Knowledge Based and Intelligent information and Engineering Systems, KES 2024 - Seville, Spain
Duration: 11 Nov 202212 Nov 2022

Keywords

  • deep learning
  • eye movements
  • saliency maps
  • video stimuli

ASJC Scopus subject areas

  • General Computer Science

Fingerprint

Dive into the research topics of 'Creating saliency maps for video stimuli and comparing with eye movement data'. Together they form a unique fingerprint.

Cite this