Self-Attention Based Vision Processing for Prosthetic Vision

Jack White; Jaime Ruiz-Serra; Stephen Petrie; Tatiana Kameneva; Chris McCarthy

doi:10.1109/EMBC40787.2023.10341053

Self-Attention Based Vision Processing for Prosthetic Vision

Annu Int Conf IEEE Eng Med Biol Soc. 2023 Jul:2023:1-4. doi: 10.1109/EMBC40787.2023.10341053.

Authors

Jack White, Jaime Ruiz-Serra, Stephen Petrie, Tatiana Kameneva, Chris McCarthy

PMID: 38083046
DOI: 10.1109/EMBC40787.2023.10341053

Abstract

We investigate Self-Attention (SA) networks for directly learning visual representations for prosthetic vision. Specifically, we explore how the SA mechanism can be leveraged to produce task-specific scene representations for prosthetic vision, overcoming the need for explicit hand-selection of learnt features and post-processing. Further, we demonstrate how the mapping of importance to image regions can serve as an explainability tool to analyse the learnt vision processing behaviour, providing enhanced validation and interpretation capability than current learning-based methods for prosthetic vision. We investigate our approach in the context of an orientation and mobility (OM) task, and demonstrate its feasibility for learning vision processing pipelines for prosthetic vision.

MeSH terms

Image Processing, Computer-Assisted / methods
Learning
Vision, Ocular
Visual Perception
Visual Prosthesis*