Understanding Character Recognition using Visual Explanations Derived from the Human Visual System and Deep Networks

Ralekar, Chetan; Choudhary, Shubham; Gandhi, Tapan Kumar; Chaudhury, Santanu

Computer Science > Computer Vision and Pattern Recognition

arXiv:2108.04558 (cs)

[Submitted on 10 Aug 2021 (v1), last revised 29 Aug 2021 (this version, v2)]

Title:Understanding Character Recognition using Visual Explanations Derived from the Human Visual System and Deep Networks

Authors:Chetan Ralekar, Shubham Choudhary, Tapan Kumar Gandhi, Santanu Chaudhury

View PDF

Abstract:Human observers engage in selective information uptake when classifying visual patterns. The same is true of deep neural networks, which currently constitute the best performing artificial vision systems. Our goal is to examine the congruence, or lack thereof, in the information-gathering strategies of the two systems. We have operationalized our investigation as a character recognition task. We have used eye-tracking to assay the spatial distribution of information hotspots for humans via fixation maps and an activation mapping technique for obtaining analogous distributions for deep networks through visualization maps. Qualitative comparison between visualization maps and fixation maps reveals an interesting correlate of congruence. The deep learning model considered similar regions in character, which humans have fixated in the case of correctly classified characters. On the other hand, when the focused regions are different for humans and deep nets, the characters are typically misclassified by the latter. Hence, we propose to use the visual fixation maps obtained from the eye-tracking experiment as a supervisory input to align the model's focus on relevant character regions. We find that such supervision improves the model's performance significantly and does not require any additional parameters. This approach has the potential to find applications in diverse domains such as medical analysis and surveillance in which explainability helps to determine system fidelity.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2108.04558 [cs.CV]
	(or arXiv:2108.04558v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2108.04558

Submission history

From: Chetan Ralekar [view email]
[v1] Tue, 10 Aug 2021 10:09:37 UTC (7,497 KB)
[v2] Sun, 29 Aug 2021 16:51:48 UTC (4,858 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Understanding Character Recognition using Visual Explanations Derived from the Human Visual System and Deep Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Understanding Character Recognition using Visual Explanations Derived from the Human Visual System and Deep Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators