paper-with-me

홈 › Papers

Convolutional Neural Networks: A Binocular Vision Perspective

2019-12-21 · Yigit Oktar, Diclehan Karakaya, Oguzhan Ulucan, Mehmet Turkan

It is arguable that whether the single camera captured (monocular) image datasets are sufficient enough to train and test convolutional neural networks (CNNs) for imitating the biological neural network structures of the human brain. As human visual system works in binocular, the collaboration of the eyes with the two brain lobes needs more investigation for improvements in such CNN-based visual imagery analysis applications. It is indeed questionable that if respective visual fields of each eye and the associated brain lobes are responsible for different learning abilities of the same scene. There are such open questions in this field of research which need rigorous investigation in order to further understand the nature of the human visual system, hence improve the currently available deep learning applications. This position paper analyses a binocular CNNs architecture that is more analogous to the biological structure of the human visual system than the conventional deep learning techniques. While taking a structure called optic chiasma into account, this architecture consists of basically two parallel CNN structures associated with each visual field and the brain lobe, fully connected later possibly as in the primary visual cortex (V1). Experimental results demonstrate that binocular learning of two different visual fields leads to better classification rates on average, when compared to classical CNN architectures.

📄 PDF Abstract BibTeX arXiv:1912.10201

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

CNN^{2}: Viewpoint Generalization via a Binocular Vision

2019-12-01 · NeurIPS 2019 12 · Wei-Da Chen, Shan-Hung (Brandon) Wu

The Convolutional Neural Networks (CNNs) have laid the foundation for many techniques in various applications. Despite achieving remarkable performance in some tasks, the 3D viewpoint generalizability of CNNs is still fa…

Two-Stream Binocular Network: Accurate Near Field Finger Detection Based On Binocular Images

2018-04-26 · Yi Wei, Guijin Wang, Cairong Zhang, Hengkai Guo 외

Fingertip detection plays an important role in human computer interaction. Previous works transform binocular images into depth images. Then depth-based hand pose estimation methods are used to predict 3D positions of fi…

Fingertip DetectionHand Pose EstimationPose Estimation

3D Motion Perception of Binocular Vision Target with PID-CNN

2025-11-25 · Jiazhao Shi, Pan Pan, Haotian Shi arxiv

This article trained a network for perceiving three-dimensional motion information of binocular vision target, which can provide real-time three-dimensional coordinate, velocity, and acceleration, and has a basic spatiot…

Computational Efficiency

Ego3DPose: Capturing 3D Cues from Binocular Egocentric Views

2023-09-21 · Taeho Kang, Kyungjin Lee, Jinrui Zhang, Youngki Lee

We present Ego3DPose, a highly accurate binocular egocentric 3D pose reconstruction system. The binocular egocentric setup offers practicality and usefulness in various applications, however, it remains largely under-exp…

Egocentric Pose EstimationPose Estimation

BidNet: Binocular Image Dehazing Without Explicit Disparity Estimation

2020-06-01 · CVPR 2020 6 · Yanwei Pang, Jing Nie, Jin Xie, Jungong Han 외

Heavy haze results in severe image degradation and thus hampers the performance of visual perception, object detection, etc. On the assumption that dehazed binocular images are superior to the hazy ones for stereo vision…

3D Object DetectionDisparity EstimationImage Dehazingobject-detection+1