Abstract
在這篇論文中,依照生理學上的一些發現建立了一個前級的雙眼知覺模型。就是要這個模型主要是以為基礎,在這個模型中我們提出一個關於在左右眼偏重性結構中視角差選擇性神經細胞是如何運作的假說。使這個模型輸出用來模擬人類知覺的影像,會符合心理學在中視覺方向的定律。 這個模型分為兩個部分,第一個是視角差產生階段,第二個是雙眼影像合成階段。在第一階段包括三個步驟,第一個作特徵抽取,第二個作視角差的估計,第三個作主要頻率及能量的估計。在第二個階段中,我們提出一個關於神經細胞間如何運作的假說,並利用一個雙眼對抗機制來實現它。選擇性注意力機制被加入用來選擇須要的資訊以合成影像。另外也使用了抑制及吸收機制使合成的影像符合幾何上的關係。而這個階段輸出的結果就是用來表示人類雙眼知覺。 在這篇論文中,我們做了五個實驗,這些實驗都是以前的學者以人為對象,探討雙眼知覺的情形。在這□,我們以所提出的模型,用電腦來模擬,看看結果是不是相同。第一個是用以簡介我們的模型,第二個實驗是來說明主要頻率的作用,第三個實驗是用以說明心理學中的視覺方向定律,第四個實驗主要是用以展示抑制及吸收機制,第五個實驗是用以說明我們模型輸出的結果可以提供資訊給大腦更高階層作判斷或動作。 我們的模型和傳統模型不一樣的地方在於傳統模型在求出視角差後,會先求出深度的資訊,再利用深度的資訊重新計算出合成的影像,而我們的模型在求出視角差後,依照生理學上的結構,將雙眼影像及深度資訊分開計算,如此不僅可以得到傳統模型的結果,也可以得到錯覺方面的資訊,而符合人類的雙眼知覺。以我們的模型為基礎,再加入彩色佑覺模型及三度空間模型,可以構成完整的人類視覺前級系統,用來瞭解人類認知運作情形。應用上,則可以推廣至立體影像或立體電視上,看看那些資料可以不必傳輸而那些是必要的。或用於娛樂的設備上,如產生錯覺影像等。In this thesis, a binocular perception model in the low levelvision processes was constructed based on physiology. Ahypothesis about the operations of binocular selectiveneurons in the ocular dominance columns was provided. Theoutputs of this model were obeyed the laws of visualdirection in psychology. There were two stages in theproposed model. The first stage was the disparitygenerating stage. Three types of information -disparities, dominant frequencies, and energies - weregenerated. The second stage was the binocular imagegenerating stage which received the outputs of the first stage.The hypothesis about the relationships between the neuronswas provided in this stage and was performed by a binocularopponent mechanism. The attention selectivity was used tochoosing the retinal image in the image formating step. Theoutputs of this stage were used to present the binocularperception. Five experiments demonstrated by Well (1792) andOno (1976,1986) were redone using the proposed model bycomputer simulations. The results of computer simulationswere consistent with the human perception.