Abstract
Scale Invariant Feature Transform (SIFT) descriptor can represent the object in detail, and is robust to variations due to image scaling and illumination changes. The challenge of using such descriptor to perform image retrieval in a large scale database is the high computational complexity. In this paper, we present the bag of words model combined with SIFT to reduce the computation cost. The average precision we get is about 30%. © 2014 IEEE.