Visual phraselet: Refining spatial constraints for large scale image search

Zheng, L; Wang, S

Visual phraselet: Refining spatial constraints for large scale image search

Zheng, L

Wang, S

Permalink

Publication Type:: Journal Article
Citation:: IEEE Signal Processing Letters, 2013, 20 (4), pp. 391 - 394
Issue Date:: 2013-03-19

Closed Access

	Filename	Description	Size
	06471751.pdf	Published Version	843.34 kB	Adobe PDF	View/Open

Copyright Clearance Process

Recently Added
In Progress
Closed Access

This item is closed access and not available.

Full metadata record

Field	Value	Language
dc.contributor.author	Zheng, L https://orcid.org/0000-0002-1464-9500	en_US
dc.contributor.author	Wang, S	en_US
dc.date.issued	2013-03-19	en_US
dc.identifier.citation	IEEE Signal Processing Letters, 2013, 20 (4), pp. 391 - 394	en_US
dc.identifier.issn	1070-9908	en_US
dc.identifier.uri	http://hdl.handle.net/10453/118042
dc.description.abstract	The Bag-of-Words (BoW) model is prone to the deficiency of spatial constraints among visual words. The state of the art methods encode spatial information via visual phrases. However, these methods discard the spatial context among visual phrases instead. To address the problem, this letter introduces a novel visual concept, the Visual Phraselet, as a kind of similarity measurement between images. The visual phraselet refers to the spatial consistent group of visual phrases. In a simple yet effective manner, visual phraselet filters out false visual phrase matches, and is much more discriminative than both visual word and visual phrase. To boost the discovery of visual phraselets, we apply the soft quantization scheme. Our method is evaluated through extensive experiments on three benchmark datasets (Oxford 5 K, Paris 6 K and Flickr 1 M). We report significant improvements as large as 54.6% over the baseline approach, thus validating the concept of visual phraselet. © 1994-2012 IEEE.	en_US
dc.relation.ispartof	IEEE Signal Processing Letters	en_US
dc.relation.isbasedon	10.1109/LSP.2013.2249513	en_US
dc.subject.classification	Networking & Telecommunications	en_US
dc.title	Visual phraselet: Refining spatial constraints for large scale image search	en_US
dc.type	Journal Article
utslib.citation.volume	4	en_US
utslib.citation.volume	20	en_US
utslib.for	0906 Electrical and Electronic Engineering	en_US
utslib.for	0801 Artificial Intelligence and Image Processing	en_US
utslib.for	1005 Communications Technologies	en_US
pubs.embargo.period	Not known	en_US
pubs.organisational-group	/University of Technology Sydney
pubs.organisational-group	/University of Technology Sydney/Faculty of Engineering and Information Technology
pubs.organisational-group	/University of Technology Sydney/Faculty of Engineering and Information Technology/School of Software
pubs.organisational-group	/University of Technology Sydney/Strength - CAI - Centre for Artificial Intelligence
utslib.copyright.status	closed_access
pubs.issue	4	en_US
pubs.publication-status	Published	en_US
pubs.volume	20	en_US

Abstract:

The Bag-of-Words (BoW) model is prone to the deficiency of spatial constraints among visual words. The state of the art methods encode spatial information via visual phrases. However, these methods discard the spatial context among visual phrases instead. To address the problem, this letter introduces a novel visual concept, the Visual Phraselet, as a kind of similarity measurement between images. The visual phraselet refers to the spatial consistent group of visual phrases. In a simple yet effective manner, visual phraselet filters out false visual phrase matches, and is much more discriminative than both visual word and visual phrase. To boost the discovery of visual phraselets, we apply the soft quantization scheme. Our method is evaluated through extensive experiments on three benchmark datasets (Oxford 5 K, Paris 6 K and Flickr 1 M). We report significant improvements as large as 54.6% over the baseline approach, thus validating the concept of visual phraselet. © 1994-2012 IEEE.

Please use this identifier to cite or link to this item:

http://hdl.handle.net/10453/118042