2014 MiningTextSnippetsforImagesonth

From GM-RKB
Jump to navigation Jump to search

Subject Headings:

Notes

Cited By

Quotes

Author Keywords

Abstract

Images are often used to convey many different concepts or illustrate many different stories. We propose an algorithm to mine multiple diverse, relevant, and interesting text snippets for images on the web. Our algorithm scales to all images on the web. For each image, all webpages that contain it are considered. The top-K text snippet selection problem is posed as combinatorial subset selection with the goal of choosing an optimal set of snippets that maximizes a combination of relevancy, interestingness, and diversity. The relevancy and interestingness are scored by machine learned models. Our algorithm is run at scale on the entire image index of a major search engine resulting in the construction of a database of images with their corresponding text snippets. We validate the quality of the database through a large-scale comparative study. We showcase the utility of the database through two web-scale applications: (a) augmentation of images on the web as webpages are browsed and (b) ~an image browsing experience (similar in spirit to web browsing) that is enabled by interconnecting semantically related images (which may not be visually related) through shared concepts in their corresponding text snippets.

References

;

 AuthorvolumeDate ValuetitletypejournaltitleUrldoinoteyear
2014 MiningTextSnippetsforImagesonthLei Zhang
Lucy Vanderwende
Anitha Kannan
Ashish Kapoor
Simon Baker
Krishnan Ramnath
Juliet Fiss
Dahua Lin
Rizwan Ansary
Qifa Ke
Matt Uyttendaele
Xin-Jing Wang
Mining Text Snippets for Images on the Web10.1145/2623330.26233462014