Passer à la navigation principale Passer à la recherche Passer au contenu principal

Multi-modal query expansion for video object instances retrieval

  • Institut Mines-Télécom

Résultats de recherche: Le chapitre dans un livre, un rapport, une anthologie ou une collectionContribution à une conférenceRevue par des pairs

Résumé

In this paper we tackle the issue of object instances retrieval in video repositories using minimum information from the user (e.g., textual description/tags). Starting for a set of tags, images containing the object of interest are crawled from popular image search engines and repositories (e.g., Bing1, Fickr2, Google3) and the positive and most representative instances of the object are automatically identified. These positive images are then used to generate a visual query descriptor and to retrieve videos containing the object of the interest. This multi-modal approach makes it possible to retrieve video content through images obtained from textual queries, without the use of any advanced learning technique. We test out method on the Flickr corpus of the TRECVID 2012 Instance Search Task.

langue originaleAnglais
titreProceedings of the 13th IAPR International Conference on Machine Vision Applications, MVA 2013
EditeurMVA Organization
Pages214-217
Nombre de pages4
ISBN (imprimé)9784901122139
étatPublié - 1 janv. 2013
Modification externeOui
Evénement13th IAPR International Conference on Machine Vision Applications, MVA 2013 - Kyoto, Japon
Durée: 20 mai 201323 mai 2013

Série de publications

NomProceedings of the 13th IAPR International Conference on Machine Vision Applications, MVA 2013

Une conférence

Une conférence13th IAPR International Conference on Machine Vision Applications, MVA 2013
Pays/TerritoireJapon
La villeKyoto
période20/05/1323/05/13

Empreinte digitale

Examiner les sujets de recherche de « Multi-modal query expansion for video object instances retrieval ». Ensemble, ils forment une empreinte digitale unique.

Contient cette citation