BossaNova at ImageCLEF 2012 Flickr Photo Annotation Task

Sandra Avila, Nicolas Thome, Matthieu Cord, Eduardo Alves do Valle Junior, Arnaldo de Albuquerque Araújo · CLEF (Online Working Notes/Labs/Workshop) · 2012

We present the BossaNova scheme for the ImageCLEF 2012 Flickr Photo Annotation Task. BossaNova is a mid-level image represen- tation, recently developed by our team, that enriches the Bag-of-Words representation, by keeping a histogram of distances between the descrip- tors found in the image and those in the codebook. Our scheme has the advantage of being conceptually simple, non-parametric, and eas- ily adaptable. Compared to other schemes existing in the literature to add information to the Bag-of-Words model, it leads to much more com- pact representations. Furthermore, it complements well the cutting-edge Fisher Vector representations, showing even better results when em- ployed in combination with them. In our participation, we submitted four purely visual runs. Our best result (MiAP = 34.37%) achieved the second rank by MiAP measure among the 28 purely visual submissions and the 18 teams.

Read the paper · More papers on PaperTik