The THUS10000 benchmark dataset comprises of 10,000 images, each of which has an unambiguous salient object and the object region is accurately annotated …
attention, saliency, salient object detection, segmentation, visualThe Caltech Game Covers dataset consists of CD/DVD covers of video games. The set was downloaded from freecovers.net during the summer of 2008. The set in…
caltech, classification, cover, game, hierarchy, retrieval, taxonomyThe ImageNET dataset is the latest dataset by Li Fei-Fei containing various dataset ranging from 1000 to 10000 categories.
image classification, object segmentation, retrievalThis dataset provides a collection of web images and 3D models for research on landmark recognition (especially for methods based on 3D models). We hope i…
3d, classification, codebook, feature, flickr, landmark, matching, recognition, reconstruction, retrievalThe domain-specific personal videos highlight dataset from the paper [1] describes a fully automatic method to train domain-specific highlight ranker for…
action, domain, human, recognition, saliency, summarization, video, wearableThe 3DVis dataset includes a set of 12 heterogeneous scenes for testing 3D scene registration and analysis methods. Models include homogeneous shapes, rep…
3d, matching, reconstruction, registration, shape, symmetryThe Compact Descriptors for Visual Search Patches Dataset (CDVS) is a dataset comprised of pairwise image patches. MPEG is a standard titled Compact Desc…
descriptor, feature, matching, mpeg, patch, retrievalHumans have used sketching to depict our visual world since prehistoric times. Even today, sketching is possibly the only rendering technique readily avai…
image retrieval, matching, partial, shape retrieval, sketchYahoo Flickr Creative Commons 100M (YFCC100M) dataset contains a list of photos and videos. This list is compiled from data available on Yahoo! Flickr. Al…
3d, clustering, community, detection, flickr, image, internet, landmark, recognition, reconstruction, socialThe San Francisco Landmark Dataset for Mobile Landmark Recognition is a set of images and query images for localization. We present the San Francisco La…
calibration, city, gps, landmark, localization, mobile, retrieval, sanfrancisco, urbanThe Webcam Interestingness dataset consists of 20 different webcam streams, with 159 images each. It is annotated with interestingness ground truth, acqui…
classification, interest, ranking, retrieval, video, weather, webcamThis material is supplementary to Michael Stark, Bernt Schiele. How Good are Local Features for Classes of Geometric Objects. Eleventh IEEE Internatio…
binary, classification, object, shape, toolThe Multi-FoV synthetic datasets are two synthetic scenes (vehicle moving in a city, and flying robot hovering in a confined room). For each scene, three …
blender, camera, fov, groundtruth, odometry, synthetic, visualThe FlickrLogos-32 dataset contains photos showing brand logos and is meant for the evaluation of multi-class logo recognition as well as logo retrieval m…
classification brand boundingbox, detection, flickr, image, logo, machine learning, object recognition, retrievalThe RGB-D Person Re-identification dataset is for person re-identification using depth information. The main motivation is that the standard techniques (s…
3d, classification, depth, identification, pedestrian, shapeAn Annotated Dataset For Near-Duplicate Detection In Personal Photo Collections Managing photo collections involves a variety of image quality assessmen…
copyright, detection, duplicate, groundtruth, retrievalThe Caltech Buildings dataset consists of images taken for 50 buildings around the Caltech campus. Five different images were taken for each building from…
building, caltech, hierarchy, retrieval, taxonomy, urbanThe 3D shape description dataset consists of multiple sub-datasets Descriptor Matching - Dataset 1 & 2 (Stanford) These datasets, created from some of …
3d, benchmark, description, matching, reconstruction, registration, shapeWe introduce a benchmark for evaluating the performance of large scale sketch-based image retrieval systems. The necessary data is acquired in a controlle…
image retrieval, matching, partial, shape retrieval, sketchThis work attempts to provide two Hand Images Databases for hand biometrics: one is created using a mobile phone camera of modest quality, which we call…
authentication, biometric/ hand geometry, identification, mobile, person, segmentation, shape, webcamWe would like to announce the release of PASCAL-Context dataset. We augmented PASCAL VOC 2010 dataset with annotations for 400+ additional categories. In …
benchmark, category, dense, pascal, recognition, segmentation, semantic, shapeGroup emotion recognition in images - Happiness Intensity labels for group of people in images. The images have been collected from Flickr using keyword s…
behavior, emotion, facial expression, flickr, group, human, wildThe VOT2016 pixel-wise annotations dataset contains pixel-wise per-frame annotations for sequences from VOT2016 dataset. The annotation is in a form of BW…
annotation, mask, object, segmentation, tracking, visualThe Salient Montages is a human-centric video summarization dataset from the paper [1]. In [1], we present a novel method to generate salient montages f…
human, montage, saliency, summarization, video, wearableThe Google Street View dataset contains 62,058 high quality Google Street View images. The images cover the downtown and neighboring areas of Pittsburgh, …
address, google, gps, localization, manhattan, panorama, pittsburgh, retrieval, sphere, streetview, urban