The High Definition Analytics (HDA) dataset is a multi-camera High-Resolution image sequence dataset for research on High-Definition surveillance: Pedestr…
benchmark, camera, detection, high-definition, human, indoor, lisbon, multiview, network, pedestrian, re-identification, surveillance, tracking, videoSVHN is a real-world image dataset for developing machine learning and object recognition algorithms with minimal requirement on data preprocessing and fo…
classification, detection, number, real, recognition, streetside, streetview, text, urban, worldDatabase contains 798 images of 114 persons, with 7 images per person and is freely available for research purposes. All images were taken in supervised c…
biometry, face, human, illumination, lighting, pedestrian, person, recognitionThe Where Who Why (WWW) dataset provides 10,000 videos with over 8 million frames from 8,257 diverse scenes, therefore offering a superior comprehensive d…
crowd, detection, flow, optical, pedestrian, recognition, surveillance, videoThe FaceScrub dataset comprises a total of 107818 unconstrained face images of 530 celebrities crawled from the Internet, with about 200 images per person…
celebrity, detection, face, human, people, recognitionThe PASCAL VOC is augmented with segmentation annotation for semantic parts of objects. For example, for the person category, we provide segmentation mask…
detection, human, object, part, pascal, pedestrian, recognition, segmentation, semanticThe KTH Multiview Football dataset contains 771 images of football players includes images taken from 3 views at 257 time instances 14 annotated body join…
camera, detection, game, multitarget, multiview, object, outdoor, pedestrian, pose, recognition, soccer, trackingThe UrbanStreet dataset used in the paper can be downloaded here [188M] . It contains 18 stereo sequences of pedestrians taken from a stereo rig mounted o…
detection, human, multitarget, pedestrian, recognition, segmentation, tracking, urban, videoThe ICG Multi-Camera and Virtual PTZ dataset contains the video streams and calibrations of several static Axis P1347 cameras and one panoramic video from…
calibration, camera, crowd, detection, graz, multitarget, multiview, network, object, outdoor, panorama, pedestrian, tracking, videoThe Our Database of Faces (ORL) dataset contains ten different images of each of 40 distinct subjects. For some subjects, the images were taken at differe…
expression, face, human, illumination, recognitionThe Stanford 40 Actions dataset contains images of humans performing 40 actions. In each image, we provide a bounding box of the person who is performing …
action, boundingbox, detection, human, recognitionThe Person Re-ID (PRID) 2011 dataset was created in co-operation with the Austrian Institute of Technology for the purpose of testing person re-identifica…
appearance, change, classification, graz, identification, illumination, multiview, pedestrian, trajectoryWe share our omnidirectional and panoramic image dataset (with annotations) to be used for human and car detection. Please reach through: http://cvrg.iyt…
car, detection, human, omnidirection, panorama, recognitionPOS Labeled Faces in the Wild, a collection of face which is proposed for studying face identification in unconstrained environment, its purpose is servin…
face, identification, recognition, registration, wildSince the publicly available face image datasets are often of small to medium size, rarely exceeding tens of thousands of images, and often without age in…
age, biometry, detection, face, imdb, recognition, wikipediaWelcome to the homepage of the gvvperfcapeva datasets. This site serves as a hub to access a wide range of datasets that have been created for projects of…
action, depth, face, human, mesh, multiview, pose, reconstruction, tracking, videoDataset contains 1000 images of 100 persons, with 10 images per person and is freely available. All images were acquired by cropping ears from images from…
biometry, ear, human, lighting, pedestrian, person, recognitionThe ICG Multi-Camera datasets consist of Easy Data Set (just one person) Medium Data Set (3-5 persons, used for the experiments) Hard Data Set (crowd…
calibration, camera, detection, graz, indoor, multitarget, multiview, object, pedestrian, tracking, videoThe PIROPO database (People in Indoor ROoms with Perspective and Omnidirectional cameras) comprises multiple sequences recorded in two different indoor ro…
detection, fisheye, human, indoor, omnidirectional, people, perspective, room, surveillanceThe PETS 2009 dataset contains 3 parts showing multi-view sequences containing pedestrians walking in an outdoor environment. The parts are used for perso…
detection, frontview, human, occlusion multitarget, outdoor, overlap, pedestrian, trackingThe ICG Lab 6 (Multi-Camera Multi-Object Tracking) dataset contains 6 indoor people tracking scenarios recorded at our laboratory using 4 static Axis P134…
calibration, camera, detection, evaluation, graz, laboratory, multiview, object, pedestrian, segmentation, trackingThe crowd datasets are collected from a variety of sources, such as UCF and data-driven crowd datasets. The sequences are diverse, representing dense crow…
anomaly, crowd, detection, human, pedestrian, scene, understanding, videoChairGest is an open challenge / benchmark. The task consists in spotting and recognizing gestures from multiple synchronized sensors: 1 Kinect and 4 Xse…
benchmark, detection, gesture, human, kinect, recognitionYahoo Flickr Creative Commons 100M (YFCC100M) dataset contains a list of photos and videos. This list is compiled from data available on Yahoo! Flickr. Al…
3d, clustering, community, detection, flickr, image, internet, landmark, recognition, reconstruction, socialThe multi-modal/multi-view datasets are created in a cooperation between University of Surrey and Double Negative within the EU FP7 IMPART project. The …
3d, action, color, dynamic, emotion, face, human, indoor, lidar, model, multi-mode, multi-view, outdoor, rgbd, videoThe Mall dataset was collected from a publicly accessible webcam for crowd counting and profiling research. Ground truth: Over 60,000 pedestrians were …
counting, crowd, detection, indoor, pedestrian, tracking, video, webcamThe TUG (Timed Up and Go test) dataset consists of actions performed three times by 20 volunteers. The people involved in the test are aged between 22 and…
accelerometer, action, depth image processing - tug, human, kinect, recognition, time, video, wearableThis dataset contains 12,995 face images which are annotated with (1) five facial landmarks, (2) attributes of gender, smiling, wearing glasses, and head …
attribute, cnn, deep learning, detection, face, landmark detectionThe Swedish Traffic Sign Recognition provides Matlab code for parsing the annotation files and displaying the results. Part0 for each set contains the ann…
city, detection, recognition, sign, traffic, urbanWe present a new large-scale dataset that contains a diverse set of stereo video sequences recorded in street scenes from 50 different cities, with high q…
car, cities, detection, pedestrian, person, segmentation, semantic, stereo, urban, video, weaklyThe MSR Action datasets is a collection of various 3D datasets for action recognition. See details http://research.microsoft.com/en-us/um/people/zliu/a…
3d, action, detection, recognition, reconstruction, videoThe RGB-D Person Re-identification dataset is for person re-identification using depth information. The main motivation is that the standard techniques (s…
3d, classification, depth, identification, pedestrian, shapeThe Pittsburgh Fast-food Image dataset (PFID) consists of 4545 still images, 606 stereo pairs, 3033600 videos for structure from motion, and 27 privacy-pr…
classification, food, laboratory, real, recognition, reconstruction, videoDataset A (former NLPR Gait Database) was created on Dec. 10, 2001, including 20 persons. Each person has 12 image sequences, 4 sequences for each of the …
action, biometry, classification, foot, gait, human, motion, pressure, recognitionSome datasets and evaluation tools are provided on this page for four different computer vision and computer graphics problems. Population counting Lin…
3d, counting, crowd, detection, groundtruth, line, network, object, pedestrian, pointcloud, reconstruction, road, surface, urbanThe dataset captures 25 people preparing 2 mixed salads each and contains over 4h of annotated accelerometer and RGB-D video data. Annotated activities co…
action, activity, classification, detection, recognition, tracking, videoRobust Multi-Person Tracking from Mobile Platforms In all cases, data was recorded using a pair of AVT Marlins F033C mounted on a chariot respectively a…
color, pedestrian, sequence, trackingThe CHALEARN Multi-modal Gesture Challenge is a dataset +700 sequences for gesture recognition using images, kinect depth, segmentation and skeleton data.…
action, depth, gesture, human, illumination, kinect, recognition, segmentation, skeletonPenn-Fudan Pedestrian Detection and Segmentation
background, detection, motion, pedestrian, segmentationThe Shefeld Kinect Gesture (SKIG) dataset contains 2160 hand gesture sequences (1080 RGB sequences and 1080 depth sequences) collected from 6 subjects. Al…
action, depth, gesture, human, illumination, kinect, recognitionThe Longterm Pedestrian dataset consists of images from a stationary camera running 24 hours for 7 days at about 1 fps. It used for adaptive detection an…
background, change, coffee, detection, graz, illumination, indoor, multitarget, pedestrian, robustA New Color Image Database for Benchmarking of Face Detection Techniques and Human Skin Segmentation Techniques. A new color face image database for be…
benchmarking, detection, face, segmentation, skinAt Udacity, we believe in democratizing education. How can we provide opportunity to everyone on the planet? We also believe in teaching really amazing an…
autonomous, car, classification, detection, driving, recognition, robot, segmentation, street, synthetic, time, urban, videoThe CVC Partial Occlusion Virtual Pedestrian datasets (CVC-01 to CVC-06) cover a range of scenarios of occluded pedestrians generated in a virtual and rea…
classification, detection, occlusion, pedestrian, synthetic, tracking, urbanThe German Traffic Sign Recognition Benchmark is a dataset for multi-class detection problem in natural images and do cordially invite you to participate.…
detection, recognition, traffic, traffic sign, urbanPN Learning - How does TLD work? Tracking estimates the object location as long as the object is visible. During tracking all observed patterns of the o…
bike, face, learning, object tracking, pedestrian, singletargetThe Daimler Mono Pedestrian Detection Benchmark dataset contains a large training and test set. The training set contains 15.560 pedestrian samples (image…
detection, mono, object, outdoor, pedestrian, scale, urbanJPL First-Person Interaction dataset (JPL-Interaction dataset) is composed of human activity videos taken from a first-person viewpoint. The dataset parti…
action, human, interactive, motion, recognition, videoThe Annotated Facial Landmarks in the Wild (AFLW) consists of a large-scale collection of annotated face images gathered from the web, exhibiting a large …
age, annotation, detection, face, landmark, poseMICCAI 2015 Challenge on Liver Ultrasound Tracking Munich, October 9, 2015 (Full Day) Outline Ultrasound (US) imaging is a widely used medical imaging…
benchmark, human, liver, medical, organ, real, therapy, tracking, ultrasoundIt is composed of ADL (activity daily living) and fall actions simulated by 11 volunteers. The people involved in the test are aged between 22 and 39, wit…
accelerometer, action, depth, fall detection - adl, human, kinect, recognition, video, wearableThe Microsoft COCO (mscoco) is an image recognition and segmentation dataset which contains more 300k images for more than 70 categories. Other features…
benchmark, context, detection, object, recognition, segmentation, semanticThe TU Berlin Multi-Object and Multi-Camera Tracking Dataset (MOCAT) is a synthetic dataset to train and test tracking and detection systems in a virtual …
animal, detection, evaluation, multi-class, multi-view, pedestrian, synthetic, tracking, vehicleMany different labeled video datasets have been collected over the past few years, but it is hard to compare them at a glance. So we have created a handy …
action, benchmark, classification, detection, object, recognition, videoBackground Models Challenge (BMC) is a complete dataset and competition for the comparison of background subtraction algorithms. The main topics concern:…
background, change, detection, modeling, motion, segmentation, surveillance, videoThe Oxford RobotCar Dataset contains over 100 repetitions of a consistent route through Oxford, UK, captured over a period of over a year. The dataset cap…
autonomous, car, classification, detection, driving, recognition, robot, segmentation, street, time, urban, video, yearThe MOT Challenge is a framework for the fair evaluation of multiple people tracking algorithms. In this framework we provide: - A large collection of d…
3d, benchmark, benhttp://motchallenge.net/chmark, dataset, evaluation, multiple, pedestrian, people, surveillance, target, tracking, videoAWS hosts a variety of public datasets that anyone can access for free. Previously, large datasets such as satellite imagery or genomic data have require…
amazon, biology, classification, deep, human, image, learning, recognition, resolution, satellite, segmentation, spaceA large dataset of geotagged face images collected from Flickr. The zip file contains text files containing urls of the images. Face2GPS: Estimating Geo…
age, classification, face, gender, geotagged, human, localizationThe Berkeley Multimodal Human Action Database (MHAD) contains 11 actions performed by 7 male and 5 female subjects in the range 23-30 years of age except …
action, classification, motion, multiview, recognitionWIDER FACE dataset is a large-scale face detection benchmark dataset with 32,203 images and 393,703 face annotations, which have high degree of variabilit…
detection, face, occlusion, pose, scaleThe INRIA People dataset from Navneet Dalal and Bill Triggs [DalalCVPR2005] consists of training and testing data. The training contains 1805 images and X…
boundingbox, frontview, human, object detection, pedestrian, sideviewThe TVPR dataset includes 23 registration sessions. Each of the 23 folders contains the video of one registration session. Acquisitions have been performe…
clothing, depth, gender, identification, indoor, people, person, recognition, reidentification, top-view, videoThe domain-specific personal videos highlight dataset from the paper [1] describes a fully automatic method to train domain-specific highlight ranker for…
action, domain, human, recognition, saliency, summarization, video, wearableISPRS Test Project on Urban Classification, 3D Building Reconstruction and Semantic Labeling. In this part of our working group site you will get further …
3d, aerial, benchmark, canada, city, germany, multiview, photogrammetry, recognition, segmentation, semantic, urbanThe TUD Crossing dataset from Micha Andriluka, Stefan Roth and Bernt Schiele consists of 201 images with 1008 highly overlapping pedestrians with signific…
detection, multitarget, overlap, pedestrian, segmentation, sideview, tracking, urbanThe QMUL Junction dataset is a busy traffic scenario for research on activity analysis and behavior understanding. Video length: 1 hour (90000 frames)…
behavior, counting, crowd, detection, motion, pedestrian, tracking, videoThis UIUC Cars dataset by Shivani Agarwal, Aatif Awan and Dan Roth contains images of side views of cars for use in evaluating object detection algorithms…
car, detection, recognition, scale, sideview, urbanTh EPFL Multi-View Car dataset contains 20 sequences of cars as they rotate by 360 degrees. There is one image approximately every 3-4 degrees. Using the …
car, detection, estimation, multiview, pose, rotationThe 3D Mask Attack Database (3DMAD) is a biometric (face) spoofing database. It currently contains 76500 frames of 17 persons, recorded using Kinect for b…
3d, biometry, emotion, face, frontview, recognition, segmentationThe Microsoft Research Cambridge-12 Kinect gesture dataset consists of sequences of human movements, represented as body-part locations, and the associate…
action, gesture, human, kinect, recognitionBrief Description of the Dataset: --------------------------------- Each of the 19 activities is performed by eight subjects (4 female, 4 male, between th…
classification, clustering, multivariate, time-seriesThe data is in the transactional form. It contains the Latin names (species or genus) and state abbreviations.
clustering, multivariateThe examined group comprised kernels belonging to three different varieties of wheat: Kama, Rosa and Canadian, 70 elements each, randomly selected for the…
classification, clustering, multivariateThe UBO 2014 consists of 7 semantic categories. Each of these 7 material categories contains measurements of 12 different material instances for being cap…
classification, illumination, light, material, recognition, texture24 scenarios recorded with 8 IP video cameras. The first 22 first scenarios contain a fall and confounding events, the last 2 ones contain only confoundin…
multiviewAbstract Scene understanding has (again) become a focus of computer vision research, leveraging advances in detection, context modeling, and tracking. In…
3d, car, classification, pedestrian, scene, segmentation, semantic, understanding3 datasets: PTZ Tracking, Thermal-visible registration, Single object tracking
pedestrian, ptz, thermal, trackingThe contour patches dataset is a large dataset of images patch matches used for contour detection. References: C. L. Zitnick and D. Parikh The Role o…
contour, detection, edge, image, lowlevel, match, patch, segmentationA Vicon motion capture camera system was used to record 12 users performing 5 hand postures with markers attached to a left-handed glove. A rigid patter…
classification, clustering, multivariateISPRS Test Project on Urban Classification and 3D Building Reconstruction The ISPRS working group III/4 announces the release of the 2D semantic labelin…
3d, building, city, classification, recognition, reconstruction, semantic, urbanKEGG Metabolic pathways can be realized into network. Two kinds of network / graph can be formed. These include Reaction Network and Relation Network. In …
classification, clustering, multivariate, regression, text, univariateScene Parsing Benchmark Scene parsing data and part segmentation data derived from ADE20K dataset could be download from MIT Scene Parsing Benchmark. m…
annotation, benchmark, recognition, scene, segmentation, semanticTreadmill gait datasets composed of 34 subjects with 9 speed variations, 68 subjects with 68 subjects, and 185 subjects with various degrees of gait fluct…
recognitionBackground information: The data set concerns the earliest history of mankind. Prehistoric men created the desired shape of a stone tool by striking on a …
causal-discovery, classification, clustering, multivariateWe developed imaging hardware consisting of a color camera, a thermal camera and a beam splitter to capture the aligned multispectral (RGB color + Thermal…
pedestrian, rgb, thermalPedestrian Color Naming (PCN) dataset contains 14,213 images, each of which hand-labeled with color label for each pixel.
color naming, pedestrian, segmentationMultiple people social interaction dataset captured by 500+ synchronized video cameras, with 3D full body skeletons and calibration data.
multiviewInstance recognition from depth data. Contains various challenges of Pose, Clutter, Occlusion and similar looking objects (Bonde, U., Badrinarayanan, V., …
depth, detection, instance, poseThe dataset is the subset of RCV1. These corpus has already been used in author identification experiments. In the top 50 authors (with respect to total s…
classification, clustering, domain-theory, multivariate, textThe PSU HUB dataset is a detection, tracking dataset. Ground truth trajectory and grouping information for pedestrians walking in the PSU student union bu…
crowd, object detection, object tracking, occlusion, overlap, pedestrian, trajectoryScene Background Initialization (SBI) dataset The SBI dataset has been assembled in order to evaluate and compare the results of background initializati…
background, benchmark, change, detection, foreground, initializationThe dataset contains 15 documentary films that are downloaded from YouTube, whose durations vary from 9 minutes to as long as 50 minutes, and the total nu…
detection, object, videoThe dataset collects data from an Android smartphone positioned in the chest pocket. Accelerometer Data are collected from 22 participants walking in the …
classification, clustering, sequential, time-series, univariateContains 6 object categories similar to object categories in Pascal VOC that are suitable for studying the abnormalities stemming from objects.
detectionScanNet is an RGB-D video dataset containing 2.5 million views in more than 1500 scans, annotated with 3D camera poses, surface reconstructions, and insta…
3d, cad, indoor, layout, object, realism, recognition, rendering, room, scene, segmentation, synthetic30000+ frames with vehicle rear annotation and classification (car and trucks) on motorway/highway sequences. Annotation semi-automatically generated usin…
detectionThis is a transnational data set which contains all the transactions occurring between 01/12/2010 and 09/12/2011 for a UK-based and registered non-store o…
classification, clustering, multivariate, sequential, time-seriesLarge population gait datasets composed of 4,016 subjects.
recognitionThis archive contains 2075259 measurements gathered between December 2006 and November 2010 (47 months). Notes: 1.(global_active_power*1000/60 - sub_mete…
clustering, multivariate, regression, time-seriesPlease see the README for the details on the data organization, and so on.
classification, clustering, multivariate, textClassification/Detection Competitions, Segmentation Competition, Person Layout Taster Competition datasets
classification, detectionThe Pedestrian Parsing dataset contains 3,673 images from 171 videos of different Surveillance Scenes (PPSS), where 2,064 images are occluded and 1,609 ar…
parsing, pedestrian, segmentationThe Robotic 3D Scan Repository from Osnabrueck contains 23 different datasets showing a veriaty of 3D scans for objects, humans, cities, university campus…
3d, aerial, bremen, city, germany, heat, human, laser, lidar, osnabrueck, reconstruction, scan, urbanFor the first few decades of the fields existence, computer vision has been focused on algorithmic, logical approaches to perception. But it was only with…
3d, depth, indoor, kinect, object, recognition, reconstructionThe mirror symmetry database contains 176 single-symmetry and 63 multyple-symmetry images (.png files) with accompanying ground-truth annotations (.mat fi…
detection, groundtruth, mirror, symmetry15 wide baseline stereo image pairs with large viewpoint change, provided ground truth homographies. Image size (~1000x700 pixels, RGB) D. Mishkin and…
description, detection, feature, matching, viewpoint, wide baseline stereoThe xawAR16 dataset is a multi-RGBD camera dataset, generated inside an operating room (IHU Strasbourg), which was designed to evaluate tracking/relocaliz…
depth, medicine, operation, recognition, surgery, table, videoThis dataset comes from the daily measures of sensors in a urban waste water treatment plant. The objective is to classify the operational state of the pl…
clustering, multivariateThe PubFig database is a large, real-world face dataset consisting of 58,797 images of 200 people collected from the internet. Unlike most other existing …
recognitionPlaces205 dataase contains 2.5 million images from 205 scene categories for the academic public. The image dataset contains 2,448,873 images from 205 sc…
feature, learning, place, recognition, scene, urbanThe experiments have been carried out with a group of 115 students of first-year, undergraduate Engineering major of the University of Genoa. We carried…
classification, clustering, multivariate, regression, sequential, time-seriesThe CERTH image blur dataset consists of 2450 digital images, 1850 out of which are photographs captured by various camera models in different shooting co…
blur, defocus, detection, image, motion, qualityCSV format where each row is a paper and each column is an attribute.
clustering, multivariateA Vicon motion capture camera system was used to record 12 users performing 5 hand postures with markers attached to a left-handed glove. A rigid patter…
classification, clustering, multivariateThe Aspect Layout dataset is designed to allow evaluation of object detection for aspect ratios in perspective images. Author text: In this project we…
aspect, detection, layout, object, perspective, ratioAn indoor action recognition dataset which consists of 18 classes performed by 20 individuals. Each action is individually performed for 8 times (4 daytim…
action, cross-view, indoor, multi-camera, open-view, recognition, videoFine-Grained Visual Classification of Aircraft (FGVC-Aircraft) is a benchmark dataset for the fine grained visual categorization of aircraft. Data, anno…
aircraft, airplane, benchmark, classification, evaluation, fine-grained, recognitionThe FAce Semantic SEGmentation (FASSEG) repository contains datasets for multi-class semantic face segmentation. The FASSEG repository is composed by tw…
face, segmentationThe Daimler Mono Pedestrian Classification Benchmark dataset consists of two parts: a base data set. The base data set contains a total of 4000 pedestri…
classification, illumination, object, outdoor, pedestrian, scale, urbanThis ETHZ CVL RueMonge 2014 dataset used for 3D reconstruction and semantic mesh labelling for urban scene understanding. It was first published in [1] …
3d, architecture, benchmark, classification, code, mesh, outdoor, paris, pointcloud, recognition, reconstruction, segmentation, semantic, source, urbanData set has no missing values. Values are in kW of each 15 min. To convert values in kWh values must be divided by 4. Each column represent one client. S…
clustering, regression, time-seriesThe FaceScrub dataset is a real-world face dataset comprising 107,818 face images of 530 male and female celebrities detected in images retrieved from the…
recognitionThe Microsoft Research Cambridge-12 Kinect gesture data set consists of sequences of human movements, represented as body-part locations, and the associat…
recognitionStyle, Price, Rating, Size, Season, NeckLine, SleeveLength, waiseline, Material, FabricType, Decoration, Pattern, Type, Recommendation are Attributes in d…
classification, clustering, textThe Visual Attributes dataset contains visual attribute annotations for over 500 object classes (animate and inanimate) which are all represented in Image…
attribute, classification, imagenet, object, recognitionThese are atlantic-mediterranean marine sponges that belong to O.Hadromerida (Demospongiae.Porifera).
clustering, multivariateThe dataset consist of the about 50 hours obtained from kindergarten surveillance videos. Dataset, totally approximately 100 videos sequences (1000GB, 50 …
action, background, behavior, human, segmentation, video surveillanceKEGG Metabolic pathways can be realized into network. Two kinds of network / graph can be formed. These include Reaction Network and Relation Network. In …
classification, clustering, multivariate, regression, text, univariateThe Video Segmentation Benchmark (VSB100) provides ground truth annotations for the Berkeley Video Dataset, which consists of 100 HD quality videos divide…
benchmark, groundtruth, motion, object, pedestrian, segmentation, tracking, videoThe CMP map2photo dataset consists of 6 pairs, where one image is satellite photo and second image is a map of the same area. The task is to match these …
baseline, description, detection, feature, map, matching, remote, sensing, wideThe PETS 2006 dataset contains 7 parts showing multi-sensor sequences containing left-luggage scenarios with increasing scene complexity at a train statio…
frontview, indoor, multitarget, object detection, object tracking, pedestrianToday, we introduce Open Images, a dataset consisting of ~9 million URLs to images that have been annotated with labels spanning over 6000 categories. We …
annotation, automatic, category, classification, deep, image, large-scale, realOver 110,000 photographic reproductions of the artworks exhibited in the Rijksmuseum (Amsterdam, the Netherlands). Offers four automatic visual recognitio…
recognitionThe FlickrLogos-32 dataset contains photos showing brand logos and is meant for the evaluation of multi-class logo recognition as well as logo retrieval m…
classification brand boundingbox, detection, flickr, image, logo, machine learning, object recognition, retrievalThe dataset represents 10 years (1999-2008) of clinical care at 130 US hospitals and integrated delivery networks. It includes over 50 features representi…
classification, clustering, multivariateChokePoint is a video dataset designed for experiments in person identification/verification under real-world surveillance conditions. The dataset consist…
recognitionThis dataset contains 600 examples of control charts synthetically generated by the process in Alcock and Manolopoulos (1999). There are six different cla…
classification, clustering, time-seriesThe Kendall Square webcam dataset consists of two streams for one sunny day and one cloudy day of a city square. It is used for tracking and analyzing col…
appearance, change, color, detection, sky, weather, webcamThe Dataset for ADL Recognition with Wrist-worn Accelerometer is a public collection of labelled accelerometer data recordings to be used for the creation…
classification, clustering, multivariate, time-seriesThe TUD Pedestrians training dataset from Micha Andriluka, Stefan Roth and Bernt Schiele consists of 210 and 400 training images with X pedestrians with s…
object detection, pedestrian, segmentation, sideviewFERG-DB is a database of stylized characters with annotated facial expressions. The database contains multiple face images of six stylized characters. The…
anger, animation, annotation emotion, cardinal classification, deep learning, disgust, face, facial expression, facial expressions, fear, human transfer, image retrieval, joy, neutral, sad, stylization, surpriseThe Caltech Pedestrian Dataset consists of approximately 10 hours of 640x480 30Hz video taken from a vehicle driving through regular traffic in an urban e…
object detection, pedestrian, urbanThe PETS 2016 IPATCH dataset contains a set of fourteen multi camera recordings (visible, themal) collected off the coast of Brest, France, in collaborati…
boat, detection, gps, maritime, multimodal, radar, thermal, tracking, vessel, visibleThe data set contains 3,425 videos of 1,595 different people. The shortest clip duration is 48 frames, the longest clip is 6,070 frames, and the average l…
recognitionIndoor localization is a key topic for mobile computing. However, it is still very difficult for the mobile sensing community to compare state-of-art Indo…
classification, clustering, multivariate, regression, sequential, time-seriesThe PD and control handwriting database consists of 62 PWP (People with parkinson) and 15 healthy individuals who appealed at the Department of Neurology …
classification, clustering, multivariate, regressionIndoor localisation is a key topic for the Ambient Intelligence (AmI) research community. In this scenarios, recent advancements in wearable technologie…
classification, clustering, multivariate, regression, sequential, time-seriesThe city planar and non-planar datset consists of urban scenes accompanied by text files describing the plane/non-plane locations. Training Set (Univer…
3d, building, detection, estimation, plane, urbanIn predicting stock prices you collect data over some period of time - day, week, month, etc. But you cannot take advantage of data from a time period unt…
classification, clustering, time-seriesAn Annotated Dataset For Near-Duplicate Detection In Personal Photo Collections Managing photo collections involves a variety of image quality assessmen…
copyright, detection, duplicate, groundtruth, retrievalThe PASCAL VOC Challenge datasets by Mark Everingham is a yearly dataset which has a central evaluation server and the final test data is not released. Th…
airplane, animal, building, car, chair, object detection, object pose, object segmentation, pedestrianThe MIT + CMU frontal face dataset from H. Rowley contains 130 images with 507 labeled frontal faces from movie, portrait and media sources. It is mostly …
detection object boundingbox, face, face detection, frontviewThe Babenko tracking dataset contains 12 video sequences for single object tracking. For each clip they provide (1) a directory with the original ima…
animal, face, object tracking, occlusion, single, videoA new large-scale PEdesTrian Attribute (PETA) dataset. The dataset is by far the largest of its kind, covering more than 60 attributes on 19000 images. I…
crowd, pedestrianThe dataset (movement_libras) contains 15 classes of 24 instances each, where each class references to a hand movement type in LIBRAS. In the video pre-p…
classification, clustering, multivariate, sequentialThis dataset is for people tracking in wide baseline camera networks and was designed as a contest at ICPR 2012. The contest consists of two challenges…
aerial, crowd, object detection, object tracking, occlusion, overlap, pedestrian, trajectory--- The dataset collects data from a wearable accelerometer mounted on the chest --- Sampling frequency of the accelerometer: 52 Hz --- Accelerom…
classification, clustering, sequential, time-series, univariateThe Western GCO Segmentation problem instances are provided to compare effects of graph size, neighborhood size, length of s to t paths, regional arc cons…
abdomen, adhead, babyface, binary, bone, face, liver, medical, optimization, segmentationThe Wide (multiple) Baseline Dataset. 31 image pairs, simultaneously combining several nuisance factors: geometry, illumination, IR-visible, etc. WxBS: …
day, description, detection, feature, ir, matching, night, viewpointIdiap/ETHZ Faces and Poses Dataset dataset by L. Jie, B. Caputo and V. Ferrari contains 1703 image-caption pairs. [author] Captions contain the names of s…
face detection, object pose, pedestrian, textMIT Pedestrian dataset from Papageorgiou and Poggio [IJCV2000] contains 509 training and 200 test images of pedestrians in city scenes (plus left-right re…
boundingbox, frontview, object detection, pedestrian, people, urban15 wide baseline stereo image pairs with large viewpoint change, provided ground truth homographies.
multiviewNews are grouped into clusters that represent pages discussing the same news story. The dataset includes also references to web pages that, at the access…
classification, clustering, multivariateThe data set gathered when we were working at project for Bahrain university between 2002 and 2003.
classification, clustering, domain-theory, univariateAutomatic identification of commercial blocks in news videos finds a lot of applications in the domain of television broadcast analysis and monitoring. Co…
classification, clustering, multivariateThe Fish4Knowledge project (groups.inf.ed.ac.uk/f4k/) is pleased to announce the availability of 2 subsets of our tropical coral reef fish video and ext…
animal, camera, classification, fish, motion, nature, recognition, video, waterThe dataset consists of a total of 3600 documents including 600 news/texts from six categories economy, culture-arts, health, politics, sports and techno…
classification, clustering, textCollect the real time readings for residential,commercial,industrial,agriculure,to find the accuracy consumption in Tamil Nadu Around Thanajvur
classification, clustering, multivariate, regressionThis work attempts to provide two Hand Images Databases for hand biometrics: one is created using a mobile phone camera of modest quality, which we call…
authentication, biometric/ hand geometry, identification, mobile, person, segmentation, shape, webcamThe TUD Campus dataset from Micha Andriluka, Stefan Roth and Bernt Schiele consists of 71 images and 303 highly overlapping pedestrians with large scale c…
object detection, object tracking, overlap, pedestrian, segmentation, sideviewFor complete information see the official challenge page: [Web Link]
causal-discovery, clustering, domain-theory, multivariate, sequential, time-seriesThe dataset is in the form of a 11463 x 5812 matrix of word counts, containing 11463 words and 5811 NIPS conference papers (the first column contains the …
clustering, textThe Yale Face dataset from A. Georghiades contains 5760 single light source images of ten subjects, each shown in 9 poses and 64 illumination setups (lead…
face detection, illumination, pedestrian, pose estimationCalifornia-ND contains 701 photos taken directly from a real user's personal photo collection, including many challenging non-identical near-duplicate cas…
detectionThe Leeds Cows dataset by Derek Magee consists of 14 different video sequences showing a total of 18 cows walking from right to left in front of different…
animal, background, cow, detection, segmentation, videoWe would like to announce the release of PASCAL-Context dataset. We augmented PASCAL VOC 2010 dataset with annotations for 400+ additional categories. In …
benchmark, category, dense, pascal, recognition, segmentation, semantic, shapeThe SUNCG dataset is a Large 3D Model Repository for Indoor Scenes. SUNCG is an ongoing effort to establish a richly-annotated, large-scale dataset of…
3d, indoor, layout, object, realism, recognition, rendering, room, scene, segmentation, syntheticThe DrivFace database contains images sequences of subjects while driving in real scenarios. It is composed of 606 samples of 640480 pixels each, acquired…
classification, clustering, multivariate, regressionPlease find the original data at '[Web Link]'
classification, clustering, multivariate, time-seriesThis dataset consists of 8000+ images of professional footballers during a match of the Allsvenskan league. It consists of two parts: one with ground trut…
multiviewThis dataset contains 250 pedestrian image pairs + 775 additional images captured in a busy underground station for the research on person re-identificati…
recognitionA database of face photographs designed for studying the problem of unconstrained face recognition
recognitionISPRS / EuroSDR Benchmark for Multi-Platform Photogrammetry In these pages you can get information about the BENCHMARK FOR MULTI-PLATFORM PHOTOGRAMMETRY…
3d, aerial, benchmark, city, germany, multiview, photogrammetry, reconstruction, switzerland, urbanGroup emotion recognition in images - Happiness Intensity labels for group of people in images. The images have been collected from Flickr using keyword s…
behavior, emotion, facial expression, flickr, group, human, wildWe wanted to have a collection of action recognition papers and results that everybody can use for reference. The site will work by the community principl…
action, benchmark, dataset, recognitionThe experiments have been carried out with a group of 30 volunteers within an age bracket of 19-48 years. Each person performed six activities (WALKING, W…
classification, clustering, multivariate, time-seriesDocuments are first obtained via a Web search using AMIEI: an integrated platform for delivering enterprise intelligence, developed by AMI Software ([Web …
clustering, multivariate, sequential, textThe Traffic Video dataset consists of X video of an overhead camera showing a street crossing with multiple traffic scenarios. The dataset can be downlo…
detection, overhead, road, tracking, traffic, urban, video, viewOpen University Learning Analytics Dataset (OULAD) contains data about courses, students and their interactions with Virtual Learning Environment (VLE) fo…
classification, clustering, multivariate, regression, sequential, time-series10000 images of natural scenes grabbed on Flickr, with 2695 logos instances cut and pasted from the BelgaLogos dataset.
detectionThe All I Have Seen (AIHS) dataset is created to study the properties of total visual input in humans, for around two weeks Nebojsa Jojic wore a camera ca…
3d, clustering, indoor, outdoor, scene, similarity, study, summary, user, videoThe Caltech Lanes dataset includes four clips taken around streets in Pasadena, CA at different times of day. The archive below includes 1225 individual…
caltech, detection, lane, pasadena, road, urbanThe Comprehensive Cars (CompCars) dataset contains data from two scenarios, including images from web-nature and surveillance-nature. The web-nature data …
attribute, car, classification, fine-grained, object, recognition, urban, vehicleThe TUD Pedestrians dataset from Micha Andriluka, Stefan Roth and Bernt Schiele [AndrilukaCVPR2008] consists of 250 images with 311 fully visible people w…
object detection, pedestrian, segmentation, sideviewThis dataset includes: camera calibration information, raw input images we have captured, radially undistorted, rectified, and cropped images, depth maps …
multiviewThe HandNet dataset contains depth images of 10 participants hands non-rigidly deforming infront of a RealSense RGB-D camera. This dataset includes 214…
articulation, classification, detection, fingertip, hand, pose, rgbd, segmentation, videoWe provide the three datasets used for testing our system for our ICCV 2007 publication, including annotations. Data was recorded using a pair of AVT Marl…
crowd, object detection, object tracking, occlusion, overlap, pedestrian, trajectoryObjects viewed from 144 calibrated viewpoints under 3 different lighting conditions
multiviewThis dataset was constructed by adding elevation information to a 2D road network in North Jutland, Denmark (covering a region of 185 x 135 km^2). Elevati…
clustering, regression, sequential, textHTRU2 is a data set which describes a sample of pulsar candidates collected during the High Time Resolution Universe Survey (South) [1]. Pulsars are a r…
classification, clustering, multivariate- The leaves were placed on a white background and then photographed. - The pictures were taken in broad daylight to ensure optimum light intensity.
classification, clustering, multivariateThe Semantic Description of Human Activities (SDHA) was a contest at ICPR 2010. The contest is composed of three different types of activity recognitio…
aerial, crowd, object detection, object tracking, occlusion, overlap, pedestrian, trajectoryCSV format where each row is a paper and each column an attribute.
clustering, multivariateFace tracks, features and shot boundaries from our latest CVPR 2013 paper. It is obtained from 6 episodes of Buffy the Vampire Slayer and 6 episodes of Bi…
recognitionFor each text collection, D is the number of documents, W is the number of words in the vocabulary, and N is the total number of words in the collection (…
clustering, textThe LabelMeFacade dataset contains buildings, windows, sky and a limited number of unlabeled regions (maximally 20% covering of the image). This procedure…
facade, recognition, rectified, segmentation, semantic, urbanThe Cholec80 dataset contains 80 videos of cholecystectomy surgeries performed by 13 surgeons. The videos are captured at 25 fps. The dataset is labeled w…
medicine, phase, recognition, surgery, tool, videot is composed of food intake movements, recorded with Kinect V1 (320240 depth frame resolution), simulated by 35 volunteers for a total of 48 tests. The d…
age, behavior, food, groundtruth, human, intake, kinect, monitoring, pointcloud, trackingCollected in a clothing store. Captured with Kinect (640*480, about 30fps)
detection, trackingProvide all relevant information about your data set.
classification, clustering, multivariateThe measurements were created to ease the development, comparison and evaluation of fingerprinting based hybrid indoor positioning methods. The measuremen…
causal-discovery, classification, clustering, textThis dataset contains 7 challenging volleyball activity classes annotated in 6 videos from professionals in the Austrian Volley League (season 2011/12). A…
action, activity recognition, analysis, detection, sport, video, volleyballThe object is a plaster reproduction of Temple of the Dioskouroi in Agrigento, Sicily. Click on thumbnail for a full-sized (640x480) image. Resolution of …
3d, 3d reconstruction, benchmark, multiview, sfmWe introduce the Shelf dataset for multiple human pose estimation from multiple views. In addition we annotate the body joints in the Campus dataset from …
3d, capture, estimation, human, motion, multiple, pose, viewPhos is a color image database of 15 scenes captured under different illumination conditions. More particularly, every scene of the database contains 15 …
detectionShakeFive2 A collection of 8 dyadic human interactions with accompanying skeleton metadata. The metadata is frame based xml data containing the skeleton…
human, interaction, kinect, videoThe data set consists of the expression levels of 77 proteins/protein modifications that produced detectable signals in the nuclear fraction of cortex. Th…
classification, clustering, multivariateThis corpus has been collected from free or free for research sources at the Internet: -> A collection of 425 SMS spam messages was manually extracted fr…
classification, clustering, domain-theory, multivariate, textThe data was collected as part of the 1990 census. There are 68 categorical attributes. This data set was derived from the USCensus1990raw data set. The…
clustering, multivariateThe CALTECH 256 dataset by Li Fei-Fei contains 30607 images for 256 categories.
centered, classification, detection, image, object, scene* Audio track (encoded as mp3) of each of the 106,574 tracks. It is on average 10 millions samples per track.* Nine audio features (consisting of 518 attr…
classification, clustering, multivariate, time-seriesThe dataset is composed by features extracted from 7 videos with people gesticulating, aiming at studying Gesture Phase Segmentation. Each video is repres…
classification, clustering, multivariate, sequential, time-seriesThe Paris Art Deco Facades dataset consists of 79 / 80 images of rectified facades of the architectural style Art Deco, which has different sizes of windo…
architecture, city, facade, grammar, paris, procedural, recognition, segmentation, semantic, urbanThe Stanford Dogs dataset contains images of 120 breeds of dogs from around the world. This dataset has been built using images and annotation from ImageN…
classification, detection, dogs, fine-grained categorizationThe characters here were used for a PhD study on primitive extraction using HMM based models. The data consists of 2858 character samples, contained in th…
classification, clustering, time-seriesThe Heterogeneity Dataset for Human Activity Recognition from Smartphone and Smartwatch sensors consists of two datasets devised to investigate sensor het…
classification, clustering, multivariate, time-series24 synthetic scenes. Available data per scene: 9x9 input images (512x512x3) , ground truth (disparity and depth), camera parameters, disparity ranges, eva…
multiviewDaimler Multi-Cue, Occluded Pedestrian Classification Benchmark Training and test samples have a resolution of 48 x 96 pixels with a 12-pixel border aro…
image classification, object detection, pedestrian, urbanThe Video Summarization (SumMe) dataset consists of 25 videos, each annotated with at least 15 human summaries (390 in total). The data consists of videos…
action, benchmark, event, groundtruth, human, summary, videoThe Graz02 dataset by Andreas Opelt and Axel Pinz contains four categories of images: bikes, people, cars and a single background class. The annotation ha…
background, bike, car, clutter, graz, object detection, pedestrianLarge-scale PEdesTrian Attribute (PETA) dataset, covering more than 60 attributes (e.g. gender, age range, hair style, casual/formal) on 19000 images.
recognitionThe object is a plaster dinosaur (stegosaurus). Click on thumbnail for a full-sized (640x480) image. Resolution of ground truth model: 0.00025m (you may w…
3d, 3d reconstruction, benchmark, multiview, sfmThe Freiburg-Berkeley Motion Segmentation Dataset (FBMS-59) is an extension of the BMS dataset with 33 additional video sequences. A total of 720 frames i…
benchmark, groundtruth, motion, object, pedestrian, segmentation, tracking, videoThe Landmark 1000 or 1k dataset is a collection of the top 1000 popular flickr landmarks mined from flickr. It is maintained by Noah Snavely and publish…
3d, estimation, landmark, location, pointcloud, pose, reconstruction, world1521 images with human faces, recorded under natural conditions, i.e. varying illumination and complex background. The eye positions have been set manuall…
detectionDaimler Stereo Pedestrian Detection Benchmark C. Keller, M. Enzweiler, and D. M. Gavrila, A New Benchmark for Stereo-based Pedestrian Detection, Proc. …
object detection, pedestrian, urbanThe BIWI Walking Pedestrians (EWAP) dataset shows walking pedestrians in busy scenarios from a bird eye view. Manually annotated. Data used for training i…
aerial, crowd, object detection, object tracking, occlusion, overlap, pedestrian, trajectorySamples (instances) are stored row-wise. Variables (attributes) of each sample are RNA-Seq gene expression levels measured by illumina HiSeq platform.
classification, clustering, multivariate10000 images of natural scenes, with 37 different logos, and 2695 logos instances, annotated with a bounding box.
detectionThe Salient Montages is a human-centric video summarization dataset from the paper [1]. In [1], we present a novel method to generate salient montages f…
human, montage, saliency, summarization, video, wearableLASIESTA is composed by many real indoor and outdoor sequences organized in different categories, each of one covering a specific challenge in moving obje…
background, camera, challenge, dataset, detection, foreground, groundtruth, motion, object, stationary, subtractionThis is a sparse data set, less than 10% of the attributes are used for each sample. The link is to a '*.tgz' file which contains two files: [amzn-anon-ac…
causal-discovery, clustering, domain-theory, regression, time-seriesThese datasets were generated for the M2CAI challenges, a satellite event of MICCAI 2016 in Athens. Two datasets are available for two different challenge…
challenge, medicine, recognition, surgery, video, workflowThe Airport MotionSeg dataset contains 12 sequences of videos of an aiprort scenario with small and large moving objects and various speeds. It is challen…
airport, camera, clustering, motion, segmentation, video, zoomDinosaur, Model House, Corridor, Aerial views, Valbonne Church, Raglan Castle, Kapel sequence
multiviewLabelMe is a web-based image annotation tool that allows researchers to label images and share the annotations with the rest of the community. If you use …
detectionThe database contains historical car images from 1920s to 1990s crawled from cardatabase.net. There are 10130 training and 3343 test images. Annotations…
car, recognition, timeThe dataset is composed of 150 synthetic scenes, captured with a (perspective) virtual camera, and each scene contains 3 to 5 objects. The model set is co…
mesh, recognition, segmentation, syntheticThis dataset comprises information regarding the ADLs performed by two users on a daily basis in their own homes. This dataset is composed by two instanc…
classification, clustering, multivariate, sequential, time-seriesThe test sequences provide interested researchers a real-world multi-view test data set captured in the blue-c portals. The data is meant to be used for t…
action, camera, multiview, segmentation, trackingThis dataset package contains the software and data used for Detection-based Object Labeling on the RGB-D Scenes Dataset as implemented in the paper: De…
3d, depth, indoor, kinect, object, recognition, reconstructionThe TRaffic ANd COngestionS (TRANCOS) dataset, a novel benchmark for (extremely overlapping) vehicle counting in traffic congestion situations. It consist…
car, detection, highway, object, spain, traffic, transportation, urban, vehicleThe CMP Facade dataset consists of facade images assembled at the Center for Machine Perception, which includes 600 rectified images of facades from vario…
classification, facade, recognition, rectification, segmentation, semantic, similarity, structure, urban52 columns for 52 weeks; normalised values of provided too.
clustering, multivariate, time-seriesThe Extreme Zoom Dataset. EZD is a 6 image sets with incleasing zoom factor from general scene view to focusing on single detail. MODS: Fast and Robust …
description, detection, feature, matching, viewpoint, zoomThe Multicamera Human Action Video Data (MuHAVi) Manually Annotated Silhouette Data (MAS) are two datasets consisting of selected action sequences for th…
action, background, behavior, human, segmentation, video15,560 pedestrian and non-pedestrian samples (image cut-outs) and 6744 additional full images not containing pedestrians for bootstrapping. The test set c…
detectionOngoing research on university faculty perceptions and practices of using Wikipedia as a teaching resource. Based on a Technology Acceptance Model, the re…
causal-discovery, clustering, multivariate, regressionThis dataset was used in several classifications tasks related to the challenge of anuran species recognition through their calls. It is a multilabel data…
classification, clustering, multivariateContains drawing pages from US patents with manually labeled figure and part labels.
detectionThe Graz01 dataset by Andreas Opelt and Axel Pinz contains four types of images: bikes, people, background with no bikes, background with no people.
background, bike, clutter, graz, object detection, occlusion, pedestrianImages from 19 sites collected from a helicopter flying around Providence, RI. USA. The imagery contains approximately a full circle around each site.
multiviewThe UMD Dynamic Scene Recognition dataset consists of 13 classes and 10 videos per class and is used to classify dynamic scenes. The dataset has been de…
classification, dynamic, motion, recognition, scene, videoThis data set contains 13,910 measurements from 16 chemical sensors exposed to 6 gases at different concentration levels. This dataset is an extension of …
causa, classification, clustering, multivariate, regression, time-seriesGlobal Symmetry Ground-truth for AVA dataset Release Date: 2016 For detailed information, please refer to: Elawady, Mohamed, Ccile Barat, Christophe …
aesthetic, bilateral, detection, global, mirror, reflection, symmetryThis dataset provides a collection of web images and 3D models for research on landmark recognition (especially for methods based on 3D models). We hope i…
3d, classification, codebook, feature, flickr, landmark, matching, recognition, reconstruction, retrievalThe Buffy dataset contains images selected from the TV series, Buffy: the Vampire Slayer. We select a set of 452 images from the first two episodes for tr…
buffy, human, movie, object detection, segmentationThe We Are Family Stickmen dataset from Eichner and Ferrari contains X images with X people in group photos for human pose estimation with annotated 2D hu…
body part, object pose, pedestrian-- The users' knowledge class were classified by the authors using intuitive knowledge classifier (a hybrid ML technique of k-NN and meta-heuristic…
classification, clustering, multivariateThe SPHERE human skeleton movements dataset was created using a Kinect camera, that measures distances and provides a depth map of the scene instead of th…
action, behavior, depth, human, kinect, motion, movement, skeleton, videoThe Ford Car dataset is joint effort of Pandey et al. (for collecting images, Lidar points, calibration etc.) and us (for annotation of 2D and 3D objects)…
3d, car, detection, groundtruth, lidar, sfmThe YouTube-Objects dataset is composed of videos collected from YouTube by querying for the names of 10 object classes. It contains between 9 and 24 vide…
detection, flow, object, optical, segmentation, videoISPRS and EuroSDR - Benchmark on High Density Aerial Image Matching Background and Scope of the project Innovations in matching algorithms as well as …
3d, aerial, benchmark, city, germany, multiview, photogrammetry, reconstruction, switzerland, urbanThe automated analysis of facial expressions has been widely used in different research areas, such as biometrics or emotional analysis. Special import…
classification, clustering, multivariate, sequential