The TU Berlin Multi-Object and Multi-Camera Tracking Dataset (MOCAT) is a synthetic dataset to train and test tracking and detection systems in a virtual …
animal, detection, evaluation, multi-class, multi-view, pedestrian, synthetic, tracking, vehicleThe UrbanStreet dataset used in the paper can be downloaded here [188M] . It contains 18 stereo sequences of pedestrians taken from a stereo rig mounted o…
detection, human, multitarget, pedestrian, recognition, segmentation, tracking, urban, videoAt Udacity, we believe in democratizing education. How can we provide opportunity to everyone on the planet? We also believe in teaching really amazing an…
autonomous, car, classification, detection, driving, recognition, robot, segmentation, street, synthetic, time, urban, videoThe TUD Crossing dataset from Micha Andriluka, Stefan Roth and Bernt Schiele consists of 201 images with 1008 highly overlapping pedestrians with signific…
detection, multitarget, overlap, pedestrian, segmentation, sideview, tracking, urbanThe dataset captures 25 people preparing 2 mixed salads each and contains over 4h of annotated accelerometer and RGB-D video data. Annotated activities co…
action, activity, classification, detection, recognition, tracking, videoThe QMUL Junction dataset is a busy traffic scenario for research on activity analysis and behavior understanding. Video length: 1 hour (90000 frames)…
behavior, counting, crowd, detection, motion, pedestrian, tracking, videoSVHN is a real-world image dataset for developing machine learning and object recognition algorithms with minimal requirement on data preprocessing and fo…
classification, detection, number, real, recognition, streetside, streetview, text, urban, worldWe present a new large-scale dataset that contains a diverse set of stereo video sequences recorded in street scenes from 50 different cities, with high q…
car, cities, detection, pedestrian, person, segmentation, semantic, stereo, urban, video, weaklyThe KTH Multiview Football dataset contains 771 images of football players includes images taken from 3 views at 257 time instances 14 annotated body join…
camera, detection, game, multitarget, multiview, object, outdoor, pedestrian, pose, recognition, soccer, trackingThe ICG Lab 6 (Multi-Camera Multi-Object Tracking) dataset contains 6 indoor people tracking scenarios recorded at our laboratory using 4 static Axis P134…
calibration, camera, detection, evaluation, graz, laboratory, multiview, object, pedestrian, segmentation, trackingSome datasets and evaluation tools are provided on this page for four different computer vision and computer graphics problems. Population counting Lin…
3d, counting, crowd, detection, groundtruth, line, network, object, pedestrian, pointcloud, reconstruction, road, surface, urbanThe Mall dataset was collected from a publicly accessible webcam for crowd counting and profiling research. Ground truth: Over 60,000 pedestrians were …
counting, crowd, detection, indoor, pedestrian, tracking, video, webcamThe Traffic Video dataset consists of X video of an overhead camera showing a street crossing with multiple traffic scenarios. The dataset can be downlo…
detection, overhead, road, tracking, traffic, urban, video, viewThe Daimler Mono Pedestrian Detection Benchmark dataset contains a large training and test set. The training set contains 15.560 pedestrian samples (image…
detection, mono, object, outdoor, pedestrian, scale, urbanThe PETS 2009 dataset contains 3 parts showing multi-view sequences containing pedestrians walking in an outdoor environment. The parts are used for perso…
detection, frontview, human, occlusion multitarget, outdoor, overlap, pedestrian, trackingThe Oxford RobotCar Dataset contains over 100 repetitions of a consistent route through Oxford, UK, captured over a period of over a year. The dataset cap…
autonomous, car, classification, detection, driving, recognition, robot, segmentation, street, time, urban, video, yearThe ICG Multi-Camera datasets consist of Easy Data Set (just one person) Medium Data Set (3-5 persons, used for the experiments) Hard Data Set (crowd…
calibration, camera, detection, graz, indoor, multitarget, multiview, object, pedestrian, tracking, videoThe High Definition Analytics (HDA) dataset is a multi-camera High-Resolution image sequence dataset for research on High-Definition surveillance: Pedestr…
benchmark, camera, detection, high-definition, human, indoor, lisbon, multiview, network, pedestrian, re-identification, surveillance, tracking, videoThe Daimler Mono Pedestrian Classification Benchmark dataset consists of two parts: a base data set. The base data set contains a total of 4000 pedestri…
classification, illumination, object, outdoor, pedestrian, scale, urbanThe ICG Multi-Camera and Virtual PTZ dataset contains the video streams and calibrations of several static Axis P1347 cameras and one panoramic video from…
calibration, camera, crowd, detection, graz, multitarget, multiview, network, object, outdoor, panorama, pedestrian, tracking, videoBelgiumTS is a large dataset with 10000+ traffic sign annotations, thousands of physically distinct traffic signs. 4 video sequences recorded with 8 high …
belgium, calibration, camera, classification, road, sign, traffic, urbanThe Semantic Description of Human Activities (SDHA) was a contest at ICPR 2010. The contest is composed of three different types of activity recognitio…
aerial, crowd, object detection, object tracking, occlusion, overlap, pedestrian, trajectoryThe PSU HUB dataset is a detection, tracking dataset. Ground truth trajectory and grouping information for pedestrians walking in the PSU student union bu…
crowd, object detection, object tracking, occlusion, overlap, pedestrian, trajectoryThe crowd datasets are collected from a variety of sources, such as UCF and data-driven crowd datasets. The sequences are diverse, representing dense crow…
anomaly, crowd, detection, human, pedestrian, scene, understanding, videoThe CMP Facade dataset consists of facade images assembled at the Center for Machine Perception, which includes 600 rectified images of facades from vario…
classification, facade, recognition, rectification, segmentation, semantic, similarity, structure, urbanDaimler Stereo Pedestrian Detection Benchmark C. Keller, M. Enzweiler, and D. M. Gavrila, A New Benchmark for Stereo-based Pedestrian Detection, Proc. …
object detection, pedestrian, urbanISPRS Test Project on Urban Classification and 3D Building Reconstruction The ISPRS working group III/4 announces the release of the 2D semantic labelin…
3d, building, city, classification, recognition, reconstruction, semantic, urbanThe BIWI Walking Pedestrians (EWAP) dataset shows walking pedestrians in busy scenarios from a bird eye view. Manually annotated. Data used for training i…
aerial, crowd, object detection, object tracking, occlusion, overlap, pedestrian, trajectoryThe city planar and non-planar datset consists of urban scenes accompanied by text files describing the plane/non-plane locations. Training Set (Univer…
3d, building, detection, estimation, plane, urbanThe Freiburg-Berkeley Motion Segmentation Dataset (FBMS-59) is an extension of the BMS dataset with 33 additional video sequences. A total of 720 frames i…
benchmark, groundtruth, motion, object, pedestrian, segmentation, tracking, videoThe Stanford Background Dataset is a new dataset introduced in Gould et al. (ICCV 2009) for evaluating methods for geometric and semantic scene understand…
classification, geometry, nature, segmentation, semantic, urbanThe Swedish Traffic Sign Recognition provides Matlab code for parsing the annotation files and displaying the results. Part0 for each set contains the ann…
city, detection, recognition, sign, traffic, urbanThe RGB-D Person Re-identification dataset is for person re-identification using depth information. The main motivation is that the standard techniques (s…
3d, classification, depth, identification, pedestrian, shapeThe Outex dataset is part of a framework for empirical evaluation of texture classification and segmentation algorithms. The framework is being construct…
benchmark, classification, segmentation, synthetic, textureWe collected a video dataset, termed ChokePoint, designed for experiments in person identification/verification under real-world surveillance conditions u…
clustering, detection, face, human, identification, multiview, pedestrian, real, recognition, sequence, surveillance, worldThe German Traffic Sign Recognition Benchmark is a dataset for multi-class detection problem in natural images and do cordially invite you to participate.…
detection, recognition, traffic, traffic sign, urbanThe TRaffic ANd COngestionS (TRANCOS) dataset, a novel benchmark for (extremely overlapping) vehicle counting in traffic congestion situations. It consist…
car, detection, highway, object, spain, traffic, transportation, urban, vehicleCollected in a clothing store. Captured with Kinect (640*480, about 30fps)
detection, trackingThe Brodatz dataset consists of 112 textures in grayscale images of various texture types. http://www.ee.oulu.fi/research/imag/texture/image_data/Brodat…
benchmark, classification, segmentation, synthetic, textureThe Video Segmentation Benchmark (VSB100) provides ground truth annotations for the Berkeley Video Dataset, which consists of 100 HD quality videos divide…
benchmark, groundtruth, motion, object, pedestrian, segmentation, tracking, videoThe Street View Text (SVT) dataset contains 647 words and 3796 letters in 249 images harvested from Google Street View. The dataset is more challenging …
classification, outdoor, text detection, text recognition, urbanThe Where Who Why (WWW) dataset provides 10,000 videos with over 8 million frames from 8,257 diverse scenes, therefore offering a superior comprehensive d…
crowd, detection, flow, optical, pedestrian, recognition, surveillance, videoDaimler Multi-Cue, Occluded Pedestrian Classification Benchmark Training and test samples have a resolution of 48 x 96 pixels with a 12-pixel border aro…
image classification, object detection, pedestrian, urbanMany different labeled video datasets have been collected over the past few years, but it is hard to compare them at a glance. So we have created a handy …
action, benchmark, classification, detection, object, recognition, videoThe GaTech VideoContext dataset consists of over 100 groundtruth annotated outdoor videos with over 20000 frames for the task of geometric context evalua…
classification, context, geometry, nature, outdoor, segmentation, semantic, supervised, unsupervised, urban, videoWIDER FACE dataset is a large-scale face detection benchmark dataset with 32,203 images and 393,703 face annotations, which have high degree of variabilit…
detection, face, occlusion, pose, scaleBelgiumTSC dataset is built for traffic sign classification purposes. Is is a subset of BelgiumTS dataset and contains cropped images around annotations f…
belgium, classification, road, sign, traffic, urbanRobust Multi-Person Tracking from Mobile Platforms In all cases, data was recorded using a pair of AVT Marlins F033C mounted on a chariot respectively a…
color, pedestrian, sequence, trackingPenn-Fudan Pedestrian Detection and Segmentation
background, detection, motion, pedestrian, segmentationThe Person Re-ID (PRID) 2011 dataset was created in co-operation with the Austrian Institute of Technology for the purpose of testing person re-identifica…
appearance, change, classification, graz, identification, illumination, multiview, pedestrian, trajectoryThe Caltech Lanes dataset includes four clips taken around streets in Pasadena, CA at different times of day. The archive below includes 1225 individual…
caltech, detection, lane, pasadena, road, urbanThe Comprehensive Cars (CompCars) dataset contains data from two scenarios, including images from web-nature and surveillance-nature. The web-nature data …
attribute, car, classification, fine-grained, object, recognition, urban, vehicleThe Stanford Dogs dataset contains images of 120 breeds of dogs from around the world. This dataset has been built using images and annotation from ImageN…
classification, detection, dogs, fine-grained categorizationThis dataset is for people tracking in wide baseline camera networks and was designed as a contest at ICPR 2012. The contest consists of two challenges…
aerial, crowd, object detection, object tracking, occlusion, overlap, pedestrian, trajectoryThe Longterm Pedestrian dataset consists of images from a stationary camera running 24 hours for 7 days at about 1 fps. It used for adaptive detection an…
background, change, coffee, detection, graz, illumination, indoor, multitarget, pedestrian, robustThe Textures volume currently contains 154 images, all monochrome, 129 512x512 and 25 1024x1024. For the Brodatz texture images, the number in parenthes…
benchmark, classification, evaluation, segmentation, synthetic, textureThis ETHZ CVL RueMonge 2014 dataset used for 3D reconstruction and semantic mesh labelling for urban scene understanding. It was first published in [1] …
3d, architecture, benchmark, classification, code, mesh, outdoor, paris, pointcloud, recognition, reconstruction, segmentation, semantic, source, urbanThe HandNet dataset contains depth images of 10 participants hands non-rigidly deforming infront of a RealSense RGB-D camera. This dataset includes 214…
articulation, classification, detection, fingertip, hand, pose, rgbd, segmentation, videoWe provide the three datasets used for testing our system for our ICCV 2007 publication, including annotations. Data was recorded using a pair of AVT Marl…
crowd, object detection, object tracking, occlusion, overlap, pedestrian, trajectoryThe Graz01 dataset by Andreas Opelt and Axel Pinz contains four types of images: bikes, people, background with no bikes, background with no people.
background, bike, clutter, graz, object detection, occlusion, pedestrianThe PASCAL VOC is augmented with segmentation annotation for semantic parts of objects. For example, for the person category, we provide segmentation mask…
detection, human, object, part, pascal, pedestrian, recognition, segmentation, semanticMIT Pedestrian dataset from Papageorgiou and Poggio [IJCV2000] contains 509 training and 200 test images of pedestrians in city scenes (plus left-right re…
boundingbox, frontview, object detection, pedestrian, people, urbanThis UIUC Cars dataset by Shivani Agarwal, Aatif Awan and Dan Roth contains images of side views of cars for use in evaluating object detection algorithms…
car, detection, recognition, scale, sideview, urbanThe MOT Challenge is a framework for the fair evaluation of multiple people tracking algorithms. In this framework we provide: - A large collection of d…
3d, benchmark, benhttp://motchallenge.net/chmark, dataset, evaluation, multiple, pedestrian, people, surveillance, target, tracking, videoAbstract Scene understanding has (again) become a focus of computer vision research, leveraging advances in detection, context modeling, and tracking. In…
3d, car, classification, pedestrian, scene, segmentation, semantic, understandingThe CALTECH 256 dataset by Li Fei-Fei contains 30607 images for 256 categories.
centered, classification, detection, image, object, sceneThe Yotta dataset consists of 70 images for semantic labeling given in 11 classes. It also contains multiple videos and camera matrices for 14km or drivin…
3d, camera, classification, reconstruction, segmentation, semantic, urban, videoClassification/Detection Competitions, Segmentation Competition, Person Layout Taster Competition datasets
classification, detectionThe Caltech Pedestrian Dataset consists of approximately 10 hours of 640x480 30Hz video taken from a vehicle driving through regular traffic in an urban e…
object detection, pedestrian, urban3 datasets: PTZ Tracking, Thermal-visible registration, Single object tracking
pedestrian, ptz, thermal, trackingThe PETS 2016 IPATCH dataset contains a set of fourteen multi camera recordings (visible, themal) collected off the coast of Brest, France, in collaborati…
boat, detection, gps, maritime, multimodal, radar, thermal, tracking, vessel, visibleThis database does NOT use a standard set of attributes per instance. Contact Ray Bareiss (rbareiss '@' uunet.uucp ?) for more information. Domain exper…
classification, multivariateThe file animals.c is a data generator of structured instances representing quadruped animals as used by Gennari, Langley, and Fisher (1989) to evaluate t…
classification, generator, multivariateEach record represents follow-up data for one breast cancer case. These are consecutive patients seen by Dr. Wolberg since 1984, and include only those c…
classification, multivariate, regressionDatabase contains 798 images of 114 persons, with 7 images per person and is freely available for research purposes. All images were taken in supervised c…
biometry, face, human, illumination, lighting, pedestrian, person, recognitionTRAINING File : I have created training file with 100+ non malacious examples and 250+ malacious samples. NON-MALACIOUS dataset is represented by +1 while…
classification, multivariateThe PD database consists of training and test files. The training data belongs to 20 PWP (6 female, 14 male) and 20 healthy individuals (10 female, 10 mal…
classification, multivariate, regressionThe main idea of this data set is to prepare the algorithm of the expert system, which will perform the presumptive diagnosis of two diseases of urinary …
classification, multivariateThere are four data sets representing different conditions of an experiment. All have the same attributes. a. adult-stretch.data Inflated is true if age…
classification, multivariateUncompressing rcv1rcv2aminigoutte.tar.bz2 will create a directory that contains 5 subdirectories EN, FR, GR, IT and SP, corresponding to the 5 languages.…
classification, multivariateThe Colosseum and San Marco are two image datasets for dense multiview stereo reconstructions used for evaluating the visual photo realism. The datasets…
3d reconstruction, aerial, flickr, landmark, photo-realism, sfm, streetside, urbanWe used preprocessing programs made available by NIST to extract normalized bitmaps of handwritten digits from a preprinted form. From a total of 43 peopl…
classification, multivariateMachine learning is used in high-energy physics experiments to search for the signatures of exotic particles. These signatures are learned from Monte Carl…
classification, multivariateThe dataset is composed by two tables. The first table go_track_tracks presents general attributes and each instance has one trajectory that is represente…
classification, multivariate, regressionEach image can be characterized by the pose, expression, eyes, and size. There are 32 images for each person capturing every combination of features. To…
classification, imageThe FaceScrub dataset comprises a total of 107818 unconstrained face images of 530 celebrities crawled from the Internet, with about 200 images per person…
celebrity, detection, face, human, people, recognition# From Garavan Institute # Documentation: as given by Ross Quinlan # 6 databases from the Garavan Institute in Sydney, Australia # Approximately the follo…
classification, domain-theory, multivariateData were extracted from images that were taken from genuine and forged banknote-like specimens. For digitization, an industrial camera usually used for …
classification, multivariateExtraction was done by Barry Becker from the 1994 Census database. A set of reasonably clean records was extracted using the following conditions: ((AAGE…
classification, multivariateThe data consist of evaluations of teaching performance over three regular semesters and two summer semesters of 151 teaching assistant (TA) assignments a…
classification, multivariateNotes: -- 3 classes of waves -- 21 attributes, all of which include noise -- See the book for details (49-55, 169) -- waveform.data.Z …
classification, generator, multivariateA description of the underlying Cargo 2000 standard and the processes reflected in the data set can be found at [Web Link].
classification, multivariate, regression, sequentialThe phishing problem is considered a vital issue in .COM industry especially e-banking and e-commerce taking the number of online transactions involving p…
classification, multivariateThe Extreme Zoom Dataset. EZD is a 6 image sets with incleasing zoom factor from general scene view to focusing on single detail. MODS: Fast and Robust …
description, detection, feature, matching, viewpoint, zoomThis dataset represents a set of possible advertisements on Internet pages. The features encode the geometry of the image (if available) as well as phras…
classification, multivariate15,560 pedestrian and non-pedestrian samples (image cut-outs) and 6744 additional full images not containing pedestrians for bootstrapping. The test set c…
detectionMany real world applications need to know the localization of a user in the world to provide their services. Therefore, automatic user localization has be…
classification, multivariate, regressionWe present the 2017 DAVIS Challenge, a public competition specifically designed for the task of video object segmentation. Following the footsteps of othe…
benchmark, code, hd, object, quality, resolution, segmentation, tracking, video segmentationThe donation includes 5 datasets, each of them defining a different learning problem: * LP1: failures in approach to grasp position * LP2: failur…
classification, multivariate, time-seriesThe Inria Aerial Image Labeling addresses a core topic in remote sensing: the automatic pixelwise labeling of aerial imagery (link to paper). Dataset fe…
aerial, building, city, footprint, groundtruth, house, segmentation, semantic, urbanThis dataset was used in several classifications tasks related to the challenge of anuran species recognition through their calls. It is a multilabel data…
classification, clustering, multivariateMining activity was and is always connected with the occurrence of dangers which are commonly called mining hazards. A special case of such threat is a s…
classification, multivariateThe dataset is about bankruptcy prediction of Polish companies. The data was collected from Emerging Markets Information Service (EMIS, [Web Link]), which…
classification, multivariateThis dataset was created for the Paper 'From Group to Individual Labels using Deep Features', Kotzias et. al,. KDD 2015 Please cite the paper if you want …
classification, textThe Berkeley DeepDrive Video Dataset contains 2x order of magnitude more video training data.
autonomous, deep, driving, endtoend, learning, urbanMICCAI 2015 Challenge on Liver Ultrasound Tracking Munich, October 9, 2015 (Full Day) Outline Ultrasound (US) imaging is a widely used medical imaging…
benchmark, human, liver, medical, organ, real, therapy, tracking, ultrasoundUSPTO Algorithm Challenge, run by NASA-Harvard Tournament Lab and TopCoder Problem: Patent Labeling
classification, domain-theorySince its launch in September 1999, Space Imaging IKONOS earth imaging satellite has provided a reliable stream of image data that has become the standard…
3d reconstruction, aerial, photogrammetry, sfm, urbanA complex modern semi-conductor manufacturing process is normally under consistent surveillance via the monitoring of signals/variables collected from sen…
causal-discovery, classification, multivariateMIT traffic data set is for research on activity analysis and crowded scenes. It includes a traffic video sequence of 90 minutes long. It is recorded by a…
trackingOne of the challenges faced by our research was the unavailability of reliable training datasets. In fact this challenge faces any researcher in the field…
classificationThe Caltech Game Covers dataset consists of CD/DVD covers of video games. The set was downloaded from freecovers.net during the summer of 2008. The set in…
caltech, classification, cover, game, hierarchy, retrieval, taxonomyThe Symmetry set dataset is a collection of images at different illuminations for the purpose of image matching using local symmetry features. Image Mat…
building, feature, illumination, image, lighting, matching, symmetry, urbanContains drawing pages from US patents with manually labeled figure and part labels.
detectionThis dataset includes the recordings of five replicates of an 8-sensor array. Each unit holds 8 MOX sensors and integrates custom-designed electronics for…
classification, domain-theory, multivariate, regression, time-seriesThe UMD Dynamic Scene Recognition dataset consists of 13 classes and 10 videos per class and is used to classify dynamic scenes. The dataset has been de…
classification, dynamic, motion, recognition, scene, videoThe INRIA People dataset from Navneet Dalal and Bill Triggs [DalalCVPR2005] consists of training and testing data. The training contains 1805 images and X…
boundingbox, frontview, human, object detection, pedestrian, sideviewThis data set contains 13,910 measurements from 16 chemical sensors exposed to 6 gases at different concentration levels. This dataset is an extension of …
causa, classification, clustering, multivariate, regression, time-seriesThe Synthetic CAD Models dataset consists of X synthetic CAD models for detection (planar) primitives. Efficient RANSAC for Point-Cloud Shape Detection …
3d object, model fitting, primitive, ransac, reconstruction, syntheticThis database contains 76 attributes, but all published experiments refer to using a subset of 14 of them. In particular, the Cleveland database is the o…
classification, multivariateIn HouseCraft, we utilize rental ads to create realistic textured 3D models of building exteriors. In particular, we exploit the address of the property a…
building, city, floorplan, house, localization, registration, segmentation, semantic, streetview, urbanThese data are the results of a chemical analysis of wines grown in the same region in Italy but derived from three different cultivars. The analysis dete…
classification, multivariateGlobal Symmetry Ground-truth for AVA dataset Release Date: 2016 For detailed information, please refer to: Elawady, Mohamed, Ccile Barat, Christophe …
aesthetic, bilateral, detection, global, mirror, reflection, symmetryThis dataset provides a collection of web images and 3D models for research on landmark recognition (especially for methods based on 3D models). We hope i…
3d, classification, codebook, feature, flickr, landmark, matching, recognition, reconstruction, retrievalZuBuD+, created in February 2017 by Federico Magliani (University of Parma), introduces many query images balancing the class evaluated from the previous …
building, image retrieval, landmark, urbanPictures of objects belonging to 256 categoriesPictures of objects belonging to 256 categories.
classification, natural-imageThe Quad 6K dataset is a Structure-from-Motion dataset taken at Arts Quad at Cornell University campus and consists of 6514 images with ground truth posit…
3d gps, 3d reconstruction, groundtruth, landmark, sfm, urbanThe We Are Family Stickmen dataset from Eichner and Ferrari contains X images with X people in group photos for human pose estimation with annotated 2D hu…
body part, object pose, pedestrian-- The users' knowledge class were classified by the authors using intuitive knowledge classifier (a hybrid ML technique of k-NN and meta-heuristic…
classification, clustering, multivariateThe Hopkins 155 Dataset has been created with the goal of providing an extensive benchmark for testing feature based motion segmentation algorithms. It co…
motion segmentation, optical flow, stereo estimation, urbanThe Ford Car dataset is joint effort of Pandey et al. (for collecting images, Lidar points, calibration etc.) and us (for annotation of 2D and 3D objects)…
3d, car, detection, groundtruth, lidar, sfmA synthetic light field dataset with 24 scenes. Data provided for each scene: - 9x9x512x512x3 light fields as individual PNGs - config files with cam…
depth, disparity, ground truth, light field, syntheticAll the patients suffered heart attacks at some point in the past. Some are still alive and some are not. The survival and still-alive variables, when ta…
classification, multivariateThe objective is to identify each of a large number of black-and-white rectangular pixel displays as one of the 26 capital letters in the English alphabet…
classification, multivariateMADELON is an artificial dataset containing data points grouped in 32 clusters placed on the vertices of a five dimensional hypercube and randomly labeled…
classification, multivariateThe dataset contains cases from a study that was conducted between 1958 and 1970 at the University of Chicago's Billings Hospital on the survival of patie…
classification, multivariateThe Daimler Urban Segmentation Dataset consists of video sequences recorded in urban traffic. The dataset consists of 5000 rectified stereo image pairs wi…
motion, outdoor, segmentation, semantic, stereo, urbanThe YouTube-Objects dataset is composed of videos collected from YouTube by querying for the names of 10 object classes. It contains between 9 and 24 vide…
detection, flow, object, optical, segmentation, videoImpedance measurements were made at the frequencies: 15.625, 31.25, 62.5, 125, 250, 500, 1000 KHz Impedance measurements of freshly excised breast tissue …
classification, multivariateThe two datasets are related to red and white variants of the Portuguese "Vinho Verde" wine. For more details, consult: [Web Link] or the reference [Corte…
classification, multivariate, regressionNotes: -- 3 classes of waves -- 40 attributes, all of which include noise -- The latter 19 attributes are all noise attributes with mean…
classification, generator, multivariateISPRS and EuroSDR - Benchmark on High Density Aerial Image Matching Background and Scope of the project Innovations in matching algorithms as well as …
3d, aerial, benchmark, city, germany, multiview, photogrammetry, reconstruction, switzerland, urbanThis data approach student achievement in secondary education of two Portuguese schools. The data attributes include student grades, demographic, social a…
classification, multivariate, regressionThe data was collected by the Magellan spacecraft over an approximately four year period from 1990--1994. The objective of the mission was to obtain globa…
classification, imageThe Microsoft COCO (mscoco) is an image recognition and segmentation dataset which contains more 300k images for more than 70 categories. Other features…
benchmark, context, detection, object, recognition, segmentation, semanticThis file concerns credit card applications. All attribute names and values have been changed to meaningless symbols to protect confidentiality of the da…
classification, multivariateThe automated analysis of facial expressions has been widely used in different research areas, such as biometrics or emotional analysis. Special import…
classification, clustering, multivariate, sequentialThe table below lists the datasets, the YouTube video ID, the amount of samples in each class and the total number of samples per dataset. Dataset --- Yo…
classification, textFor Further information about the variables see the file in the data folder.
classification, textBrief Description of the Dataset: --------------------------------- Each of the 19 activities is performed by eight subjects (4 female, 4 male, between th…
classification, clustering, multivariate, time-seriesThis dataset contains the features extracted from a database of colonoscopic videos showing gastrointestinal lesions. It also contains the ground truth co…
classification, multivariateThe examined group comprised kernels belonging to three different varieties of wheat: Kama, Rosa and Canadian, 70 elements each, randomly selected for the…
classification, clustering, multivariateThe automotive multi-sensor (AMUSE) dataset consists of inertial and other complementary sensor data combined with monocular, omnidirectional, high frame …
api, city, image, inertial, streetside, traffic, urban, videoMIL data sets used in our 2002 NIPS paper for Elepphant, Musk, TREC http://www.cs.cmu.edu/~juny/MILL/MIL-experiments.htm
classification, machine learningThe classification task of this database is to determine where patients in a postoperative recovery area should be sent to next. Because hypothermia is a…
classification, multivariateThe UBO 2014 consists of 7 semantic categories. Each of these 7 material categories contains measurements of 12 different material instances for being cap…
classification, illumination, light, material, recognition, textureWe create a digit database by collecting 250 samples from 44 writers. The samples written by 30 writers are used for training, cross-validation and writer…
classification, multivariateDataset contains 1000 images of 100 persons, with 10 images per person and is freely available. All images were acquired by cropping ears from images from…
biometry, ear, human, lighting, pedestrian, person, recognitionThis challenge is set up around three tasks: Text Localisation, Text Segmentation and Word Recognition. Participation in any or all tasks is welcome. Chec…
classification, text detection, text recognitionThe KU Leuven Facade dataset is used for architectural styles classification. M. Mathias, A. Martinovic, J. Weissenberg, S. Haegler, L. Van Gool: Automa…
architecture, image classification, procedural reconstruction, urbanThe records represent individual data including first and family name, sex, date of birth and postal code, which were collected through iterative insertio…
classification, multivariateThe Zurich Building dataset (ZuBud) from Hao Shao, Tomas Svoboda and Luc Van Gool [?] contains 1005 images with 201 buildings each in five views. There is…
image retrieval, procedural, rectification, urban--- By using a tweet crawler, we collect 2000 labelled tweets (1000 positive tweets and 1000 negative ones) on various topics such as: politics an…
classification, textThe contour patches dataset is a large dataset of images patch matches used for contour detection. References: C. L. Zitnick and D. Parikh The Role o…
contour, detection, edge, image, lowlevel, match, patch, segmentationAIMS AND PURPOSES This corpus is intended to do cleaning (or binarization) and enhancement of noisy grayscale printed text images using supervised learni…
classification, multivariate, regressionAn Inductive Logic Programming (ILP) or relational learning framework is assumed (Muggleton, 1992). The learning system is provided with examples of chess…
classification, multivariateA Vicon motion capture camera system was used to record 12 users performing 5 hand postures with markers attached to a left-handed glove. A rigid patter…
classification, clustering, multivariateIMPORTANT: we have lower performance on 'leave-one-subject-out' tests. The performance baseline index we established is for 10-fold cross-validation tests…
classification, sequentialSince the publicly available face image datasets are often of small to medium size, rarely exceeding tens of thousands of images, and often without age in…
age, biometry, detection, face, imdb, recognition, wikipediaWelcome to the homepage of the gvvperfcapeva datasets. This site serves as a hub to access a wide range of datasets that have been created for projects of…
action, depth, face, human, mesh, multiview, pose, reconstruction, tracking, videoData Characteristics: -------------------- This data was created by selecting 20 files each from the 10 largest classes in the Reuters-21578 collection …
classification, textThis dataset consists of features of handwritten numerals (`0'--`9') extracted from a collection of Dutch utility maps. 200 patterns per class (for a tota…
classification, multivariateThe PIROPO database (People in Indoor ROoms with Perspective and Omnidirectional cameras) comprises multiple sequences recorded in two different indoor ro…
detection, fisheye, human, indoor, omnidirectional, people, perspective, room, surveillanceThere are 19 classes, only the first 15 of which have been used in prior work. The folklore seems to be that the last four classes are unjustified by the …
classification, multivariateThe database contains, for each of the 100 examples: (1) the uncompressed frames, up to the 10th frame after the appearance of the 8th cell; (2) a text fi…
biology, cell, circle, mouse, tracking, trajectoryDataset A (former NLPR Gait Database) was created on Dec. 10, 2001, including 20 persons. Each person has 12 image sequences, 4 sequences for each of the …
action, biometry, classification, foot, gait, human, motion, pressure, recognitionThis data set contains weighted census data extracted from the 1994 and 1995 Current Population Surveys conducted by the U.S. Census Bureau. The data cont…
classification, multivariateData Type: GrayScale Image The image dataset can be used to benchmark classification algorithm for OCR systems. The highest accuracy obtained in the Test…
classificationThis database contains 279 attributes, 206 of which are linear valued and the rest are nominal. Concerning the study of H. Altay Guvenir: "The aim is to…
classification, multivariateThis simple domain contains 7 Boolean attributes and 10 concepts, the set of decimal digits. Recall that LED displays contain 7 light-emitting diodes -- …
classification, generator, multivariateKEGG Metabolic pathways can be realized into network. Two kinds of network / graph can be formed. These include Reaction Network and Relation Network. In …
classification, clustering, multivariate, regression, text, univariateYahoo Flickr Creative Commons 100M (YFCC100M) dataset contains a list of photos and videos. This list is compiled from data available on Yahoo! Flickr. Al…
3d, clustering, community, detection, flickr, image, internet, landmark, recognition, reconstruction, socialThe Infra-Red Astronomy Satellite (IRAS) was the first attempt to map the full sky at infra-red wavelengths. This could not be done from ground observato…
classification, multivariateThis database has been artificially generated by using a first order theory which describes the structure of ten capital letters of the English alphabet a…
classification, multivariateThe PAMAP2 Physical Activity Monitoring dataset contains data of 18 different physical activities (such as walking, cycling, playing soccer, etc.), perfor…
classification, multivariate, time-seriesBackground information: The data set concerns the earliest history of mankind. Prehistoric men created the desired shape of a stone tool by striking on a …
causal-discovery, classification, clustering, multivariateThe 5473 examples comes from 54 distinct documents. Each observation concerns one block. All attributes are numeric. Data are in a format readable by C4.5.
classification, multivariateWe developed imaging hardware consisting of a color camera, a thermal camera and a beam splitter to capture the aligned multispectral (RGB color + Thermal…
pedestrian, rgb, thermalPedestrian Color Naming (PCN) dataset contains 14,213 images, each of which hand-labeled with color label for each pixel.
color naming, pedestrian, segmentationFormat: Each observation concerns one university. In some cases, more information is provided about the attribute (e.g., units or domain). Some duplicates…
classification, multivariateFrom the original readme file (please consult it for more information): ------------------------- The documents in the Reuters-21578 collection appeared o…
classification, textInstance recognition from depth data. Contains various challenges of Pose, Clutter, Occlusion and similar looking objects (Bonde, U., Badrinarayanan, V., …
depth, detection, instance, poseZurich City Hall dataset (also CIPA dataset) nformation: Place: City Hall, Zurich, Switzerland Number of Images: 15, 1280 x 1000 pixels Camera: Fuji …
3d reconstruction, photogrammetry, sfm, urban, zurich* The articles were published by Mashable (www.mashable.com) and their content as the rights to reproduce it belongs to them. Hence, this dataset does not…
classification, multivariate, regressionProblem Description: Splice junctions are points on a DNA sequence at which `superfluous' DNA is removed during the process of protein creation in highe…
classification, domain-theory, sequentialThe "spam" concept is diverse: advertisements for products/web sites, make money fast schemes, chain letters, pornography... Our collection of spam e-mai…
classification, multivariateThe dataset is the subset of RCV1. These corpus has already been used in author identification experiments. In the top 50 authors (with respect to total s…
classification, clustering, domain-theory, multivariate, textFor further details on this dataset and/or its attributes, please read the 'ReadMe.pdf' file included and/or consult the Master's Thesis 'Development of a…
classification, multivariateThis database is a standardized version of the original audiology database (see audiology.* in this directory). The non-standard set of attributes have b…
classification, multivariateScene Background Initialization (SBI) dataset The SBI dataset has been assembled in order to evaluate and compare the results of background initializati…
background, benchmark, change, detection, foreground, initializationContains training and testing data for classifying a high resolution aerial image into 9 types of urban land cover. Multi-scale spectral, size, shape, and…
classification, multivariateThe REALDISP (REAListic sensor DISPlacement) dataset has been originally collected to investigate the effects of sensor displacement in the activity recog…
classification, multivariate, time-seriesThis database contains all legal 8-ply positions in the game of connect-4 in which neither player has won yet, and in which the next move is not forced. …
classification, multivariate, spatialThe San Francisco Landmark Dataset for Mobile Landmark Recognition is a set of images and query images for localization. We present the San Francisco La…
calibration, city, gps, landmark, localization, mobile, retrieval, sanfrancisco, urbanThe dataset contains 15 documentary films that are downloaded from YouTube, whose durations vary from 9 minutes to as long as 50 minutes, and the total nu…
detection, object, videoThe Berkeley Multimodal Human Action Database (MHAD) contains 11 actions performed by 7 male and 5 female subjects in the range 23-30 years of age except …
action, classification, motion, multiview, recognitionThis dataset contains 1800 stereo pairs with ground truth disparity maps, occlusion maps and discontinuity maps that will help to further develop the stat…
graphics, optical, optical flow, stereo depth, syntheticThere are two versions to the database: - V1 contains the original examples and - V2 contains descriptions after discretizing numeric properti…
classification, multivariateThis is a dataset of rectified facade images and semantic labels. The goal of the annotation is to study the layout of the facades. It contains 50 imag…
graz, procedural reconstruction, semantic, semantic segmentation, urbanThe dataset collects data from an Android smartphone positioned in the chest pocket. Accelerometer Data are collected from 22 participants walking in the …
classification, clustering, sequential, time-series, univariateContains 6 object categories similar to object categories in Pascal VOC that are suitable for studying the abnormalities stemming from objects.
detectionThe goal of LabelMe is to provide an online annotation tool to build image databases for computer vision research. You can contribute to the database by v…
object detection, outdoor, semantic, semantic segmentation, software, urbanThe dataset has been enriched during the Nomao Challenge: [Web Link] organized along with the ALRA workshop (Active Learning in Real-world Applications): …
classification, univariateThis data was used by Hong and Young to illustrate the power of the optimal discriminant plane even in ill-posed settings. Applying the KNN method in the …
classification, multivariateScanNet is an RGB-D video dataset containing 2.5 million views in more than 1500 scans, annotated with 3D camera poses, surface reconstructions, and insta…
3d, cad, indoor, layout, object, realism, recognition, rendering, room, scene, segmentation, synthetic30000+ frames with vehicle rear annotation and classification (car and trucks) on motorway/highway sequences. Annotation semi-automatically generated usin…
detectionThis is a transnational data set which contains all the transactions occurring between 01/12/2010 and 09/12/2011 for a UK-based and registered non-store o…
classification, clustering, multivariate, sequential, time-series* The dataset was acquired and annotated by professional physicians at 'Hospital Universitario de Caracas'. * The subjective judgments (target variables) …
classification, multivariateThe Webcam Interestingness dataset consists of 20 different webcam streams, with 159 images each. It is annotated with interestingness ground truth, acqui…
classification, interest, ranking, retrieval, video, weather, webcamThis dataset has been developed to help evaluate a "hybrid" learning algorithm ("KBANN") that uses examples to inductively refine preexisting knowledge. …
classification, domain-theory, sequentialKNAPSACK_01 is a dataset directory which contains some examples of data for 01 Knapsack problems. In the 01 Knapsack problem, we are given a knapsack of…
classification, machine learningDota 2 is a popular computer game with two teams of 5 players. At the start of the game each player chooses a unique hero with different strengths and wea…
classification, multivariateThe Rent3D dataset comprises floorplans and images. The goal of this work is to enable a 3D virtual-tour of an apartment given a small set of monocular im…
apartment, building, floorplan, indoor, layout, reconstruction, urbanThree data sets are submitted, for training and testing. Ground-truth occupancy was obtained from time stamped pictures that were taken every minute. For …
classification, multivariate, time-seriesPlease see the README for the details on the data organization, and so on.
classification, clustering, multivariate, textThe Chars74K dataset consists of 64 classes (0-9, A-Z, a-z), 7705 characters obtained from natural images, 3410 hand drawn characters using a tablet PC, 6…
classification, text detection, text recognitionThe Text and Vision (TVGraz) dataset is an annotated multi-modal dataset which currently contains 10 visual object categories, 4030 images and associated …
appearance, classification, evaluation, textThe ECP Paris 2011 dataset consists of 104 images taken from rue Monge in the fifth district of Paris, we kept only 20 for training and 10 for testing. …
paris, procedural reconstruction, semantic, semantic segmentation, urbanThe Pedestrian Parsing dataset contains 3,673 images from 171 videos of different Surveillance Scenes (PPSS), where 2,064 images are occluded and 1,609 ar…
parsing, pedestrian, segmentation1. Protocol: Three male and one female subjects (age 25 to 30), who have experienced aggression in scenarios such as physical fighting, took part in…
classification, time-seriesThis site is dedicated to provide datasets for the Robotics community with the aim to facilitate result evaluations and comparisons. The datasets presente…
3d, city, laser, nature, urbanThe Robotic 3D Scan Repository from Osnabrueck contains 23 different datasets showing a veriaty of 3D scans for objects, humans, cities, university campus…
3d, aerial, bremen, city, germany, heat, human, laser, lidar, osnabrueck, reconstruction, scan, urbanThis is one of three domains provided by the Oncology Institute that has repeatedly appeared in the machine learning literature. (See also breast-cancer a…
classification, multivariateSeveral constraints were placed on the selection of these instances from a larger database. In particular, all patients here are females at least 21 year…
classification, multivariateThis data set includes descriptions of hypothetical samples corresponding to 23 species of gilled mushrooms in the Agaricus and Lepiota Family (pp. 500-52…
classification, multivariateThe mirror symmetry database contains 176 single-symmetry and 63 multyple-symmetry images (.png files) with accompanying ground-truth annotations (.mat fi…
detection, groundtruth, mirror, symmetryThe parameters which we used for collecting the dataset is referred from the paper 'The discovery of experts decision rules from qualitative bankruptcy da…
classification, multivariateAll data is from one continuous EEG measurement with the Emotiv EEG Neuroheadset. The duration of the measurement was 117 seconds. The eye state was detec…
classification, multivariate, sequential, time-series15 wide baseline stereo image pairs with large viewpoint change, provided ground truth homographies. Image size (~1000x700 pixels, RGB) D. Mishkin and…
description, detection, feature, matching, viewpoint, wide baseline stereoThis dataset contains 12,995 face images which are annotated with (1) five facial landmarks, (2) attributes of gender, smiling, wearing glasses, and head …
attribute, cnn, deep learning, detection, face, landmark detectionMammography is the most effective method for breast cancer screening available today. However, the low positive predictive value of breast biopsy resultin…
classification, multivariateThe Pittsburgh Fast-food Image dataset (PFID) consists of 4545 still images, 606 stereo pairs, 3033600 videos for structure from motion, and 27 privacy-pr…
classification, food, laboratory, real, recognition, reconstruction, videoThis data set contains the acquired time series from 16 chemical sensors exposed to gas mixtures at varying concentration levels. In particular, we genera…
classification, multivariate, regression, time-seriesThis database contains 5 numeric-valued attributes. Only a subset of 3 are used during testing (the latter 3). Furthermore, only 2 of the 3 concepts are…
classification, multivariateNorthix is designed to be a schema matching benchmark problem for data integration of two entity relationship databases. Northix is the resulting schema m…
classification, multivariate, text, univariateThis dataset is a subset of the 1987 National Indonesia Contraceptive Prevalence Survey. The samples are married women who were either not pregnant or do …
classification, multivariatePlaces205 dataase contains 2.5 million images from 205 scene categories for the academic public. The image dataset contains 2,448,873 images from 205 sc…
feature, learning, place, recognition, scene, urbanThe experiments have been carried out with a group of 115 students of first-year, undergraduate Engineering major of the University of Genoa. We carried…
classification, clustering, multivariate, regression, sequential, time-seriesThe CERTH image blur dataset consists of 2450 digital images, 1850 out of which are photographs captured by various camera models in different shooting co…
blur, defocus, detection, image, motion, qualityApproximately 80% of the data belongs to class 1. Therefore the default accuracy is about 80%. The aim here is to obtain an accuracy of 99 - 99.9%. The e…
classification, multivariateA Vicon motion capture camera system was used to record 12 users performing 5 hand postures with markers attached to a left-handed glove. A rigid patter…
classification, clustering, multivariateProvide all relevant informatioThe data has been produced using Monte Carlo simulations. The first 8 features are kinematic properties measured by the par…
classificationThe Zurich Summer v1.0 dataset is a collection of 20 chips (crops), taken from a QuickBird acquisition of the city of Zurich (Switzerland) in August 2002.…
aerial, annotation, city, gsd, nir, pan, rgb, satellite, segmentation, semantic, superpixel, urban, zurichThe MHEALTH (Mobile HEALTH) dataset comprises body motion and vital signs recordings for ten volunteers of diverse profile while performing several physic…
classification, multivariate, time-seriesPredicting forest cover type from cartographic variables only (no remotely sensed data). The actual forest cover type for a given observation (30 x 30 me…
classification, multivariateAll the 504 reviews were collected between January and August of 2015.
classification, regressionThis data set was generated to model psychological experimental results. Each example is classified as having the balance scale tip to the right, tip to …
classification, multivariateThe Aspect Layout dataset is designed to allow evaluation of object detection for aspect ratios in perspective images. Author text: In this project we…
aspect, detection, layout, object, perspective, ratioThis material is supplementary to Michael Stark, Bernt Schiele. How Good are Local Features for Classes of Geometric Objects. Eleventh IEEE Internatio…
binary, classification, object, shape, toolFine-Grained Visual Classification of Aircraft (FGVC-Aircraft) is a benchmark dataset for the fine grained visual categorization of aircraft. Data, anno…
aircraft, airplane, benchmark, classification, evaluation, fine-grained, recognitionCost Matrix _______ abse pres absence 0 1 presence 5 0 where the rows represent the true values and the columns the predicted.
classification, multivariateA New Color Image Database for Benchmarking of Face Detection Techniques and Human Skin Segmentation Techniques. A new color face image database for be…
benchmarking, detection, face, segmentation, skinThe Annotated Facial Landmarks in the Wild (AFLW) consists of a large-scale collection of annotated face images gathered from the web, exhibiting a large …
age, annotation, detection, face, landmark, poseOur dataset is used by us to explore spammers in microblog and you can access our demo system at [Web Link]Please add :8080 after the domain name as port.…
causal-discovery, classification, multivariate, sequential, text, univariateThe main goal of this data set is providing clean and valid signals for designing cuff-less blood pressure estimation algorithms. The raw electrocardiogra…
classification, multivariate, regressionThis dataset was derived from geospatial data from two sources: 1) Landsat time-series satellite imagery from the years 2014-2015, and 2) crowdsourced geo…
classification, multivariateThe file "sonar.mines" contains 111 patterns obtained by bouncing sonar signals off a metal cylinder at various angles and under various conditions. The …
classification, multivariateStyle, Price, Rating, Size, Season, NeckLine, SleeveLength, waiseline, Material, FabricType, Decoration, Pattern, Type, Recommendation are Attributes in d…
classification, clustering, textThe Visual Attributes dataset contains visual attribute annotations for over 500 object classes (animate and inanimate) which are all represented in Image…
attribute, classification, imagenet, object, recognitionThe purpose is to classify a given silhouette as one of four types of vehicle, using a set of features extracted from the silhouette. The vehicle may be …
classification, multivariateThis data set contains 416 liver patient records and 167 non liver patient records.The data set was collected from north east of Andhra Pradesh, India. Se…
classification, multivariateBackground Models Challenge (BMC) is a complete dataset and competition for the comparison of background subtraction algorithms. The main topics concern:…
background, change, detection, modeling, motion, segmentation, surveillance, videoThe Multi-FoV synthetic datasets are two synthetic scenes (vehicle moving in a city, and flying robot hovering in a confined room). For each scene, three …
blender, camera, fov, groundtruth, odometry, synthetic, visualThis dataset is an addition to the dataset at [Web Link] We collected more dataset to improve the accuracy of our HAR algorithms applied in …
classification, time-seriesI collected 64 e-mails from DBWorld newsletter and I used them to train different algorithms in order to classify between 'announces of conferences' and '…
classification, textDataset from 8800(10 digits x 10 repetitions x 88 speakers) time series of 13 Frequency Cepstral Coefficients (MFCCs) had taken from 44 males and 44 femal…
classification, multivariate, time-seriesOur repetitive pattern dataset with 106 images of app. 30 buildings from Pankrac, Prague and Marseille appearing in more than one image, number of appeara…
image classification, image retrieval, repetition, symmetry, urbanData was captured using a setup that consisted of: - Two Fifth Dimension Technologies (5DT) gloves, one right and one left - Two Ascension Flock-of-Bir…
classification, multivariate, time-seriesThis MALDI-TOF dataset consists in:A) A reference panel of 20 Gram positive and negative bacterial species covering 9 genera among which several species a…
classification, multivariateThis data file contains details of various nations and their flags. In this file the fields are separated by spaces (not commas). With this data you can …
classification, multivariateKEGG Metabolic pathways can be realized into network. Two kinds of network / graph can be formed. These include Reaction Network and Relation Network. In …
classification, clustering, multivariate, regression, text, univariateBiophysical models of mutant p53 proteins yield features which can be used to predict p53 transcriptional activity. All class labels are determined via i…
classification, multivariateA large dataset of geotagged face images collected from Flickr. The zip file contains text files containing urls of the images. Face2GPS: Estimating Geo…
age, classification, face, gender, geotagged, human, localizationThe MONK's problem were the basis of a first international comparison of learning algorithms. The result of this comparison is summarized in "The MONK's P…
classification, multivariateThe Dubrovnik6K and Rome16K datasets are image collections for SfM reconstruction, where the suffix refers to the number of images in the dataset. Dubro…
3d reconstruction, dubrovnik, landmark, rome, sfm, urbanThis data set contains training and testing data from a remote sensing study which mapped different forest types based on their spectral characteristics a…
classification, multivariateThe CMP map2photo dataset consists of 6 pairs, where one image is satellite photo and second image is a map of the same area. The task is to match these …
baseline, description, detection, feature, map, matching, remote, sensing, wideThe PETS 2006 dataset contains 7 parts showing multi-sensor sequences containing left-luggage scenarios with increasing scene complexity at a train statio…
frontview, indoor, multitarget, object detection, object tracking, pedestrianToday, we introduce Open Images, a dataset consisting of ~9 million URLs to images that have been annotated with labels spanning over 6000 categories. We …
annotation, automatic, category, classification, deep, image, large-scale, realThe data has been produced using Monte Carlo simulations. The first 21 features (columns 2-22) are kinematic properties measured by the particle detectors…
classificationThe instances were drawn randomly from a database of 7 outdoor images. The images were handsegmented to create a classification for every pixel. Each …
classification, multivariateUncompressing the archive url_svmlight.tar.gz will yield a directory url_svmlight/ containing the following files: * FeatureTypes --- A text file list…
classification, multivariate, time-seriesSee the file bridge-holden-paulson-details.txt in the submitted tarball.
classification, multivariateThe data are MC generated (see below) to simulate registration of high energy gamma particles in a ground-based atmospheric Cherenkov gamma telescope usin…
classification, multivariateThe FlickrLogos-32 dataset contains photos showing brand logos and is meant for the evaluation of multi-class logo recognition as well as logo retrieval m…
classification brand boundingbox, detection, flickr, image, logo, machine learning, object recognition, retrievalThis archive contains 13910 measurements from 16 chemical sensors utilized in simulations for drift compensation in a discrimination task of 6 gases at va…
classification, multivariateThe submitted file is set up as follows. In the first line is the the number of signal events followed by the number of background events. The signal even…
classification, multivariateThe dataset represents 10 years (1999-2008) of clinical care at 130 US hospitals and integrated delivery networks. It includes over 50 features representi…
classification, clustering, multivariateThe original paper demonstrated that it is possible to correctly replicate the experts' binary assessment with approximately 90% accuracy using both 10-fo…
classification, multivariateThe database consists of the multi-spectral values of pixels in 3x3 neighbourhoods in a satellite image, and the classification associated with the centra…
classification, multivariateTh EPFL Multi-View Car dataset contains 20 sequences of cars as they rotate by 360 degrees. There is one image approximately every 3-4 degrees. Using the …
car, detection, estimation, multiview, pose, rotationThis dataset contains 600 examples of control charts synthetically generated by the process in Alcock and Manolopoulos (1999). There are six different cla…
classification, clustering, time-seriesThe Prague Texture Segmentation Datagenerator and Benchmark is designed to mutually compare and rank different (dynamic/static) texture segmenters (superv…
benchmark, synthetic, texture classification, texture segmentationThis dataset contains records of simulation crashes encountered during climate model uncertainty quantification (UQ) ensembles. Ensemble members were co…
classification, multivariateFeatures are extracted from electric current drive signals. The drive has intact and defective components. This results in 11 different classes with diffe…
classification, multivariateThe source of the data is the raw measurements from a Nintendo PowerGlove. It was interfaced through a PowerGlove Serial Interface to a Silicon Graphics 4…
classification, multivariate, time-seriesThe MSR RGB-D Dataset 7-Scenes dataset is a collection of tracked RGB-D camera frames. The dataset may be used for evaluation of methods for different app…
depth, kinect, location, reconstruction, tracking, videoWe create a character database by collecting samples from 11 writers. Each writer contributed with letters (lower and uppercase), digits, and other chara…
classification, multivariate, sequentialThe Kendall Square webcam dataset consists of two streams for one sunny day and one cloudy day of a city square. It is used for tracking and analyzing col…
appearance, change, color, detection, sky, weather, webcamThe Dataset for ADL Recognition with Wrist-worn Accelerometer is a public collection of labelled accelerometer data recordings to be used for the creation…
classification, clustering, multivariate, time-series21 bioassay datasets generated from Pubchem. Both Primary and confirmatory bioassays (12 bioassays, 21 mixes)The data is provided in the same train/test s…
classification, multivariateThe TUD Pedestrians training dataset from Micha Andriluka, Stefan Roth and Bernt Schiele consists of 210 and 400 training images with X pedestrians with s…
object detection, pedestrian, segmentation, sideviewThe York Urban Line Segment Database is a compilation of 102 images (45 indoor, 57 outdoor) of urban environments consisting mostly of scenes from the cam…
geometry, manhattan, outdoor, pose estimation, reconstruction, urban, vanishing pointEach record is an example of a hand consisting of five playing cards drawn from a standard deck of 52. Each card is described using two attributes (suit a…
classification, multivariateThe companion file is a Common Lisp demonstration file that generates knight-pin Chess end-game samples. Start up Lisp and load the file. It generates 10…
classification, generator, multivariateCar Evaluation Database was derived from a simple hierarchical decision model originally developed for the demonstration of DEX, M. Bohanec, V. Rajkovic: …
classification, multivariateThe provided files comprise three different data sets. The first one contains the raw values of the measurements of all 24 ultrasound sensors and the cor…
classification, multivariate, sequentialThe measurements were created to ease the development, comparison and evaluation of fingerprinting based hybrid indoor positioning methods. The measuremen…
classification, multivariate, sequential, time-seriesThe tracking environment consists of multiple 3D range sensors, covering an area of about 900 m2, in the "ATC" shopping center in Osaka, Japan.
trackingdataset are derived from the customers reviews in Amazon Commerce Website for authorship identification. Most previous studies conducted the identifi…
classification, domain-theory, multivariate, textIndoor localization is a key topic for mobile computing. However, it is still very difficult for the mobile sensing community to compare state-of-art Indo…
classification, clustering, multivariate, regression, sequential, time-seriesThe PD and control handwriting database consists of 62 PWP (People with parkinson) and 15 healthy individuals who appealed at the Department of Neurology …
classification, clustering, multivariate, regression-- The users' knowledge class were classified by the authors using intuitive knowledge classifier (a hybrid ML technique of k-NN and meta-heuristic …
classification, multivariateIndoor localisation is a key topic for the Ambient Intelligence (AmI) research community. In this scenarios, recent advancements in wearable technologie…
classification, clustering, multivariate, regression, sequential, time-seriesThe DynTex dataset consists of a comprehensive set of Dynamic Textures. Dynamic, or temporal, texture is a spatially repetitive, time-varying visual patte…
dynamic, segmentation, synthetic, texture, video repetitionThe BEOID dataset includes object interactions ranging from preparing a coffee to operating a weight lifting machine and opening a door. The dataset is re…
3d, egocentric, interaction, object, pose, tracking, videoHallway Corridor - Multiple Camera Tracking: An indoor camera network dataset with 6 cameras (contains ground plane homography).
trackingIn predicting stock prices you collect data over some period of time - day, week, month, etc. But you cannot take advantage of data from a time period unt…
classification, clustering, time-seriesAn Annotated Dataset For Near-Duplicate Detection In Personal Photo Collections Managing photo collections involves a variety of image quality assessmen…
copyright, detection, duplicate, groundtruth, retrievalThe PASCAL VOC Challenge datasets by Mark Everingham is a yearly dataset which has a central evaluation server and the final test data is not released. Th…
airplane, animal, building, car, chair, object detection, object pose, object segmentation, pedestrianThe MPI Sintel Flow dataset is an optical flow / stereo dataset based on the Blender movie Sintel: http://www.sintel.org The goal of this project was to…
graphics, optical flow, stereo depth, syntheticThe most important feature of this dataset is its simplicity to use and its being well-documented, which can be widely used in various studies of text ana…
classification, multivariate, regression, textA dataset for testing object class detection algorithms. It contains 255 test images and features five diverse shape-based classes (apple logos, bottles, …
classificationThe Caltech Buildings dataset consists of images taken for 50 buildings around the Caltech campus. Five different images were taken for each building from…
building, caltech, hierarchy, retrieval, taxonomy, urbanInstrumentation: The data were collected at a sampling rate of 500 Hz, using as a programming kernel the National Instruments (NI) Labview. The signals we…
classification, time-seriesThis dataset contains 503 sponges belonging to the Demospongiae class collected from the Mediterranean (451 sponges) and Atlantic oceans (52 sponges). Eac…
classification, multivariateWe use the following representation to collect the dataset age - age bp - blood pressure sg - specific gravity al - …
classification, multivariateThe data was collected for examining our newly developed classifier for multidimensional curves (multidimensional time series). Nine male speakers uttered…
classification, multivariate, time-seriesPitch classes information has been extracted from MIDI sources downloaded from (JSB Chorales)[[Web Link]]. Meter information has been computed through the…
classification, sequentialThe Babenko tracking dataset contains 12 video sequences for single object tracking. For each clip they provide (1) a directory with the original ima…
animal, face, object tracking, occlusion, single, videoA small subset of the original soybean database. See the reference for Fisher and Schlimmer in soybean-large.names for more information. Steven Souders …
classification, multivariateThe data was collected retrospectively at Wroclaw Thoracic Surgery Centre for patients who underwent major lung resections for primary lung cancer in the …
classification, multivariateChairGest is an open challenge / benchmark. The task consists in spotting and recognizing gestures from multiple synchronized sensors: 1 Kinect and 4 Xse…
benchmark, detection, gesture, human, kinect, recognitionThis dataset describes a set of 92 molecules of which 47 are judged by human experts to be musks and the remaining 45 molecules are judged to be non-musks…
classification, multivariateA new large-scale PEdesTrian Attribute (PETA) dataset. The dataset is by far the largest of its kind, covering more than 60 attributes on 19000 images. I…
crowd, pedestrianThis research aimed at the case of customers default payments in Taiwan and compares the predictive accuracy of probability of default among six data mini…
classification, multivariateWe perform energy analysis using 12 different building shapes simulated in Ecotect. The buildings differ with respect to the glazing area, the glazing are…
classification, multivariate, regressionThe dataset (movement_libras) contains 15 classes of 24 instances each, where each class references to a hand movement type in LIBRAS. In the video pre-p…
classification, clustering, multivariate, sequentialThe Stanford 40 Actions dataset contains images of humans performing 40 actions. In each image, we provide a bounding box of the person who is performing …
action, boundingbox, detection, human, recognitionThe Paris dataset consists of 6412 images. Images have high resolution and are in JPEG format. http://www.robots.ox.ac.uk/~vgg/data/parisbuildings/pari…
image retrieval, landmark, paris, urbanBiomedical data set built by Dr. Henrique da Mota during a medical residence period in the Group of Applied Research in Orthopaedics (GARO) of the Centre …
classification, multivariateThis data was collected from text ads found on twelve websites that deal with various farm animal related topics. Information from the ad creative and th…
classification, text--- The dataset collects data from a wearable accelerometer mounted on the chest --- Sampling frequency of the accelerometer: 52 Hz --- Accelerom…
classification, clustering, sequential, time-series, univariateThe USAA dataset includes 8 different semantic class videos which are home videos of social occassions which feature activities of group of people. It con…
classificationThe Wide (multiple) Baseline Dataset. 31 image pairs, simultaneously combining several nuisance factors: geometry, illumination, IR-visible, etc. WxBS: …
day, description, detection, feature, ir, matching, night, viewpointIdiap/ETHZ Faces and Poses Dataset dataset by L. Jie, B. Caputo and V. Ferrari contains 1703 image-caption pairs. [author] Captions contain the names of s…
face detection, object pose, pedestrian, textThis data set includes votes for each of the U.S. House of Representatives Congressmen on the 16 key votes identified by the CQA. The CQA lists nine diff…
classification, multivariateThis dataset contains features extracted from the Messidor image set to predict whether an image contains signs of diabetic retinopathy or not. All featur…
classification, multivariateNews are grouped into clusters that represent pages discussing the same news story. The dataset includes also references to web pages that, at the access…
classification, clustering, multivariateThe data set gathered when we were working at project for Bahrain university between 2002 and 2003.
classification, clustering, domain-theory, univariateDatabase contains records for 1885 respondents. For each respondent 12 attributes are known: Personality measurements which include NEO-FFI-R (neuroticism…
classification, multivariateAutomatic identification of commercial blocks in news videos finds a lot of applications in the domain of television broadcast analysis and monitoring. Co…
classification, clustering, multivariateThe Fish4Knowledge project (groups.inf.ed.ac.uk/f4k/) is pleased to announce the availability of 2 subsets of our tropical coral reef fish video and ext…
animal, camera, classification, fish, motion, nature, recognition, video, waterThe dataset consists of a total of 3600 documents including 600 news/texts from six categories economy, culture-arts, health, politics, sports and techno…
classification, clustering, textCollect the real time readings for residential,commercial,industrial,agriculure,to find the accuracy consumption in Tamil Nadu Around Thanajvur
classification, clustering, multivariate, regressionThis dataset is composed of a range of biomedical voice measurements from 31 people, 23 with Parkinson's disease (PD). Each column in the table is a parti…
classification, multivariateSceneNet RGB-D is dataset comprised of 5 million Photorealistic Images of Synthetic Indoor Trajectories with Ground Truth. It expands the previous work of…
3d, indoor, lighting, navigation, reconstruction, rendering, robot, scene, segmentation, slam, synthetic, trajectoryThe TUD Campus dataset from Micha Andriluka, Stefan Roth and Bernt Schiele consists of 71 images and 303 highly overlapping pedestrians with large scale c…
object detection, object tracking, overlap, pedestrian, segmentation, sideviewThe experiments were carried out with a group of 30 volunteers within an age bracket of 19-48 years. They performed a protocol of activities composed of s…
classification, multivariate, time-seriesStroke Width Transform Text dataset is by Boris Epstein and consists of 307 images and XXX text instances. Detecting Text in Natural Scenes with Stroke…
classification, text detection, text recognitionThe Yale Face dataset from A. Georghiades contains 5760 single light source images of ten subjects, each shown in 9 poses and 64 illumination setups (lead…
face detection, illumination, pedestrian, pose estimationPeople used for recording of the data were wearing four tags (ankle left, ankle right, belt and chest). Each instance is a localization data for one of t…
classification, sequential, time-series, univariateThis is the database of biological images (from the genetics model system, Drosophila melanogaster, a fruit fly) across multiple levels of variation. we…
animal, biology, classification, fly, genetic, variationCalifornia-ND contains 701 photos taken directly from a real user's personal photo collection, including many challenging non-identical near-duplicate cas…
detection1593 handwritten digits from around 80 persons were scanned, stretched in a rectangular box 16x16 in a gray scale of 256 values.Then each pixel of each im…
classification, multivariateThe Leeds Cows dataset by Derek Magee consists of 14 different video sequences showing a total of 18 cows walking from right to left in front of different…
animal, background, cow, detection, segmentation, videoThe references below describe a predecessor to this dataset and its development. They also give results (not cross-validated) for classification by a rule…
classification, multivariatePart of the problem in using an automated program to discover the unknown target function is to decide how to encode names such that the program can be us…
classification, text, univariate1. Protocol: Seven male and three female subjects (age 25 to 30), who have experienced aggression in scenarios such as physical fighting, took part …
classification, time-seriesF. Bergadano supplied this database. Each instance contains many components, each of which has 8 attributes. Different instances in this database have d…
classification, multivariateThe SUNCG dataset is a Large 3D Model Repository for Indoor Scenes. SUNCG is an ongoing effort to establish a richly-annotated, large-scale dataset of…
3d, indoor, layout, object, realism, recognition, rendering, room, scene, segmentation, syntheticThe DrivFace database contains images sequences of subjects while driving in real scenarios. It is composed of 606 samples of 640480 pixels each, acquired…
classification, clustering, multivariate, regressionPlease find the original data at '[Web Link]'
classification, clustering, multivariate, time-seriesThis is a data set containing 1080 documents of free text business descriptions of Brazilian companies categorized into a subset of 9 categories cataloged…
classification, multivariate, textThe original data were formatted by Thorsten Joachims in the bag-of-words representation. There were 9947 features (of which 2562 are always zeros for all…
classification, multivariateTwo datasets are provided. the original dataset, in the form provided by Prof. Hofmann, contains categorical/symbolic attributes and is in the file "germ…
classification, multivariateISPRS Test Project on Urban Classification, 3D Building Reconstruction and Semantic Labeling. In this part of our working group site you will get further …
3d, aerial, benchmark, canada, city, germany, multiview, photogrammetry, recognition, segmentation, semantic, urbanThe set was recorded in Zurich, using a pair of cameras mounted on a mobile platform. It contains 12'298 annotated pedestrians in roughly 2'000 frames.
tracking2126 fetal cardiotocograms (CTGs) were automatically processed and the respective diagnostic features measured. The CTGs were also classified by three exp…
classification, multivariateISPRS / EuroSDR Benchmark for Multi-Platform Photogrammetry In these pages you can get information about the BENCHMARK FOR MULTI-PLATFORM PHOTOGRAMMETRY…
3d, aerial, benchmark, city, germany, multiview, photogrammetry, reconstruction, switzerland, urbanThe experiments have been carried out with a group of 30 volunteers within an age bracket of 19-48 years. Each person performed six activities (WALKING, W…
classification, clustering, multivariate, time-seriesThis database contains 18000 video frames of 640x480 resolution from 60 video sequences, each of which recorded from a different subject (31 female and 29…
classificationThe dataset was built from a personal collection of 1059 tracks covering 33 countries/area. The music used is traditional, ethnic or `world' only, as cla…
classification, multivariate, regressionThis dataset represents a real-life benchmark in the area of Ambient Assisted Living applications, as described in [1]. The binary classification task con…
classification, multivariate, sequential, time-seriesThe VOT2016 pixel-wise annotations dataset contains pixel-wise per-frame annotations for sequences from VOT2016 dataset. The annotation is in a form of BW…
annotation, mask, object, segmentation, tracking, visualThe dataset is used in the evaluation of EER, FRR and FAR metrics using a new anomaly detector model (Med-Min-Diff). The typed text in the experiment is t…
classification, multivariateOpen University Learning Analytics Dataset (OULAD) contains data about courses, students and their interactions with Virtual Learning Environment (VLE) fo…
classification, clustering, multivariate, regression, sequential, time-series10000 images of natural scenes grabbed on Flickr, with 2695 logos instances cut and pasted from the BelgaLogos dataset.
detectionThis is a subset of the dataset introduced in the SIGGRAPH Asia 2009 paper, Webcam Clip Art: Appearance and Illuminant Transfer from Time-lapse Sequences.…
camera, change, illumination, light, nature, static, time, urban, video, webcamThe TUD Pedestrians dataset from Micha Andriluka, Stefan Roth and Bernt Schiele [AndrilukaCVPR2008] consists of 250 images with 311 fully visible people w…
object detection, pedestrian, segmentation, sideviewThis file concerns credit card applications. All attribute names and values have been changed to meaningless symbols to protect confidentiality of the da…
classification, multivariateExamples represent positive and negative instances of people who were and were not granted credit. The theory was generated by talking to the individual…
classification, domain-theory, multivariate1. Title of Database: Machine Learning based ZZAlpha Stock Recommendations 2. Sources: (a) Original owners of data: ZZAlpha Ltd., 4729 E. Sunrise #109,…
classification, sequential, time-seriesPlease ask Gail Gong for further information on this database.
classification, multivariateThe Malaya Abrupt Motion (MAMo) dataset is targeted for visual tracking, particularly for abrupt motion tracking. It was collected from publicly accessibl…
abrupt motion tracking, tracking, visual trackingHere's the abstract from the above reference: ABSTRACT: Machine learning tools show significant promise for knowledge acquisition, particularly when huma…
classification, multivariateThis radar data was collected by a system in Goose Bay, Labrador. This system consists of a phased array of 16 high-frequency antennas with a total trans…
classification, multivariateThe eTrims dataset is comprised of two datasets, the 4-Class eTRIMS Dataset with 4 annotated object classes and the 8-Class eTRIMS Dataset with 8 annotate…
procedural reconstruction, semantic segmentation, urbanVina conducted a comparison test of her rule-based system, BEAGLE, the nearest-neighbor algorithm, and discriminant analysis. BEAGLE is a product availab…
classification, multivariateThis DataSet is about the results of Statlog project. The project performed a comparative study between Statistical, Neural and Symbolic learning algorith…
classification, multivariateThe examples are complete and noise free. The examples highly simplified the problem. The attributes do not fully describe all the factors affecting the d…
classification, multivariateThe MSR Action datasets is a collection of various 3D datasets for action recognition. See details http://research.microsoft.com/en-us/um/people/zliu/a…
3d, action, detection, recognition, reconstruction, videoThe ICG Graz240 dataset consists of 240 buildings with 5400 redundant images with a total of 5542 window instances. Window detection itself is difficult d…
graz, object detection, semantic, semantic segmentation, urbanProvide all relevant information about your data set.
classification, multivariate, regressionThe digits have been size-normalized and centered in a fixed-size image of dimension 28x28. The original data were modified for the purpose of the feature…
classification, multivariateHTRU2 is a data set which describes a sample of pulsar candidates collected during the High Time Resolution Universe Survey (South) [1]. Pulsars are a r…
classification, clustering, multivariateThe dataset describes diagnosing of cardiac Single Proton Emission Computed Tomography (SPECT) images. Each of the patients is classified into two categor…
classification, multivariateThe data can be used to try to predict student learning in SE teamwork based on observation of their team activity **** README FILE from the submitted d…
classification, sequential, time-seriesCMU/VMR Urban Image+Laser dataset contains 372 images linked with 3D laser points projections. There are additional images (due to the laser scanner being…
3d reconstruction, laser, semantic segmentation, sfm, urbanThis dataset describes a set of 102 molecules of which 39 are judged by human experts to be musks and the remaining 63 molecules are judged to be non-musk…
classification, multivariateVideo data sets to train machines to recognise objects in our environment. e-VDS35 has 35 classes and a total of 2050 videos of roughly 10 seconds each.
classificationWe share our omnidirectional and panoramic image dataset (with annotations) to be used for human and car detection. Please reach through: http://cvrg.iyt…
car, detection, human, omnidirection, panorama, recognitionWe have created the UJIpenchars2 character database by collecting samples from 60 writers at two different sites in two phases: 1st phase, 11 writers, car…
classification, multivariate, sequentialThe user first creates a classification model and then generates classified examples from it. To create a model, the following are specified: the number o…
classification, multivariate- The leaves were placed on a white background and then photographed. - The pictures were taken in broad daylight to ensure optimum light intensity.
classification, clustering, multivariateThe ECP New York dataset contains 10 manually segmented buildings from New York City, USA. Segmentation evaluating using Dice coefficient is calculated fo…
newyork, procedural reconstruction, semantic, semantic segmentation, urbanThe QSAR biodegradation dataset was built in the Milano Chemometrics and QSAR Research Group (Universit degli Studi Milano Bicocca, Milano, Italy). The r…
classification, multivariateThe Weather and Illumination Database (WILD) is an extensive database of high quality images of an outdoor urban scene, acquired every hour over all seaso…
camera, change, depth, estimation, illumination, light, newyork, static, time, urban, video, weather, webcamThe LabelMeFacade dataset contains buildings, windows, sky and a limited number of unlabeled regions (maximally 20% covering of the image). This procedure…
facade, recognition, rectified, segmentation, semantic, urbanThis is a tiny database. Michie reports that Burke's group used RULEMASTER to generate comprehendable rules for determining the conditions under which an…
classification, multivariatet is composed of food intake movements, recorded with Kinect V1 (320240 depth frame resolution), simulated by 35 volunteers for a total of 48 tests. The d…
age, behavior, food, groundtruth, human, intake, kinect, monitoring, pointcloud, trackingYouTube Comedy Slam ([Web Link]) is a video discovery experiment running on YouTube's version of labs (called TestTube) for a few months in 2011 and 2012.…
classification, textLow-resolution RGB videos + ground truth trajectories from multiple fixed and moving cameras monitoring the same scenes (indoor and outdoor) to improve ob…
trackingThe measured data was collected using a chemical sensing system based on an array of 16 metal-oxide gas sensors and an external mechanical ventilator to s…
classification, multivariate, regression, time-seriesFor Each feature, a 64 element vector is given per sample of leaf. These vectors are taken as a contigous descriptors (for shape) or histograms (for textu…
classificationCMP Dataset by Ondra Chum contains 5 million images collected from the internet.
image retrieval, large scale, urbanThis is one of three domains provided by the Oncology Institutenthat has repeatedly appeared in the machine learning literature. (See also breast-cancer …
classification, multivariateThe ICDAR 2003 datasets available for download on this site: Robust Reading , Robust Word Recognition , Robust OCR , Text Locating and Cursive Script . …
classification, text detection, text recognitionThis repository contains labeled 3-D point cloud laser data collected from a moving platform in a urban environment. Data are provided for research purpos…
3d reconstruction, laser, semantic segmentation, sfm, urbanThe Oxford Buildings dataset by James Philbin and Andrew Zisserman consists of 5062 images collected from Flickr by searching for particular Oxford landma…
image retrieval, landmark, oxford, urbanThe skin dataset is collected by randomly sampling B,G,R values from face images of various age groups (young, middle, and old), race groups (white, black…
classification, univariateThis dataset has recordings of a gas sensor array composed of 8 MOX gas sensors, and a temperature and humidity sensor. This sensor array was exposed to b…
classification, multivariate, time-seriesProvide all relevant information about your data set.
classification, clustering, multivariateThe measurements were created to ease the development, comparison and evaluation of fingerprinting based hybrid indoor positioning methods. The measuremen…
causal-discovery, classification, clustering, textThe instances were drawn randomly from a database of 7 outdoor images. The images were handsegmented to create a classification for every pixel. Ea…
classification, multivariateThe Extreme Classification Repository: Multi-label Datasets & Code Kush Bhatia Himanshu Jain Prateek Jain Manik Varma The objective in extreme mult…
benchmark, classification, evaluation, learning, machine, multilabelThis dataset contains 7 challenging volleyball activity classes annotated in 6 videos from professionals in the Austrian Volley League (season 2011/12). A…
action, activity recognition, analysis, detection, sport, video, volleyballThe Deformed Lattice Detection In Real-World Images dataset is used for regular grid detection. The authors have developed a robust and fast lattice detec…
lattice detection, symmetry, texture segmentation, urbanA dataset for Attribute Based Classification. It consists of 30475 images of 50 animals classes with six pre-extracted feature representations for each im…
classificationIn this paper, we look for to recognize the causes of users tend to cyber space in Kohkiloye and Boyer Ahmad Province in Iran. Collecting information to f…
classification, multivariateThe problem is specified by the accompanying data file, "vowel.data". This consists of a three dimensional array: voweldata [speaker, vowel, input]. The …
classificationARCENE was obtained by merging three mass-spectrometry datasets to obtain enough training and test data for a benchmark. The original features indicate th…
classification, multivariateFeatures are computed from a digitized image of a fine needle aspirate (FNA) of a breast mass. They describe characteristics of the cell nuclei present i…
classification, multivariateThe 1DSfM Landmarks is a collection of community-based image reconstruction by Kyle Wilson and is comprised of 14 datasets with comparison to bundler grou…
3d, benchmark, city, groundtruth, landmark, reconstruction, urbanThe Google Street View dataset contains 62,058 high quality Google Street View images. The images cover the downtown and neighboring areas of Pittsburgh, …
address, google, gps, localization, manhattan, panorama, pittsburgh, retrieval, sphere, streetview, urbanSheffield Building Image Dataset consists of over 3,000 low-resolution images of forty different buildings typically between 70 and 120 images per buildi…
image classification, image retrieval, sheffield, urbanPhos is a color image database of 15 scenes captured under different illumination conditions. More particularly, every scene of the database contains 15 …
detectionPN Learning - How does TLD work? Tracking estimates the object location as long as the object is visible. During tracking all observed patterns of the o…
bike, face, learning, object tracking, pedestrian, singletargetEEG record contains many regular oscillations, which are believed to reflect synchronized rhythmic activity in a group of neurons. Most activity related E…
classification, univariateThe data set consists of the expression levels of 77 proteins/protein modifications that produced detectable signals in the nuclear fraction of cortex. Th…
classification, clustering, multivariateThe OPPORTUNITY Dataset for Human Activity Recognition from Wearable, Object, and Ambient Sensors is a dataset devised to benchmark human activity recogni…
classification, multivariate, time-seriesTo demonstrate the RFMTC marketing model (a modified version of RFM), this study adopted the donor database of Blood Transfusion Service Center in Hsin-Ch…
classification, multivariateThis dataset contains Australian legal cases from the Federal Court of Australia (FCA). The cases were downloaded from AustLII ([Web Link]). We included a…
classification, textWe present a dataset to address the problem of visual privacy - where users unintentionally leak private information when sharing personal images online, …
classification, flickr, multilabel, privacy, regression, sceneThe data is related with direct marketing campaigns of a Portuguese banking institution. The marketing campaigns were based on phone calls. Often, more th…
classification, multivariateA dataset of online handwritten assamese characters by collecting samples from 45 writers is created. Each writer contributed 52 basic characters, 10 nume…
classification, multivariate, sequentialThis corpus has been collected from free or free for research sources at the Internet: -> A collection of 425 SMS spam messages was manually extracted fr…
classification, clustering, domain-theory, multivariate, text* Audio track (encoded as mp3) of each of the 106,574 tracks. It is on average 10 millions samples per track.* Nine audio features (consisting of 518 attr…
classification, clustering, multivariate, time-seriesThe dataset is composed by features extracted from 7 videos with people gesticulating, aiming at studying Gesture Phase Segmentation. Each video is repres…
classification, clustering, multivariate, sequential, time-seriesThe Paris Art Deco Facades dataset consists of 79 / 80 images of rectified facades of the architectural style Art Deco, which has different sizes of windo…
architecture, city, facade, grammar, paris, procedural, recognition, segmentation, semantic, urbanThe Symmetry Facades dataset contains 9 building facades with multiple images. It used for coupled symmetry and structure from motion detection. Couple…
3d, building, facade, reconstruction, repetition, sfm, symmetry, urbanThe Leuven Stereo Scene dataset is a scene and depth dataset. There exist two variants of this dataset - a CVPR 2007 paper [1] by Leibe et al. for detecti…
3d, depth, leuven, reconstruction, segmentation, semantic, sfm, stereo, urbanThe Google Street View Pittsburgh Research dataset is a street-level image collection provided by Google for research purposes. The dataset provided he…
3d reconstruction, panorama, pittsburgh, sfm, urbanThis data set was generated as follows. 150 subjects spoke the name of each letter of the alphabet twice. Hence, we have 52 training examples from each sp…
classification, multivariateExtraction was done by Barry Becker from the 1994 Census database. A set of reasonably clean records was extracted using the following conditions: ((AAGE…
classification, multivariateThe characters here were used for a PhD study on primitive extraction using HMM based models. The data consists of 2858 character samples, contained in th…
classification, clustering, time-seriesNotes: - Additional "background" knowledge is supplied that provides a partial ordering on some of the attribute values. - We are providing this dataset…
classification, multivariateThis is perhaps the best known database to be found in the pattern recognition literature. Fisher's paper is a classic in the field and is referenced fre…
classification, multivariateType of dependent variables (7 Types of Steel Plates Faults): 1.Pastry 2.Z_Scratch 3.K_Scatch 4.Stains 5.Dirtiness 6.Bumps 7.Other_Faults
classification, multivariateThe Symmetric Bundle Adjustment dataset contains four sequences of the CAB building, Barcelona, Redmond and Capitole for 3D reconstruction considering sym…
3d reconstruction, bundle adjustment, sfm, symmetry, urbanEach record represents 100 points on a two-dimensional graph. When plotted in order (from 1 through 100) as the Y co-ordinate, the points will create eith…
classification, sequentialThe Heterogeneity Dataset for Human Activity Recognition from Smartphone and Smartwatch sensors consists of two datasets devised to investigate sensor het…
classification, clustering, multivariate, time-series2 data files: -- horse-colic.data: 300 training instances -- horse-colic.test: 68 test instances Possible class attributes: 24 (whether lesi…
classification, multivariateThe HTML source of a web page is given. Users looked at each web page and inidated on a 3 point scale (hot medium cold) 50-100 pages per domain. However, …
classification, multivariate, textThis dataset consists of more than 22,000 images of 24 people which are captured by 16 cameras installed in a shopping mall "Shinpuh-kan". All images are …
trackingThis data set contains some training and testing data from a remote sensing study by Johnson et al. (2013) that involved detecting diseased trees in Quick…
classification, multivariateThe dataset was collected at 'Hospital Universitario de Caracas' in Caracas, Venezuela. The dataset comprises demographic information, habits, and histori…
classification, multivariateDrugs are typically small organic molecules that achieve their desired activity by binding to a target site on a receptor. The first step in the discovery…
classification, multivariateFor a list of attributes, please refer to those two .names files. They use the following naming convention: All the attribute start with T means the tem…
classification, multivariate, sequential, time-seriesThe Daphnet Freezing of Gait Dataset is a dataset devised to benchmark automatic methods to recognize gait freeze from wearable acceleration sensors plac…
classification, multivariate, time-seriesThe Graz02 dataset by Andreas Opelt and Axel Pinz contains four categories of images: bikes, people, cars and a single background class. The annotation ha…
background, bike, car, clutter, graz, object detection, pedestrianA chemical detection platform composed of 8 chemo-resistive gas sensors was exposed to turbulent gas mixtures generated naturally in a wind tunnel. The ac…
classification, multivariate, regression, time-seriesThe Cambridge-driving Labeled Video Database (CamVid) dataset from Gabriel Brostow [?] contains ten minutes of video footage and corresponding semanticall…
3d reconstruction, depth, semantic, semantic segmentation, sfm, urbanThis database encodes the complete set of possible board configurations at the end of tic-tac-toe games, where "x" is assumed to have played first. The t…
classification, multivariateData is collected from imkb.gov.tr and finance.yahoo.com. Data is organized with regard to working days in Istanbul Stock Exchange.
classification, multivariate, regression, time-series, univariate1521 images with human faces, recorded under natural conditions, i.e. varying illumination and complex background. The eye positions have been set manuall…
detectionThis database contains 34 attributes, 33 of which are linear valued and one of them is nominal. The differential diagnosis of erythemato-squamous diseas…
classification, multivariateSamples (instances) are stored row-wise. Variables (attributes) of each sample are RNA-Seq gene expression levels measured by illumina HiSeq platform.
classification, clustering, multivariateThe dataset format is described below. Note: the format of this database was modified on 2/26/90 to conform with the format of all the other databases in…
classification, multivariateThis dataset represents a real-life benchmark in the area of Activity Recognition applications, as described in [1]. The classification tasks consist in…
classification, multivariate, sequential, time-series10000 images of natural scenes, with 37 different logos, and 2695 logos instances, annotated with a bounding box.
detectionSamples arrive periodically as Dr. Wolberg reports his clinical cases. The database therefore reflects this chronological grouping of the data. This group…
classification, multivariateThe Eurasian Cities dataset contains 103 images of outdoor urban scenes taken in Eurasian cities. It is annotated with horizontal and vertical vanishing p…
geometry, line, manhattan, outdoor, point, pose, reconstruction, urban, vanishingLASIESTA is composed by many real indoor and outdoor sequences organized in different categories, each of one covering a specific challenge in moving obje…
background, camera, challenge, dataset, detection, foreground, groundtruth, motion, object, stationary, subtractionThe New College Data Set contains 30GB of data intended for use by the mobile robotics and vision research communities. Our anticipated users are parties …
3d, navigation, odometry, panorama, path, reconstruction, stereo, urbanPredicted Attribute: Localization site of protein. ( non-numeric ). The references below describe a predecessor to this dataset and its development. They…
classification, multivariateThe Ecole Centrale Paris 2010 (Paris 2010) dataset consists of 30 images of densely annotated building facades in seven classes - wall, window, sky, shop,…
paris, procedural reconstruction, semantic, semantic segmentation, urbanPredicting the age of abalone from physical measurements. The age of abalone is determined by cutting the shell through the cone, staining it, and counti…
classification, multivariateThe data consist of 16 binary inputs and one 'four-bit' one-hot classification output. The 16-bit inputs are binary-valued attack-point vectors. 1 indicat…
classification, multivariatePast Usage: (a) Rgnvaldsson, You and Garwicz (2015) 'State of the art prediction of HIV-1 protease cleavage sites', Bioinformatics, vol 31 (8), p…
classification, multivariateThis dataset was collected for training and validation of machine learning algorithm for classification regions of documents on text, picture and backgrou…
classificationVelloso, E.; Bulling, A.; Gellersen, H.; Ugulino, W.; Fuks, H. Qualitative Activity Recognition of Weight Lifting Exercises. Proceedings of 4th Internatio…
classification, multivariateWordNet is a large lexical database of English. Nouns, verbs, adjectives and adverbs are grouped into sets of cognitive synonyms (synsets), each expressin…
category, classification, hierarchy, imagenet, languageLabelMe is a web-based image annotation tool that allows researchers to label images and share the annotations with the rest of the community. If you use …
detectionThe dataset consists of eight unique scenes in crowded spaces such as a university campus or the sidewalks of a busy street.
trackingA simple database containing 17 Boolean-valued attributes. The "type" attribute appears to be the class attribute. Here is a breakdown of which animals …
classification, multivariateThis is a data set used by Ning Qian and Terry Sejnowski in their study using a neural net to predict the secondary structure of certain globular proteins…
classification, sequentialNursery Database was derived from a hierarchical decision model originally developed to rank applications for nursery schools. It was used during several …
classification, multivariateThis is one of three domains provided by the Oncology Institute that has repeatedly appeared in the machine learning literature. (See also lymphography an…
classification, multivariateAWS hosts a variety of public datasets that anyone can access for free. Previously, large datasets such as satellite imagery or genomic data have require…
amazon, biology, classification, deep, human, image, learning, recognition, resolution, satellite, segmentation, spaceThe dataset is composed of 150 synthetic scenes, captured with a (perspective) virtual camera, and each scene contains 3 to 5 objects. The model set is co…
mesh, recognition, segmentation, syntheticThis dataset comprises information regarding the ADLs performed by two users on a daily basis in their own homes. This dataset is composed by two instanc…
classification, clustering, multivariate, sequential, time-seriesParis-rue-Madame dataset contains 3D Mobile Laser Scanning (MLS) data from rue Madame, a street in the 6th Parisian district (France). The test zone conta…
3d, classification, laser, pointcloud, segmentation, semanticThe test sequences provide interested researchers a real-world multi-view test data set captured in the blue-c portals. The data is meant to be used for t…
action, camera, multiview, segmentation, trackingNumber of instances: 18000 times-series measurements recorded from a 72 metal-oxide gas sensor array-based chemical detection platform. Number of attribu…
classification, multivariate, time-seriesWe have downloaded 15 months worth of daily data from the California Department of Transportation PEMS website, [Web Link], The data describes the occupan…
classification, multivariate, time-seriesThe dataset describes diagnosing of cardiac Single Proton Emission Computed Tomography (SPECT) images. Each of the patients is classified into two categor…
classification, multivariate