Data set has no missing values. Values are in kW of each 15 min. To convert values in kWh values must be divided by 4. Each column represent one client. Some clients were created after 2011. In these cases consumption were considered zero. All time labels report to Portuguese hour. However all days present 96 measures (24*15). Every year in March time change day (which has only 23 hours) the values between 1:00 am and 2:00 am are zero for all points. Every year in October time change day (which has 25 hours) the values between 1:00 am and 2:00 am aggregate the consumption of two hours.
Indoor localisation is a key topic for the Ambient Intelligence (AmI) research community. In this scenarios, recent advancements in wearable technologie…
classification, clustering, multivariate, regression, sequential, time-seriesThe experiments have been carried out with a group of 115 students of first-year, undergraduate Engineering major of the University of Genoa. We carried…
classification, clustering, multivariate, regression, sequential, time-seriesThis data set contains 13,910 measurements from 16 chemical sensors exposed to 6 gases at different concentration levels. This dataset is an extension of …
causa, classification, clustering, multivariate, regression, time-seriesThis is a sparse data set, less than 10% of the attributes are used for each sample. The link is to a '*.tgz' file which contains two files: [amzn-anon-ac…
causal-discovery, clustering, domain-theory, regression, time-seriesIndoor localization is a key topic for mobile computing. However, it is still very difficult for the mobile sensing community to compare state-of-art Indo…
classification, clustering, multivariate, regression, sequential, time-seriesOpen University Learning Analytics Dataset (OULAD) contains data about courses, students and their interactions with Virtual Learning Environment (VLE) fo…
classification, clustering, multivariate, regression, sequential, time-seriesThis archive contains 2075259 measurements gathered between December 2006 and November 2010 (47 months). Notes: 1.(global_active_power*1000/60 - sub_mete…
clustering, multivariate, regression, time-seriesCollect the real time readings for residential,commercial,industrial,agriculure,to find the accuracy consumption in Tamil Nadu Around Thanajvur
classification, clustering, multivariate, regressionFor complete information see the official challenge page: [Web Link]
causal-discovery, clustering, domain-theory, multivariate, sequential, time-seriesData is collected from imkb.gov.tr and finance.yahoo.com. Data is organized with regard to working days in Istanbul Stock Exchange.
classification, multivariate, regression, time-series, univariate52 columns for 52 weeks; normalised values of provided too.
clustering, multivariate, time-seriesThis data set contains time series of greenhouse gas (GHG) concentrations at 2921 grid cells in California created using simulations of the Weather Resear…
multivariate, regression, time-seriesThe dataset collects data from an Android smartphone positioned in the chest pocket. Accelerometer Data are collected from 22 participants walking in the …
classification, clustering, sequential, time-series, univariateThe DrivFace database contains images sequences of subjects while driving in real scenarios. It is composed of 606 samples of 640480 pixels each, acquired…
classification, clustering, multivariate, regressionPlease find the original data at '[Web Link]'
classification, clustering, multivariate, time-seriesThe experiments have been carried out with a group of 30 volunteers within an age bracket of 19-48 years. Each person performed six activities (WALKING, W…
classification, clustering, multivariate, time-seriesThis is a transnational data set which contains all the transactions occurring between 01/12/2010 and 09/12/2011 for a UK-based and registered non-store o…
classification, clustering, multivariate, sequential, time-seriesKEGG Metabolic pathways can be realized into network. Two kinds of network / graph can be formed. These include Reaction Network and Relation Network. In …
classification, clustering, multivariate, regression, text, univariateBrief Description of the Dataset: --------------------------------- Each of the 19 activities is performed by eight subjects (4 female, 4 male, between th…
classification, clustering, multivariate, time-series* Audio track (encoded as mp3) of each of the 106,574 tracks. It is on average 10 millions samples per track.* Nine audio features (consisting of 518 attr…
classification, clustering, multivariate, time-seriesThe dataset could contain missing values. The data was sampled every minute, computing and uploading it smoothed with 15 minute means. The header of the d…
multivariate, regression, sequential, text, time-seriesThis dataset was constructed by adding elevation information to a 2D road network in North Jutland, Denmark (covering a region of 185 x 135 km^2). Elevati…
clustering, regression, sequential, textThis dataset contains 600 examples of control charts synthetically generated by the process in Alcock and Manolopoulos (1999). There are six different cla…
classification, clustering, time-seriesThe time period is between Jan 1st, 2010 to Dec 31st, 2015. Missing data are denoted as NA.
multivariate, regression, time-seriesThe Heterogeneity Dataset for Human Activity Recognition from Smartphone and Smartwatch sensors consists of two datasets devised to investigate sensor het…
classification, clustering, multivariate, time-seriesThe Dataset for ADL Recognition with Wrist-worn Accelerometer is a public collection of labelled accelerometer data recordings to be used for the creation…
classification, clustering, multivariate, time-seriesThe dataset contains 9358 instances of hourly averaged responses from an array of 5 metal oxide chemical sensors embedded in an Air Quality Chemical Multi…
multivariate, regression, time-seriesThe PD and control handwriting database consists of 62 PWP (People with parkinson) and 15 healthy individuals who appealed at the Department of Neurology …
classification, clustering, multivariate, regressionThe measured data was collected using a chemical sensing system based on an array of 16 metal-oxide gas sensors and an external mechanical ventilator to s…
classification, multivariate, regression, time-seriesOngoing research on university faculty perceptions and practices of using Wikipedia as a teaching resource. Based on a Technology Acceptance Model, the re…
causal-discovery, clustering, multivariate, regressionThe dataset is composed by features extracted from 7 videos with people gesticulating, aiming at studying Gesture Phase Segmentation. Each video is repres…
classification, clustering, multivariate, sequential, time-seriesIn predicting stock prices you collect data over some period of time - day, week, month, etc. But you cannot take advantage of data from a time period unt…
classification, clustering, time-seriesThis dataset comprises information regarding the ADLs performed by two users on a daily basis in their own homes. This dataset is composed by two instanc…
classification, clustering, multivariate, sequential, time-seriesThis data set contains the acquired time series from 16 chemical sensors exposed to gas mixtures at varying concentration levels. In particular, we genera…
classification, multivariate, regression, time-seriesKEGG Metabolic pathways can be realized into network. Two kinds of network / graph can be formed. These include Reaction Network and Relation Network. In …
classification, clustering, multivariate, regression, text, univariateThe characters here were used for a PhD study on primitive extraction using HMM based models. The data consists of 2858 character samples, contained in th…
classification, clustering, time-seriesA chemical detection platform composed of 8 chemo-resistive gas sensors was exposed to turbulent gas mixtures generated naturally in a wind tunnel. The ac…
classification, multivariate, regression, time-seriesThis dataset includes the recordings of five replicates of an 8-sensor array. Each unit holds 8 MOX sensors and integrates custom-designed electronics for…
classification, domain-theory, multivariate, regression, time-seriesThe data set is at 10 min for about 4.5 months. The house temperature and humidity conditions were monitored with a ZigBee wireless sensor network. Each w…
multivariate, regression, time-series--- The dataset collects data from a wearable accelerometer mounted on the chest --- Sampling frequency of the accelerometer: 52 Hz --- Accelerom…
classification, clustering, sequential, time-series, univariateThe datas time period is between Jan 1st, 2010 to Dec 31st, 2014. Missing data are denoted as NA.
multivariate, regression, time-seriesAIMS AND PURPOSES This corpus is intended to do cleaning (or binarization) and enhancement of noisy grayscale printed text images using supervised learni…
classification, multivariate, regressionA Vicon motion capture camera system was used to record 12 users performing 5 hand postures with markers attached to a left-handed glove. A rigid patter…
classification, clustering, multivariateYou should respect the following train / test split: train: first 463,715 examples test: last 51,630 examples It avoids the 'producer effect' by making su…
multivariate, regressionIn [Cortez and Morais, 2007], the output 'area' was first transformed with a ln(x+1) function. Then, several Data Mining methods were applied. After fi…
multivariate, regressionYahoo Flickr Creative Commons 100M (YFCC100M) dataset contains a list of photos and videos. This list is compiled from data available on Yahoo! Flickr. Al…
3d, clustering, community, detection, flickr, image, internet, landmark, recognition, reconstruction, socialSequential (time-series) domain. Single-line melodies of 100 Bach chorales (originally 4 voices). The melody line can be studied independently of other …
time-series, univariateThe PAMAP2 Physical Activity Monitoring dataset contains data of 18 different physical activities (such as walking, cycling, playing soccer, etc.), perfor…
classification, multivariate, time-seriesBackground information: The data set concerns the earliest history of mankind. Prehistoric men created the desired shape of a stone tool by striking on a …
causal-discovery, classification, clustering, multivariate* The articles were published by Mashable (www.mashable.com) and their content as the rights to reproduce it belongs to them. Hence, this dataset does not…
classification, multivariate, regressionThis data arises from a large study to examine EEG correlates of genetic predisposition to alcoholism. It contains measurements from 64 electrodes placed …
multivariate, time-seriesThe dataset is the subset of RCV1. These corpus has already been used in author identification experiments. In the top 50 authors (with respect to total s…
classification, clustering, domain-theory, multivariate, textThe REALDISP (REAListic sensor DISPlacement) dataset has been originally collected to investigate the effects of sensor displacement in the activity recog…
classification, multivariate, time-seriesThe dataset contains 9568 data points collected from a Combined Cycle Power Plant over 6 years (2006-2011), when the power plant was set to work with full…
multivariate, regressionThree data sets are submitted, for training and testing. Ground-truth occupancy was obtained from time stamped pictures that were taken every minute. For …
classification, multivariate, time-seriesPlease see the README for the details on the data organization, and so on.
classification, clustering, multivariate, text1. Protocol: Three male and one female subjects (age 25 to 30), who have experienced aggression in scenarios such as physical fighting, took part in…
classification, time-seriesThis is a custom generated dataset designed for the task of action co-segmentation in pairs of action sequences. The dataset contains 101 pairs of acti…
action cosegmentation, motion-capture-data, temporal segmentation, time-seriesAll data is from one continuous EEG measurement with the Emotiv EEG Neuroheadset. The duration of the measurement was 117 seconds. The eye state was detec…
classification, multivariate, sequential, time-seriesThis dataset comes from the daily measures of sensors in a urban waste water treatment plant. The objective is to classify the operational state of the pl…
clustering, multivariateNumber of instances 1030 Number of Attributes 9 Attribute breakdown 8 quantitative input variables, and 1 quantitative output variable Missing Attribute …
multivariate, regressionThe NASA data set comprises different size NACA 0012 airfoils at various wind tunnel speeds and angles of attack. The span of the airfoil and the observer…
multivariate, regressionThe experiments have been carried out by means of a numerical simulator of a naval vessel (Frigate) characterized by a Gas Turbine (GT) propulsion plant. …
multivariate, regressionCSV format where each row is a paper and each column is an attribute.
clustering, multivariateA Vicon motion capture camera system was used to record 12 users performing 5 hand postures with markers attached to a left-handed glove. A rigid patter…
classification, clustering, multivariateThe MHEALTH (Mobile HEALTH) dataset comprises body motion and vital signs recordings for ten volunteers of diverse profile while performing several physic…
classification, multivariate, time-seriesAll the 504 reviews were collected between January and August of 2015.
classification, regressionThe main goal of this data set is providing clean and valid signals for designing cuff-less blood pressure estimation algorithms. The raw electrocardiogra…
classification, multivariate, regressionStyle, Price, Rating, Size, Season, NeckLine, SleeveLength, waiseline, Material, FabricType, Decoration, Pattern, Type, Recommendation are Attributes in d…
classification, clustering, textThese are atlantic-mediterranean marine sponges that belong to O.Hadromerida (Demospongiae.Porifera).
clustering, multivariateThis dataset is an addition to the dataset at [Web Link] We collected more dataset to improve the accuracy of our HAR algorithms applied in …
classification, time-seriesDataset from 8800(10 digits x 10 repetitions x 88 speakers) time series of 13 Frequency Cepstral Coefficients (MFCCs) had taken from 44 males and 44 femal…
classification, multivariate, time-seriesData was captured using a setup that consisted of: - Two Fifth Dimension Technologies (5DT) gloves, one right and one left - Two Ascension Flock-of-Bir…
classification, multivariate, time-seriesUncompressing the archive url_svmlight.tar.gz will yield a directory url_svmlight/ containing the following files: * FeatureTypes --- A text file list…
classification, multivariate, time-seriesBike sharing systems are new generation of traditional bike rentals where whole process from membership, rental and return back has become automatic. Thro…
regression, univariateThe dataset represents 10 years (1999-2008) of clinical care at 130 US hospitals and integrated delivery networks. It includes over 50 features representi…
classification, clustering, multivariateThis dataset is a slightly modified version of the dataset provided in the StatLib library. In line with the use by Ross Quinlan (1993) in predicting the…
multivariate, regressionThe source of the data is the raw measurements from a Nintendo PowerGlove. It was interfaced through a PowerGlove Serial Interface to a Silicon Graphics 4…
classification, multivariate, time-seriesThe measurements were created to ease the development, comparison and evaluation of fingerprinting based hybrid indoor positioning methods. The measuremen…
classification, multivariate, sequential, time-seriesThis data set is designed for testing indexing schemes in time series databases. It is a much larger dataset than has been used in any published study (Th…
time-series, univariateThe most important feature of this dataset is its simplicity to use and its being well-documented, which can be widely used in various studies of text ana…
classification, multivariate, regression, textDiabetes patient records were obtained from two sources: an automatic electronic recording device and paper records. The automatic device had an interna…
multivariate, time-seriesInstrumentation: The data were collected at a sampling rate of 500 Hz, using as a programming kernel the National Instruments (NI) Labview. The signals we…
classification, time-seriesThe data was collected for examining our newly developed classifier for multidimensional curves (multidimensional time series). Nine male speakers uttered…
classification, multivariate, time-seriesThis data set consists of three types of entities: (a) the specification of an auto in terms of various characteristics, (b) its assigned insurance risk r…
multivariate, regressionWe perform energy analysis using 12 different building shapes simulated in Ecotect. The buildings differ with respect to the glazing area, the glazing are…
classification, multivariate, regressionThe dataset (movement_libras) contains 15 classes of 24 instances each, where each class references to a hand movement type in LIBRAS. In the video pre-p…
classification, clustering, multivariate, sequentialPrediction of residuary resistance of sailing yachts at the initial design stage is of a great value for evaluating the ships performance and for estimati…
multivariate, regressionNews are grouped into clusters that represent pages discussing the same news story. The dataset includes also references to web pages that, at the access…
classification, clustering, multivariateThe data set gathered when we were working at project for Bahrain university between 2002 and 2003.
classification, clustering, domain-theory, univariateAutomatic identification of commercial blocks in news videos finds a lot of applications in the domain of television broadcast analysis and monitoring. Co…
classification, clustering, multivariateThe dataset consists of a total of 3600 documents including 600 news/texts from six categories economy, culture-arts, health, politics, sports and techno…
classification, clustering, textThe experiments were carried out with a group of 30 volunteers within an age bracket of 19-48 years. They performed a protocol of activities composed of s…
classification, multivariate, time-seriesThe dataset is in the form of a 11463 x 5812 matrix of word counts, containing 11463 words and 5811 NIPS conference papers (the first column contains the …
clustering, textPeople used for recording of the data were wearing four tags (ankle left, ankle right, belt and chest). Each instance is a localization data for one of t…
classification, sequential, time-series, univariateThe data is related to posts' published during the year of 2014 on the Facebook's page of a renowned cosmetics brand. This dataset contains 500 of the 790…
multivariate, regression1. Protocol: Seven male and three female subjects (age 25 to 30), who have experienced aggression in scenarios such as physical fighting, took part …
classification, time-seriesThe dataset was built from a personal collection of 1059 tracks covering 33 countries/area. The music used is traditional, ethnic or `world' only, as cla…
classification, multivariate, regressionThis dataset represents a real-life benchmark in the area of Ambient Assisted Living applications, as described in [1]. The binary classification task con…
classification, multivariate, sequential, time-seriesDocuments are first obtained via a Web search using AMIEI: an integrated platform for delivering enterprise intelligence, developed by AMI Software ([Web …
clustering, multivariate, sequential, textThe All I Have Seen (AIHS) dataset is created to study the properties of total visual input in humans, for around two weeks Nebojsa Jojic wore a camera ca…
3d, clustering, indoor, outdoor, scene, similarity, study, summary, user, video1. Title of Database: Machine Learning based ZZAlpha Stock Recommendations 2. Sources: (a) Original owners of data: ZZAlpha Ltd., 4729 E. Sunrise #109,…
classification, sequential, time-seriesProvide all relevant information about your data set.
classification, multivariate, regressionHTRU2 is a data set which describes a sample of pulsar candidates collected during the High Time Resolution Universe Survey (South) [1]. Pulsars are a r…
classification, clustering, multivariateThe data can be used to try to predict student learning in SE teamwork based on observation of their team activity **** README FILE from the submitted d…
classification, sequential, time-series- The leaves were placed on a white background and then photographed. - The pictures were taken in broad daylight to ensure optimum light intensity.
classification, clustering, multivariateCSV format where each row is a paper and each column an attribute.
clustering, multivariateFor each text collection, D is the number of documents, W is the number of words in the vocabulary, and N is the total number of words in the collection (…
clustering, textRoss Quinlan: This data was given to me by Karl Ulrich at MIT in 1986. I didn't record his description at the time, but here's his subsequent (1992) rec…
multivariate, regressionThis dataset has recordings of a gas sensor array composed of 8 MOX gas sensors, and a temperature and humidity sensor. This sensor array was exposed to b…
classification, multivariate, time-seriesObservations come from 2 data streams (people flow in and out of the building), over 15 weeks, 48 time slices per day (half hour count aggregates). The…
multivariate, time-seriesProvide all relevant information about your data set.
classification, clustering, multivariateThe measurements were created to ease the development, comparison and evaluation of fingerprinting based hybrid indoor positioning methods. The measuremen…
causal-discovery, classification, clustering, textProvide all relevant information about your data set.
multivariate, regressionThe OPPORTUNITY Dataset for Human Activity Recognition from Wearable, Object, and Ambient Sensors is a dataset devised to benchmark human activity recogni…
classification, multivariate, time-seriesThe data set consists of the expression levels of 77 proteins/protein modifications that produced detectable signals in the nuclear fraction of cortex. Th…
classification, clustering, multivariate2. Information database: 2.1. Protocol: 22 male subjects , 11 with different knee abnormalities previously diagnosed by a professional. They undergo thre…
multivariate, time-seriesWe present a dataset to address the problem of visual privacy - where users unintentionally leak private information when sharing personal images online, …
classification, flickr, multilabel, privacy, regression, sceneThis corpus has been collected from free or free for research sources at the Internet: -> A collection of 425 SMS spam messages was manually extracted fr…
classification, clustering, domain-theory, multivariate, textThe data was collected as part of the 1990 census. There are 68 categorical attributes. This data set was derived from the USCensus1990raw data set. The…
clustering, multivariateThere are three disadvantages of weighted scoring stock selection models. First, they cannot identify the relations between weights of stock-picking conce…
multivariate, regressionFor a list of attributes, please refer to those two .names files. They use the following naming convention: All the attribute start with T means the tem…
classification, multivariate, sequential, time-seriesThe Daphnet Freezing of Gait Dataset is a dataset devised to benchmark automatic methods to recognize gait freeze from wearable acceleration sensors plac…
classification, multivariate, time-seriesThe Dataset is uploaded in ZIP format. The dataset contains 5 variants of the dataset, for the details about the variants and detailed analysis read and c…
multivariate, regressionInformation about customers consists of 86 variables and includes product usage data and socio-demographic data derived from zip area codes. The data was …
description, multivariate, regressionThis dataset is composed of a range of biomedical voice measurements from 42 people with early-stage Parkinson's disease recruited to a six-month trial of…
multivariate, regressionSamples (instances) are stored row-wise. Variables (attributes) of each sample are RNA-Seq gene expression levels measured by illumina HiSeq platform.
classification, clustering, multivariateThis dataset represents a real-life benchmark in the area of Activity Recognition applications, as described in [1]. The classification tasks consist in…
classification, multivariate, sequential, time-seriesThere are two databases: (both use the same set of 5 attributes): 1. Primary o-ring erosion and/or blowby 2. Primary o-ring erosion only The two database…
multivariate, regressionThis loop sensor data was collected for the Glendale on ramp for the 101 North freeway in Los Angeles. It is close enough to the stadium to see unusual t…
multivariate, time-seriesThis data originates from blog posts. The raw HTML-documents of the blog posts were crawled and processed. The prediction task associated with the data …
multivariate, regressionThe Airport MotionSeg dataset contains 12 sequences of videos of an aiprort scenario with small and large moving objects and various speeds. It is challen…
airport, camera, clustering, motion, segmentation, video, zoomThe data set includes 103 data points. There are 7 input variables, and 3 output variables in the data set. The initial data set included 78 data. After s…
multivariate, regressionWe collected a video dataset, termed ChokePoint, designed for experiments in person identification/verification under real-world surveillance conditions u…
clustering, detection, face, human, identification, multiview, pedestrian, real, recognition, sequence, surveillance, worldThe data were collected over a series of specifically designed trials. Our hope was to cover most of the types of sensory interactions that a Pioneer migh…
multivariate, time-seriesThe source datasets needed to be combined via programming. Many variables are included so that algorithms that select or learn weights for attributes coul…
multivariate, regressionNumber of instances: 18000 times-series measurements recorded from a 72 metal-oxide gas sensor array-based chemical detection platform. Number of attribu…
classification, multivariate, time-seriesWe have downloaded 15 months worth of daily data from the California Department of Transportation PEMS website, [Web Link], The data describes the occupan…
classification, multivariate, time-seriesThe data was retrieved from a set of 53500 CT images from 74 different patients (43 male, 31 female). Each CT slice is described by two histograms in pol…
domain-theory, regressionEach record represents follow-up data for one breast cancer case. These are consecutive patients seen by Dr. Wolberg since 1984, and include only those c…
classification, multivariate, regressionThe PD database consists of training and test files. The training data belongs to 20 PWP (6 female, 14 male) and 20 healthy individuals (10 female, 10 mal…
classification, multivariate, regressionThe presented dataset is composed of two tsv files named 'youtube_videos.tsv' and 'transcoding_mesurment.tsv'. The first contains 10 columns of fundament…
multivariate, regressionThe dataset is composed by two tables. The first table go_track_tracks presents general attributes and each instance has one trajectory that is represente…
classification, multivariate, regressionA description of the underlying Cargo 2000 standard and the processes reflected in the data set can be found at [Web Link].
classification, multivariate, regression, sequentialMany real world applications need to know the localization of a user in the world to provide their services. Therefore, automatic user localization has be…
classification, multivariate, regressionThe donation includes 5 datasets, each of them defining a different learning problem: * LP1: failures in approach to grasp position * LP2: failur…
classification, multivariate, time-seriesThis dataset was used in several classifications tasks related to the challenge of anuran species recognition through their calls. It is a multilabel data…
classification, clustering, multivariateNotes: -- The database contains 3 potential classes, one for the number of times a certain type of solar flare occured in a 24 hour period. -- Each…
multivariate, regressionMany variables are included so that algorithms that select or learn weights for attributes could be tested. However, clearly unrelated attributes were n…
multivariate, regression-- We aggregated screen movements into screen-fixations using a Salvucci & Goldberg (2000) dispersion-threshold algorithm, and defined Perception Actio…
multivariate, regressionThe estimated relative performance values were estimated by the authors using a linear regression method. See their article (pp 308-313) for more details…
multivariate, regression-- The users' knowledge class were classified by the authors using intuitive knowledge classifier (a hybrid ML technique of k-NN and meta-heuristic…
classification, clustering, multivariateThe two datasets are related to red and white variants of the Portuguese "Vinho Verde" wine. For more details, consult: [Web Link] or the reference [Corte…
classification, multivariate, regressionThis data approach student achievement in secondary education of two Portuguese schools. The data attributes include student grades, demographic, social a…
classification, multivariate, regressionThe automated analysis of facial expressions has been widely used in different research areas, such as biometrics or emotional analysis. Special import…
classification, clustering, multivariate, sequentialThe data is in the transactional form. It contains the Latin names (species or genus) and state abbreviations.
clustering, multivariateThe examined group comprised kernels belonging to three different varieties of wheat: Kama, Rosa and Canadian, 70 elements each, randomly selected for the…
classification, clustering, multivariate