Beliebiger Eintrag,

Clustering Time Series Data through Autoencoder-based Deep Learning Models

N. Tavakoli, S. Siami-Namini, M. Khanghah, F. Soltani, und A. Namin.
(2020)cite arxiv:2004.07296.

Zusammenfassung

Machine learning and in particular deep learning algorithms are the emerging approaches to data analysis. These techniques have transformed traditional data mining-based analysis radically into a learning-based model in which existing data sets along with their cluster labels (i.e., train set) are learned to build a supervised learning model and predict the cluster labels of unseen data (i.e., test set). In particular, deep learning techniques are capable of capturing and learning hidden features in a given data sets and thus building a more accurate prediction model for clustering and labeling problem. However, the major problem is that time series data are often unlabeled and thus supervised learning-based deep learning algorithms cannot be directly adapted to solve the clustering problems for these special and complex types of data sets. To address this problem, this paper introduces a two-stage method for clustering time series data. First, a novel technique is introduced to utilize the characteristics (e.g., volatility) of given time series data in order to create labels and thus be able to transform the problem from unsupervised learning into supervised learning. Second, an autoencoder-based deep learning model is built to learn and model both known and hidden features of time series data along with their created labels to predict the labels of unseen time series data. The paper reports a case study in which financial and stock time series data of selected 70 stock indices are clustered into distinct groups using the introduced two-stage procedure. The results show that the proposed procedure is capable of achieving 87.5\% accuracy in clustering and predicting the labels for unseen time series data.

BibTeX-Schlüssel: tavakoli2020clustering
Eintragstyp: misc
Jahr: 2020
URL: http://arxiv.org/abs/2004.07296
Hinweis: cite arxiv:2004.07296

Nutzer

Kommentare und Rezensionenanzeigen / verbergen

Bitte melden Sie sich an um selbst Rezensionen oder Kommentare zu erstellen.

Zitieren Sie diese Publikation

%0 Generic %1 tavakoli2020clustering %A Tavakoli, Neda %A Siami-Namini, Sima %A Khanghah, Mahdi Adl %A Soltani, Fahimeh Mirza %A Namin, Akbar Siami %D 2020 %K autoencoder timeseries %T Clustering Time Series Data through Autoencoder-based Deep Learning Models %U http://arxiv.org/abs/2004.07296 %X Machine learning and in particular deep learning algorithms are the emerging approaches to data analysis. These techniques have transformed traditional data mining-based analysis radically into a learning-based model in which existing data sets along with their cluster labels (i.e., train set) are learned to build a supervised learning model and predict the cluster labels of unseen data (i.e., test set). In particular, deep learning techniques are capable of capturing and learning hidden features in a given data sets and thus building a more accurate prediction model for clustering and labeling problem. However, the major problem is that time series data are often unlabeled and thus supervised learning-based deep learning algorithms cannot be directly adapted to solve the clustering problems for these special and complex types of data sets. To address this problem, this paper introduces a two-stage method for clustering time series data. First, a novel technique is introduced to utilize the characteristics (e.g., volatility) of given time series data in order to create labels and thus be able to transform the problem from unsupervised learning into supervised learning. Second, an autoencoder-based deep learning model is built to learn and model both known and hidden features of time series data along with their created labels to predict the labels of unseen time series data. The paper reports a case study in which financial and stock time series data of selected 70 stock indices are clustered into distinct groups using the introduced two-stage procedure. The results show that the proposed procedure is capable of achieving 87.5\% accuracy in clustering and predicting the labels for unseen time series data.

@misc{tavakoli2020clustering, abstract = {Machine learning and in particular deep learning algorithms are the emerging approaches to data analysis. These techniques have transformed traditional data mining-based analysis radically into a learning-based model in which existing data sets along with their cluster labels (i.e., train set) are learned to build a supervised learning model and predict the cluster labels of unseen data (i.e., test set). In particular, deep learning techniques are capable of capturing and learning hidden features in a given data sets and thus building a more accurate prediction model for clustering and labeling problem. However, the major problem is that time series data are often unlabeled and thus supervised learning-based deep learning algorithms cannot be directly adapted to solve the clustering problems for these special and complex types of data sets. To address this problem, this paper introduces a two-stage method for clustering time series data. First, a novel technique is introduced to utilize the characteristics (e.g., volatility) of given time series data in order to create labels and thus be able to transform the problem from unsupervised learning into supervised learning. Second, an autoencoder-based deep learning model is built to learn and model both known and hidden features of time series data along with their created labels to predict the labels of unseen time series data. The paper reports a case study in which financial and stock time series data of selected 70 stock indices are clustered into distinct groups using the introduced two-stage procedure. The results show that the proposed procedure is capable of achieving 87.5\% accuracy in clustering and predicting the labels for unseen time series data.}, added-at = {2022-02-11T14:37:00.000+0100}, author = {Tavakoli, Neda and Siami-Namini, Sima and Khanghah, Mahdi Adl and Soltani, Fahimeh Mirza and Namin, Akbar Siami}, biburl = {https://www.bibsonomy.org/bibtex/2df5b03a43a380183a4b53a040c0cfeb9/albinzehe}, description = {Clustering Time Series Data through Autoencoder-based Deep Learning Models}, interhash = {516bea8ef662e3e6c710c8c5b4f1eb42}, intrahash = {df5b03a43a380183a4b53a040c0cfeb9}, keywords = {autoencoder timeseries}, note = {cite arxiv:2004.07296}, timestamp = {2022-02-11T14:37:00.000+0100}, title = {Clustering Time Series Data through Autoencoder-based Deep Learning Models}, url = {http://arxiv.org/abs/2004.07296}, year = 2020 }

BibSonomy

Clustering Time Series Data through Autoencoder-based Deep Learning Models

Zusammenfassung

Tags

Nutzer

Kommentare und Rezensionenanzeigen / verbergen

Zitieren Sie diese Publikation

Mehr Zitationsstile

Suchen auf