Assessing the Performance of a Long Short-Term Memory Algorithm in the Dataset with Missing Values

Hyun-Geoun Park,Jinuk Jang,Seo Jin Ki,Sang-Ik Suh,Gyeong Cheol Jo

doi:10.4491/ksee.2022.44.12.636

Abstract

This study was conducted to assess the performance of a long short-term memory algorithm (LSTM), which was suitable for time series prediction, in the multivariate dataset with missing values. The full dataset for the adopted LSTM model was prepared by running a popular watershed model Hydrological Simulation Program-Fortran (HSPF) in the upper Nam River Basin for 3 years from 2016 to 2018, excluding a one-year warm-up period, on a daily time step. The accuracy of prediction for the LSTM model was evaluated in response to various interpolation methods as well as changes in the number of missing values (for dependent variables) and independent variables (containing a fixed number of missing values for either single or multiple variables). Note that the entire dataset is divided into training and test datasets at a ratio of 7:3. Results showed that different interpolation methods resulted in a considerable variation in performance of the LSTM model. Out of them, StructTS and RPART were selected as the best imputation methods recovering missing values for discharge and total phosphorus, respectively. The prediction error of the LSTM model increased gradually with increasing the number of missing values from 300 to 700. The LSTM model, however, appeared to maintain its performance fairly well even in data sets with a large amount of missing values as long as adequate interpolation methods were adopted for each dependent variable. The performance of the LSTM model degraded further as the number of independent variables containing the fixed number of missing values increased from 1 to 7. We believe that the proposed methodology can be used not only to reconstruct missing values in a real-time monitoring dataset with excellent performance, but also to improve the accuracy of prediction for (time series) deep learning models.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Assessing the Performance of a Long Short-Term Memory Algorithm in the Dataset with Missing Values

Abstract

Talk to us

Similar Papers

More From: Journal of Korean Society of Environmental Engineers

Lead the way for us

Journal: Journal of Korean Society of Environmental Engineers	Publication Date: Dec 31, 2022
License type: cc-by-nc

Similar Papers

Automatic Stock Market Prediction using Novel Long Short Term Memory Algorithm compared with Logistic Regression for improved F1 score
P Venkata Sairam ... Logu K
-
P Venkata Sairam, et. al.P Venkata Sairam ... Logu K
23 Feb 2022
23 Feb 2022

Fault Prediction for Software System in Industrial Internet: A Deep Learning Algorithm via Effective Dimension Reduction
Siqi Yang ... Lanlan Rui
-
Siqi Yang, et. al.Siqi Yang ... Lanlan Rui
01 Jan 2019
01 Jan 2019

An Improved Log Loss Stock Market Prediction Using Novel Long Short Term Memory Algorithm in Comparison with Support Vector Machine Algorithm
P Venkata Sairam ... K Logu
-
P Venkata Sairam, et. al.P Venkata Sairam ... K Logu
03 Nov 2022
03 Nov 2022

Forecasting Demand Using ARIMA Model and LSTM Neural Network: a Case of Detergent Manufacturing Industry
Imen Mejri ... Sourour Bacha
-
Imen Mejri, et. al.Imen Mejri ... Sourour Bacha
29 Sep 2021
29 Sep 2021

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Assessing the Performance of a Long Short-Term Memory Algorithm in the Dataset with Missing Values

Abstract

Talk to us

Similar Papers

More From: Journal of Korean Society of Environmental Engineers