How does machine learning compare to conventional econometrics for transport data sets? A test of ML versus MLE

Weijia (Vivian) Li,Kara M Kockelman

doi:10.1111/grow.12587

Abstract

AbstractMachine learning (ML) is being used regularly in many different fields. This paper compares traditional econometric methods that have better explanations of data analysis to ML methods, focusing on predicting, understanding and unpacking ML methods which have higher prediction accuracies of four key transport‐planning variables: household vehicle‐miles traveled (continuous variable), household vehicle ownership (count variable), mode choice (categorical variable), and land use change (categorical variable with strong spatial interactions). Here, the results of ten ML methods are compared to methods of ordinary least squares (OLS), multinomial logit (MNL), negative binomial and spatial auto‐regressive (SAR). The U.S.’s 2017 National Household Travel Survey and land use data sets from the Dallas‐Ft. Worth region of Texas are used. Results suggest traditional econometric methods work pretty well on the more continuous responses (VMT and vehicle ownership), but the random forest (RF), gradient boosting decision trees (GBDT), and extreme gradient boosting (XGBoost) methods delivered the best results, though the RF model required 30 to almost 60 times more computing time than XGBoost and GBDT methods. The RF, GBDT, XGBoost, light gradient boosting method (lightGBM), and catboost offer better results than other methods for the two “classification” cases, with lightGBM being the most time‐efficient. Importantly, ML methods captured the plateauing effect modelers may expect when extrapolating covariate effects.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

How does machine learning compare to conventional econometrics for transport data sets? A test of ML versus MLE

Abstract

Talk to us

Similar Papers

More From: Growth and Change

Lead the way for us

Journal: Growth and Change	Publication Date: Nov 19, 2021
Citations: 12

Similar Papers

Identification of risk factors for infection after mitral valve surgery through machine learning approaches.
Ningjie Zhang ... Yongjun Wang
Frontiers in Cardiovascular Medicine | VOL. 10
Ningjie Zhang, et. al.Ningjie Zhang ... Yongjun Wang
13 Jun 2023
Frontiers in Cardiovascular Medicine | VOL. 10

Machine learning algorithms realized soil stoichiometry prediction and its driver identification in intensive agroecosystems across a north-south transect of eastern China
Xintong Xu ... Zhengqin Xiong
Science of The Total Environment | VOL. 906
Xintong Xu, et. al.Xintong Xu ... Zhengqin Xiong
30 Sep 2023
Science of The Total Environment | VOL. 906

A robust indoor localization method with calibration strategy based on joint distribution adaptation
Yujie Wang ... Yi Lei
Wireless Networks | VOL. 27
Yujie Wang, et. al.Yujie Wang ... Yi Lei
19 Jan 2021
Wireless Networks | VOL. 27

Machine learning-based models to support decision-making in emergency department triage for patients with suspected cardiovascular disease
Huilin Jiang ... Xiaohui Chen
International Journal of Medical Informatics | VOL. 145
Huilin Jiang, et. al.Huilin Jiang ... Xiaohui Chen
03 Nov 2020
International Journal of Medical Informatics | VOL. 145

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

How does machine learning compare to conventional econometrics for transport data sets? A test of ML versus MLE

Abstract

Talk to us

Similar Papers

More From: Growth and Change