Skip to content
Research Article Open access CC BY 4.0

Assessment of the Different Machine Learning Models for Prediction of Cluster Bean (Cyamopsis tetragonoloba L. Taub.) Yield

Darshan Jagannath Pangarkar, Rajesh Sharma, Amita Sharma, Madhu Sharma

Advances in Research · pp. 98–105 · Published 27 Aug 2020

10.9734/air/2020/v21i930238

Abstract

Prediction of crop yield can help traders, agri-business and government agencies to plan their activities accordingly. It can help government agencies to manage situations like over or under production. Traditionally statistical and crop simulation methods are used for this purpose. Machine learning models can be great deal of help. Aim of present study is to assess the predictive ability of various machine learning models for Cluster bean (Cyamopsis tetragonoloba L. Taub.) yield prediction. Various machine learning models were applied and tested on panel data of 19 years i.e. from 1999-2000 to 2017-18 for the Bikaner district of Rajasthan. Various data mining steps were performed before building a model. K- Nearest Nighbors (K-NN), Support Vector Regression (SVR) with various kernels, and Random forest regression were applied. Cross validation was also performed to know extra sampler validity. The best fitted model was chosen based cross validation scores and R2 values. Besides the coefficient of determination (R2), root mean squared error (RMSE), mean absolute error (MAE), and root relative squared error (RRSE) were calculated for the testing set. Support vector regression with linear kernel has the lowest RMSE (23.19), RRSE (0.14), MAE (19.27) values followed by random forest regression and second-degree polynomial support vector regression with the value of gamma = auto. Instead there was a little difference with R2, placing support vector regression first (98.31%), followed by second-degree polynomial support vector regression with value of gamma = auto (89.83%) and second-degree polynomial support vector regression with value of gamma = scale (88.83%). On two-fold cross validation, support vector regression with a linear kernel had the highest cross validation score explaining 71% (+/-0.03) followed by second-degree polynomial support vector regression with a value of gamma = auto and random forest regression. KNN and support vector regression with radial basis function as a kernel function had negative cross validation scores. Support vector regression with linear kernel was found to be the best-fitted model for predicting the yield as it had higher sample validity (98.31%) and global validity (71%).

Yield machine learning K-NN SVR random forest

Cited by 7

Identifying salient features of cooling energy usage of commercial buildings using explainable artificial intelligence

Lakmini Rangana Senarathne, Gaurav Nanda, Raji Sundararajan · Advances in Building Energy Research · 2023

Precision Geolocation of Medicinal Plants: Assessing Machine Learning Algorithms for Accuracy and Efficiency

Maria Concepcion Suarez Vera · Advances in Technology Innovation · 2024

Transforming agriculture with Machine Learning, Deep Learning, and IoT: perspectives from Ethiopia—challenges and opportunities

Natei Ermias Benti, Mesfin Diro Chaka, Addisu Gezahegn Semie · Discover Agriculture · 2024

Predictive modeling of hydrogen production and methane conversion from biomass-derived methane using machine learning and optimisation techniques

Adegboyega Bolu Ehinmowo, Bright Ikechukwu Nwaneri, Joseph Oluwatobi Olaide · Next Energy · 2025

Ensemble learning prediction of soybean yields in China based on meteorological data

Qian-chuan LI, Shi-wei XU, Jia-yu ZHUANG · Journal of Integrative Agriculture · 2023

Advancing Soil Management and Crop Yield Optimization: A Comprehensive Review of Precision Agriculture Techniques Using Machine Learning, Deep Learning and IoT

Arko Bagchi, Prashant Johri · 2024 2nd International Conference on Advances in Computation, Communication and Information Technology (ICAICCIT) · 2024

Article metrics

Real usage data collected on this platform.

0

Page views

0

PDF downloads

0

Outbound clicks

7

Citations

Views by country

Approximate, from request IP at view time — not citizenship or institution. Countries with fewer than 5 views are grouped as "Other".

No views recorded yet.

Traffic sources

Referring site, by host.

No traffic recorded yet.

Views and downloads exclude known bots/crawlers. Citations combines this platform's own DOI-resolved index with each external source's own reported total — see Cited by above for individually listed citing works. Last refreshed 0 seconds ago.