IntelliPaper
Abstract
A time series is a predetermination of data points that happen in repeated order of time. Forecasting productions play a necessary part in several fields such as, meteorological data, weather data, stock market data, rainfall data, agriculture data and so on. In recent years, fuzzy time series is used for forecasting. Song and Chissom (1993) proposed fuzzy time series for forecasting enrollments of data. In this paper, Autoregressive Integrated Moving Average Model (ARIMA), Neural networks for Radial Basis Function (RBF) and Multilayer Perceptran (MLP) and fuzzy time series for predicting wheat production of India were compared. Mean Absolute Error (MAE), Mean Absolute Percentage Error (MAPE) and Root Mean Square Error (RMSE) were compared. The results were displayed numerically and graphically.
Explore Digital Article Text
I. INTRODUCTION
The time series is a category of variables ordered in a specific order of time. Forecasting believes the future values of the time series. For the past decades, fuzzy time series has been widely used for predicting the historic data. Fuzzy time series is used to planning with forecasting difficulties in which the historical data are linguistic values. Song and Chissom (1993) proposed the fuzzy time series for the enrollments of university of Alabama. Wheat is the chief cereal crop in India. Wheat crop has well-known flexibility. Wheat is developed in a variety of soils of India. Yanpeng Zhang et.al(2020) proposed a novel fuzzy time series forecasting model by multiple linear regression and time series clustering for forecasting market prices. Singh, P. (2018) offered a new model to deal with four major issues of fuzzy time series (FTS) forecasting, viz., Wangren Qiu et.al(2015) proposed model was implemented in forecasting enrollment data at the University of Alabama. Ozge Cagcag Yolcu et.al(2016) proposed a novel high-order fuzzy time series approach that considers the membership values, where artificial neural networks are employed to identify the fuzzy relations. Adesh Kumar Pandey(2008) proposed fuzzy time series and neural network. The proposed method has been implemented in the historical data. Paarth Thadani(2021) presented non-linear forecasting models, including artificial neural networks, are popularly adopted in financial forecasting. Yousif Alyousifi et.al (2021) proposed Fuzzy Time Series Markov Chain – Transition Probability Matrix model is tested using two types of time series data, namely, air pollution index (API) data, and yearly enrollments for the University of Alabama.Wang et.al (2017) applied autoregressive integrated moving average model and artificial neural network for air pollution data. Alyousi et. al (2019) applied artificial neural network and Markov chain are applied for air pollution forecasting. In this work, ARIMA, Neural networks for Multilayer Perceptron and Radial Basics Function and fuzzy time series algorithms are used for wheat production prediction in India. Residual analysis for Mean
Absolute Error (MAE), Mean Absolute Percentage Error (MAPE) and Root Mean Square Error (RMSE) were compared.
II. METHODOLOGY
2.1 Autoregressive Integrated Moving Average Model
Autoregressive Integrated Moving Average Model (p, d, q), where p is Autoregressive and q is the Moving Average Model and d is the differencing. If d = 0, the data exhibits stationary and the order is denoted as (p, q), which is called ARMA process. If the data does not exhibit stationary, the first order differencing is carried out in converting it into a stationary, hence the model is denoted as (p, d, q).
2.2 Radial Basis Function (RBF)
RBF networks, a class of feed forward networks, called radial basis function that compute activations at the hidden neurons in a way that is different from what we have seen in the case of feed forward neural networks. Rather than employing an inner product between the input vector and the weight vector The RBF output layer results in a linear fashion. The output y is computed by
For where is the output of the RBF. is the connection weight from the hidden to the output unit is the prototype or center of the hidden unit, and denotes the Euclidean norms. The RBF typically selected as the Gaussian function
Where is the center of the associated field, and is the width of the Gaussian function.
2.3 Multi Layer Perceptron
Artificial Neural Network(ANN) is an operational model which consists of a large number of interconnected nodes(neurons). Each node contains a specific output function which is called an activation function. The connection between every two nodes represents a weighted value that passes through the connection signal, which is called weight. Weight is equivalent to the memory of ANN. Multi Layer Perceptron(MLP) has many layers, the first layer is the input layer, the last layer is the output layer, the middle layers are called hidden layers, each layer includes several neurons. This calculation process is called feed forward process of MLP. If there is an MLP, which contains m hidden layers, its input and output dimensions are respectively equal to and . The number of nodes in each hidden layer is respectively. In the feed forward process of this MLP, each node value is calculated using the following formula
fis the activation function.
Where represents the value of the j neuron i layer. represents the weight vector of the j neuron in layer i-1 to layer i. represents the value vector of all neurons in layer i-1. represents the bias of the i-1 layer, and f is the activation function.
2.4 Fuzzy Time Series
It is the values of the observations of a special dynamic process are represented by linguistic values.
Computational Algorithm for Fuzzy Time Series
The step by step process is as follows:
Step1: Calculate the first order variation of the historical data.
Step2: Define the universe of discourse, U based on the range of available historical data.
Where is the minimum value of the first order variation of the historical data, is the maximum value of the historical data and , are two positive integers.
Step3: Partition the universe of discourse U into equal length intervals: .
Step4: The number of intervals will be in accordance with the number of linguistic variables (fuzzy sets) to be considered.
Step5: Fuzzify the variations of the historical data and establish the fuzzy logical relationship is represented by .
Step6: Rules for forecasting:
is corresponding interval for which membership in is supremum (i.e., 1)
is the highest value of the interval having supremum value in .
M[A_{j}] is the mid value of the interval u_{j} having supremum value in A_{j}.
For a fuzzy logical relationship .
is the fuzzified wheat production of the current year n;
is the fuzzified wheat production of the next year ;
is the actual wheat production of the current year n;
is the actual wheat production of the previous year n-1;
is the variation wheat production of the current year n;
is the variation wheat production of the previous year n-1;
is the forecasted wheat production of the next year ;
Step7: Forecasting wheat production for the year is obtained from modified computational algorithm as follows;
Obtain the fuzzy logical relationship .
Step8: Obtain the mean absolute error using actual values and forecasted values
Where n is the number of years and \left|u_{t}\right|=Y_{t}-\hat{Y}{t}. Y{t} is actual values at time t. \hat{Y}_{t} is predicted values at time t.
Step9: Obtain the mean absolute percentage error using actual values and forecasted values
Step 10: Obtain the root mean square error using actual values and forecasted values
III. RESULTS AND DISCUSSIONS
Step1: Compute the first order variation of the historical data.
Step2: The universe of discourse U is defined as+
Where is the minimum value of the first order variation of the historical data.
is the maximum value of the first order variation of the historical data,
and are two positive integers. , are choosing arbitrarily for the rounded off U value.
Step3: The universe of discourse U is partitioned into five equal length of intervals.
Step4: Define five fuzzy sets having some linguistic values on the universe of discourse U. The linguistic values are as follows:
Table 1: Fuzzified Wheat Production(tonnes) on Variations
| Year | Wheat production(tonnes) | Variations | Fuzzified variations |
| 2001 | 69681 | - | - |
| 2002 | 72766 | 3085 | A3 |
| 2003 | 65761 | -7005 | A1 |
| 2004 | 72156 | 6395 | A4 |
| 2005 | 68637 | -3519 | A2 |
| 2006 | 69355 | 718 | A3 |
| 2007 | 75807 | 6452 | A4 |
| 2008 | 78570 | 2763 | A3 |
| 2009 | 80679 | 2109 | A3 |
| 2010 | 80804 | 125 | A3 |
| 2011 | 86874 | 6070 | A4 |
| 2012 | 94882 | 8008 | A5 |
| 2013 | 93506 | -1376 | A2 |
| 2014 | 95850 | 2344 | A3 |
| 2015 | 86527 | -9323 | A1 |
| 2016 | 87000 | 473 | A3 |
| 2017 | 98510 | 11510 | A5 |
| 2018 | 99870 | 1360 | A3 |
| 2019 | 103600 | 3730 | A4 |
| 2020 | 107860 | 4260 | A4 |
| 2021 | 109520 | 1660 | A3 |
Step6: The historical variations of the time series data are fuzzified in order to have the fuzzy logical relations obtained as follows: Variations in the fuzzy logic relationships
Step-7: The forecasted values have been obtained by using the computational algorithm. Then the forecasted output while comparing with different models given as table 2. Table 2: Forecasted Wheat Production(tonnes) by Different Models Wheat Production Prediction in India using ARIMA, Neural Network and Fuzzy Time Series
| Year | Actual Wheat Production | ARIMA | RBF | MLP | Fuzzy Time Series |
| 2001 | 69681 | 70444 | 71136 | 68012 | --- |
| 2002 | 72766 | 72919 | 70476 | 68511 | 71314 |
| 2003 | 65761 | 69333 | 68013 | 69209 | 66333 |
| 2004 | 72156 | 72296 | 67256 | 70171 | 72294 |
| 2005 | 68637 | 71247 | 67285 | 71465 | 69223 |
| 2006 | 69355 | 71458 | 69203 | 73150 | 69204 |
| 2007 | 75807 | 75788 | 76098 | 75258 | 74655 |
| 2008 | 78570 | 79249 | 78514 | 77761 | 77440 |
| 2009 | 80679 | 81977 | 79655 | 80561 | 80203 |
| 2010 | 80804 | 83173 | 82161 | 83488 | 81246 |
| 2011 | 86874 | 87615 | 86382 | 86339 | 87337 |
| 2012 | 94882 | 94466 | 90346 | 88932 | 94441 |
| 2013 | 93506 | 96192 | 93535 | 91147 | 91949 |
| 2014 | 95850 | 98445 | 91769 | 92941 | 95139 |
| 2015 | 86527 | 93432 | 85975 | 94331 | 87683 |
| 2016 | 87000 | 92056 | 87682 | 95372 | 87094 |
| 2017 | 98510 | 99045 | 98117 | 96132 | 98433 |
| 2018 | 99870 | 102568 | 100114 | 96676 | 100143 |
| 2019 | 103600 | 106355 | 104423 | 97061 | 103937 |
| 2020 | 107860 | 110576 | 108673 | 97331 | 107667 |
| 2021 | 109520 | 113292 | 108709 | 97518 | 109493 |
{"image_source":{"path":"images/51d7bcac9781f89a4d099814c099dc83ccd70e88542453ea471aa51dd46a5f96.jpg"},"content":"","chart_caption":[{"type":"text","content":"Table 2 shows that wheat production of actual values and forecasted values using different models, namely, ARIMA, RBF, MLP and fuzzy time series."},{"type":"text","content":"Figure 1 Shows that actual and forecasted values of wheat production of various models were compared by using line chart."},{"type":"text","content":"Figure 1: Actual and Forecasted Values of Wheat Production (tonnes)"}],"chart_footnote":[]}
Table 3: Residual Analysis by Different Models
| Models | MAE | MAPE | RMSE |
| ARIMA | 3453.9 | 4.188 | 4907.1 |
| RBF | 1361.19 | 1.668 | 1972.72 |
| MLP | 4033.857 | 4.5134 | 5171.001 |
| Fuzzy Time Series | 571.4 | 0.6976 | 733.569 |
{"image_source":{"path":"images/dff69b08ea5646db7df81cda1deaed59ebe104115b931ad5a00250de85373193.jpg"},"content":"","chart_caption":[{"type":"text","content":"Figure 2: Residual Analysis by Using ARIMA, RBF, MLP and Fuzzy Time Series"}],"chart_footnote":[]}
Figure 2 shows that the mean absolute error, mean absolute percentage error and root mean square error obtained by using ARIMA and neural networks for Radial Basics Function and Multilayer Perceptron and fuzzy time series for wheat production prediction. Mean absolute error, mean absolute percentage error and root mean square error of fuzzy time series is less values when compared to ARIMA and neural networks. Mean absolute error, mean absolute percentage error and root mean square error show that the performance of fuzzy time series is better than that of ARIMA and neural networks.
III. CONCLUSION
In this work, three models, namely ARIMA, neural networks and fuzzy time series were used for wheat production prediction in India. Residual analysis for mean absolute error, mean absolute percentage error and root mean square error were compared using bar charts. Mean absolute error, mean absolute percentage error and root mean square error were minimum for fuzzy time series when compared to ARIMA and neural networks. Fuzzy time series is performed better than that of ARIMA and neural networks.
Conflict of Interest
The authors declare no conflict of interest.
Ethical Approval
Not applicable
Data Availability
The datasets used in this study are openly available at [repository link] and the source code is available on GitHub at [GitHub link].
Funding
This work did not receive any external funding.
References
Cite this article
Special Issue
Launch a focused special issue to highlight research, emerging trends, and expert insights in your academic field.
