Purpose

Accurate prediction serves as a critical foundation for formulating and optimizing industrial development strategies, particularly in projecting the output of the high-tech industry through a scientific approach.

Design/methodology/approach

A novel nonlinear time-varying delay grey multivariable model (NTVDGM(1,N)) is designed in this work. First, we propose a new time-varying delay function, which can overcome the constraint of a fixed time delay in the existing grey time-delay models. On this basis, a new cumulative time-delay function combined with the Gamma function is constructed, indicating a decreasing trend and an inverted U-shaped trend. Then, the power parameter is designed to express the nonlinear relationship between the inputs and outputs, and the Grey Wolf Optimization (GWO) algorithm is selected to determine the hyperparameter.

Findings

Simulations validate the robustness and stability of the (NTVDGM(1,N)) model, demonstrating that the proposed model possesses superior forecasting accuracy.

Practical implications

The proposed model is applied to project the output of the high-tech industry in Guangdong province, the grey and non-grey prediction models are selected for comparison to the new model, and the results show that the (NTVDGM(1,N)) model achieves a satisfactory prediction accuracy in forecasting the high-tech industry's output.

Originality/value

This work proposes a novel nonlinear time-varying delay grey multivariable model with a cumulative function and variable time-delay parameters to capture the dynamic time-delay effect.

Currently, the traditional mathematical and econometric models, machine learning models, and grey models are attracting significant attention. Many econometric models have played significant roles in the field, such as distributed lag models (Pang, 2019), autoregressive moving average models, and regression models (Ávila et al., 2019). Machine learning models also have a significant involvement in the field, such as artificial neural networks (Bonyadi and Jalilian, 2022) and support vector machines (Kouziokas, 2020). However, these methods require large amounts of data, which is not always available in the real world. Thus, as a tool working with limited data, grey models are receiving more and more attention. Grey prediction models, being a portion of grey theory, possess a relatively simple structure and fewer parameters that can handle problems with insufficient samples (Deng, 1982). As the grey models are valuable, a series of improvement models have been proposed and widely applied in various filed (Wang and Zhang, 2022; Wang et al., 2022; Mao et al., 2020; Xu and Li, 2023). Multivariable grey prediction models have gained popularity in recent years due to their capacity to characterize the connection between inputs and outputs.

Many scholars have made lots of improvements from various perspectives for the grey multivariable model GM(1,N) to broaden its applicability (Zeng et al., 2019; Yan et al., 2023;Li et al., 2023a, b, c). For instance, Xie and Liu (2008) developed a discrete grey multivariable model (DGM(1,N)). Based on the DGM(1,N), Zeng et al. (2019) created a grey multivariable model with a flexible structure (NMGM(1,N)). To enhance the precision of predicting the values in the marine emerging industry, Li et al. (2023a, b, c) proposed a grey multivariable model incorporating interactive effects. Moonchai and Chutsagulprom (2020) put forward a modified grey multivariable model and used Kalman filtering to optimize the parameters. Ma et al. (2023) established a novel multivariable prediction model by integrating wavelet kernel learning with the grey model. Kiran et al. (2021) incorporated econometric methods into the grey model to propose an enhanced grey multivariable model.

Moreover, to address the issue of predicting nonlinearity characteristics, some researchers have incorporated a power exponent into the grey multivariable model to elucidate the nonlinear connection between influencing factors and outputs (Wang and Lv, 2021). For example, a discrete grey power multivariable model (DPGM(1,N)) was constructed by Ding et al. (2018), through a driver control function to identify the parameters. Zhang et al. (2022) presented a new grey multivariable method (FGM(1,N)) with power exponent and correct term, which was proven to be able to provide a satisfactory prediction accuracy in projecting energy consumption. Additionally, Ma et al. (2019) developed an enhanced Bernoulli grey model (NGBMC(1,n)) and utilized the convolution integral approach for calculating the time response function. Besides, Duman et al. (2019) used Particle Swarm Optimization to further improve the NGBMC(1,n) model's performance and applied the optimized model in electronic waste estimation.

An observed delay is a common phenomenon in economic systems. Policy effects, technological innovation, and investment require time before their effects manifest. Furthermore, the time-delay in control systems reflects the feedback delay, effect delay, and so on (Li et al., 2023a, b, c, 2024a, b). Considering the presence of a potential time-delay effect between the influencing factors and outputs, many scholars have proposed grey forecasting models and incorporated this time-delay effect. Ye et al. proposed a series of grey time-delay models (Ye et al., 2021a, b). Shi et al. (2020) combined the time delay parameters and interval grey numbers to construct a new time delay GM(1,N), and then used it to forecast air pollution. In addition, Zhou et al. (2023) designed an optimized discrete grey time-delay multivariable model (DMTGM(1,N|τ)) and utilized the traversal technique to determine the time delay parameters., and then, considering the time-varying parameters, a novel time-delay model (TVDPGM(1,N,τ)) is proposed (Zhou et al., 2024). Duan et al. (2022) developed a dynamic grey model incorporating time delay factors to investigate the interaction between energy supply and demand. Meanwhile, they adjusted the background value through the simulated annealing method. Ding et al. (2023) introduced a novel model by considering the time-delay and interaction effects simultaneously (CTGM(1,N)). Besides, Wan et al. (2022) established a mixed-frequency time-delay grey model (AMTGM(1,N)) to deal with the different frequency data. To explore the relationships between the marine and land economy-energy-environment, Li et al. (2023a, b, c) designed a grey prediction model (FTDNSGM(1,m)) involving fractional order and time-delay effect. In addition, Yan et al. (2023) built a new grey time-delay multivariable model (FTDGM(1,N,τi)) containing fractional order for predicting online public opinion. To accurately forecast the agricultural drought hazard, Zhang and Luo (2023) coupled the time-delay with the existing grey multivariable model, and put forward a new grey prediction approach with a time-delayed driving term (TDMGM(1,m,N)).

Parameter estimation plays an essential role in model optimization, and various researchers have used intelligent algorithms to find the optimal parameters. The Grey wolf optimizer (GWO) (Mirjalili et al., 2014) is s swarm intelligence optimization algorithm, which has received wide attention. It is easy to use, and many scholars have used GWO to solve their research issues. Both Rajakumar et al. (2021) and Pan et al. (2023) have demonstrated its superiority to other algorithms. The GWO algorithm has the advantages of not requiring complex parameter settings, simple implementation, and strong global search capability. Thus, in this work, we will select the GWO to search for the parameters in the proposed model, and the detailed comparison will be discussed in the empirical study.

As presented in previous studies, significant progress has been made by the developed grey time-delay models recently. However, certain limitations remain in addressing the time-delay prediction issues. For the model construction, although current grey models account for time-delay effects, they all consider the lag period to be a fixed constant and ignore the time-varying characteristics of the time-delay, which may lead to unsatisfactory prediction results. Besides, some studies have focused on developing a nonlinear framework for modeling processes, but with limited attention to their combined effect. Moreover, the existing grey projection research does not consider both the cumulative time-delay effect with varying parameters and the nonlinear relationship between factors and outputs.

To address these issues, an improved grey time-delay multivariable model is proposed, which integrates a flexible time-varying function with a cumulative time-delay effect and nonlinear characteristics. This new model can explore the varying accumulative time-delay effect and its nonlinear feature simultaneously, and extend the model's flexibility and applicability in real-world scenarios. The major innovations of this study can be stated as the following aspects:

  1. A novel time-varying delay grey multivariable model with power exponent, abbreviated as NTVDGM(1,N), is developed. The new model not only takes into account the time-delay effect with variable parameters but also incorporates its nonlinear relationship between inputs and outputs into the proposed model, which further improves the applicability of NTVDGM(1,N). Furthermore, this newly-designed model not only improves prediction accuracy but also offers a unified framework for incorporating time-delay and discrete grey models, enhancing the generalizability of the approach.

  2. The time-varying delay function is incorporated into the proposed model so that NTVDGM(1,N) can capture the dynamic time-delay. The new-design function with dynamic lag parameters in the proposed model overcomes the constraint of a fixed time delay in the existing grey time-delay models, thereby strikingly enhancing the flexibility of this new model. In addition, in light of the nonlinear connection between the inputs and outputs, the power exponent is also integrated to further improve the adaptability of NTVDGM(1,N). Overall, the new grey prediction model can simultaneously explore the variable time delay parameters and nonlinear interaction between influencing factors and outputs.

  3. The Grey Wolf Optimization (GWO) algorithm is employed to determine the time-varying lag parameters and power exponents, which further augments its prediction capability. Simulations validate the robustness and stability of NTVDGM(1,N). Consequently, this new model is validated with feasible forecasting capacity.

  4. To demonstrate the reliability of NTVDGM(1,N), this work utilizes the new model to forecast the high-tech industry's output in Guangdong province, exhibiting satisfactory performance over the compared models.

The remainder of this work is organized as follows. Section 2 presents the methodology of the proposed model. Simulations with two types of random data, including stochastic fluctuating and exponential data, are conducted to illustrate the stability of the new approach in Section 3. Subsequently, a case study on forecasting the high-tech industry's output in Guangdong province is conducted to illustrate the practicability of the new model in Section 4. Lastly, Section 5 summarizes the conclusion and future work.

Definition 1.

Xie and Liu (2008) Let X1(0)=(x1(0)(1),x1(0)(2),,x1(0)(n)) be the dependent sequence, and the sequences

represent the independent variables. Xi(1)=(xi(1)(1),xi(1)(2),,xi(1)(n))(i=1,2,,N) is first order sequence, which can be defined by the first-order accumulation generation operation (1-AGO): xi(1)(k)=m=1kxi(0)(m),k=1,2,,n. Then

(1)

is the discrete grey multivariable model (DGM(1,N)), where the parameter vector P=(β1,β2,,βN+1)T can be determined by the least square method (LSM). Notably, DGM(1,N) can describe the multivariable system with N1 independent variables. However, the independent factors have an instantaneous effect on the dependent factor, and the effects between inputs and outputs are synchronized, ignoring the lag relationship in the system. Nevertheless, the outputs may not immediately experience the effects of the inputs. Considering the delay relationship between inputs and outputs would be more reasonable to explore the time-delay effect in the system. Besides, DGM(1,N) does not consider the time-varying characteristic of the model's parameters. Particularly, the parameters might not be fixed values with the dynamic development of the system. In addition, DGM(1,N) can not explore the nonlinear relationship between independent and dependent variables. Given some data that exhibits a nonlinear characteristic, it might fail to process such a series.

Therefore, when dealing with the time series characterized by time-delay effect, accumulative effect, and the nonlinear relationship between inputs and outputs, we will construct a new grey time-delay multivariable model, considering the accumulative effect, time-varying lag parameters, and nonlinear characteristic between variables.

Definition 2.

Assuming that X1(0),Xi(0),X1(1),Xi(1) are defined as Definition 1, then the following equation

(2)

is called the nonlinear grey multivariable discrete model with time-varying delay function, denoted as NTVDGM(1,N). In this model, j=kτkfi(jτi(j))(xi(1)(jτi(j)))γi represents the driving term with an accumulative time-delay effect. τ=maxi=2N{τi}, and τi=max{τi(k)}, which τi(k) is the time-varying delay function, indicating that the independent variable Xi(0) lags behind the dynamics of the dependent variable X1(0), and k=τ+2,τ+3,,n. fi denotes the strength of the cumulative time delay effect of the ith variable on the output. In addition, γi is the power item of the ith variable, representing the nonlinear effect of ith variable on the dependent variable. βN+1k is the linear term, and βN+2 is the action grey item.

The time-varying delay function τi(k) is an essential component of NTVDGM(1,N). In this work, τi(k) is constructed as

(3)

Besides, τi0 is a lag of static time-delay, and τi0{0,1,2,,n2}. The time-varying lag function τi(k) proposed in this paper is dynamic over time and is a decreasing function of the time variable, indicating that the delay effect of the drivers on the system's output gradually decreases over time. The construction of the time-varying lag function enhances the fit of the model to the series, and the trend of the series subject to the time-delay effect can be further investigated by analyzing the properties of the time-varying delay function.

By establishing the time-varying delay function τi(k), reflecting the dynamic change process of the time delay function over time, the static time delay in the existing grey multivariate time-delay model is expanded to dynamic time-delay, and the value of the lag parameter is expanded from integer to non-integer, which expands the scope of application of the grey time-delay forecasting model. When λi=0, the proposed model decreases the static grey time-delay model. Thus, NTVDGM(1,N) enables the conversion of dynamic time delay to static time delay.

In accordance with Eq. (3), the values of the time-varying delay function τi(k) are determined by the static time delay parameters τi0 and index parameters λi. In this work, we will use the traversal algorithm (Zhou et al., 2023) to obtain the τi0, and the λi can be determined by intelligent algorithms, which will be introduced in Section 2.3.

Due to the time-varying delay function, the kτi(k) is a non-integer, thus, the value of xi(1)(kτi(k)) in Eq. (2) requires to be determined. In this paper, we will utilize the cubic spline interpolation (Li et al., 2018) to ascertain the value of xi(1)(kτi(k)). The spline function generated by cubic spline interpolation has better smoothness than linear spline interpolation, and the sequence Xi(kτi(k)) with time delay parameters is fitted by using cubic spline interpolation.

In the NTVDGM(1,N) model, fi represents the cumulative time-delay effect of independent variables. The commonly used cumulative time-delay effects can be divided into two categories:

  1. The declining type: The influence of lagged variables on the system's outputs is weighted more heavily on the trend change in the early years and tends to decrease gradually over time.

  2. Inverted U type: This type indicates that the effect of the lag factors on the dependent variable is low in the early stages, but the cumulative time-delay effect gradually increases over time, and when it peaks at a certain point, its effect gradually decreases.

The existing grey time-delay model generally constructs two functions to characterize the above two different cumulative time-delay effects. This work constructs a new type of cumulative time-delay function, which unifies the characterization of decreasing and inverted U-type cumulative time-delay, and reduces the computational complexity of the new model. fi can be expressed as follows:

(4)

where Γ(ξi)=0sξi1esds represents the Gamma function, and ξi is the shape parameter and unknown. The curves of the cumulative time-delay function are shown as in Figure 1. When ξi1, fi(t) is a monotonically decreasing function. When ξi>1, fi(t) is the single-peak function. The designed cumulative time-delay function can adaptively calculate the cumulative time delay intensity under different time-delay types with different parameters ξi, which strengthens the adaptive prediction ability of the model for complex uncertain systems.

Figure 1
A line graph showing the curves of the cumulative time-delay function with different shape parameters.A line graph showing the curves of the cumulative time-delay function with different shape parameters. The x-axis represents time (t) ranging from 0 to 20, and the y-axis represents the time-delay function ranging from 0 to 1.8. The graph includes six lines, each representing a different shape parameter (gamma) with values 0.5, 1, 2, 3, 4, and 5. The blue line represents gamma equals 0.5, the black line represents gamma equals 1, the red line represents gamma equals 2, the green line represents gamma equals 3, the magenta line represents gamma equals 4, and the cyan line represents gamma equals 5. All values are approximated.

The curves of the cumulative time-delay function with different shape parameters. Source: Author’s own creation/work

Figure 1
A line graph showing the curves of the cumulative time-delay function with different shape parameters.A line graph showing the curves of the cumulative time-delay function with different shape parameters. The x-axis represents time (t) ranging from 0 to 20, and the y-axis represents the time-delay function ranging from 0 to 1.8. The graph includes six lines, each representing a different shape parameter (gamma) with values 0.5, 1, 2, 3, 4, and 5. The blue line represents gamma equals 0.5, the black line represents gamma equals 1, the red line represents gamma equals 2, the green line represents gamma equals 3, the magenta line represents gamma equals 4, and the cyan line represents gamma equals 5. All values are approximated.

The curves of the cumulative time-delay function with different shape parameters. Source: Author’s own creation/work

Close modal
Theorem 1.

Suppose that X1(1) and Xi(1)(i=2,3,,N) are defined as in Definition 1, P=(β1,β2,,βN,βN+1,βN+2)T is the parameter vector of NTVDGM(1,N), then the structure parameters satisfy

(5)

where

Proof. Substituting the k=τ+2,τ+3,,n into Eq. (2), we can obtain the equations

(6)

Eq. (3) can be expressed as

(7)

As the matrix B is a column-full rank matrix, Eq. (7) can be presented as BTY=BTBP, and then we can get P=(BTB)1BTY.

Theorem 2.

Assuming that X1(1) and Xi(1)(i=2,3,,N) are described as Definition 1, P=(β1,β2,,βN,βN+1,βN+2)T is shown as Theorem 1. Given the initial value xˆ1(1)(τ+1)=x1(1)(τ+1), the NTVDGM(1,N) model's solution can be determined in the following equation

(8)

Proof. Theorem 2 can be proven by using the mathematical induction method. When k=τ+2, by Eq. (2), we can obtain

(9)

Eq. (9) can be expressed as

(10)

Eq. (10) is right when k=τ+2. Assuming that Eq. (5) holds true when k=τ+m+1, we get

(11)

Based on Eq. (2), when k=τ+m+2, we have

(12)

Substituting Eq. (11) into Eq. (12), it can be obtained

(13)

According to Eq. (13), we get that Eq. (8) holds true when k=τ+m+2. Thus, providing proof for the conclusion of Theorem 2. As for the first-order inverse accumulation generating operation, we have

(14)

This section focuses on determining the unknown parameters of the proposed model. NTVDGM(1,N) model includes the parameters γi,λi,τi0,ξ, and the structure parameters P=(β1,β2,,βN,βN+1,βN+2)T. Among them, the structure parameters can be obtained by Theorem 1, the parameter τi0 can be calculated by the traversal algorithm, and then the other parameters γi,λi,ξi will be solved by the GWO algorithm. This this paper, the mean absolute percentage error is used as the objective function. Based on the constraints, the objective function of NTVDGM(1,N) is proposed as

(15)

The GWO algorithm is a search method for optimization designed based on the predation behavior of grey wolves. This algorithm demonstrates notable convergence performance, requires minimal parameters, and may be easily implemented. In GWO, the social hierarchy is organized from highest to lowest using the symbols: α,β, and δ from high to low. The hunting behavior of the grey wolf primarily involves the pursuit and capture of prey by searching, encircling, and engaging in predatory attacks.

To formulate the encircling behaviors of wolves during a hunt, the following equations are presented:

where t represents the current iteration, A and C stand for coefficient vectors, Xp is the position of the prey. Adjustments are made to the positions of each search agent in accordance with potential locations “α”, “β”, and “δ” of the prey. The subsequent equations are used to determine the optimal search space.

where Xα,Xβ,Xδ are the three best positions of the best search agents and Dα,Dβ,Dδ are the distances from the search agents to the three best solutions.

To assess the predictive capabilities of NTVDGM(1,N), four critical indicators are presented in Table 1, where x1(0)(k) and xˆ1(0)(k) denote the raw data and estimated results.

Table 1

Comparison of prediction model performance

MetricDefinitionFormula
APEAbsolute percentage errorAPE(k)=|x1(0)(k)xˆ1(0)(k)x1(0)(k)|×100%
MAPEMean absolute percentage errorMAPE=1nτk=τ+1n|x1(0)(k)xˆ1(0)(k)x1(0)(k)|×100%
RMSERoot mean squared errorRMSE=1nτk=τ+1n(x1(0)(k)xˆ1(0)(k))2
MAEMean absolute errorMAE=1nτk=τ+1n|x1(0)(k)xˆ1(0)(k)|

The NTVDGM(1,N) model's modeling steps can be described as follows.

  • Step 1. Collect the raw data X1(0) and Xi(0)(i=2,3,,N), and then obtain the AGO sequences X1(1) and Xi(1) by processing the original data.

  • Step 2. Establish the NTVDGM(1,N) model based on Definition 2, based on Eq. (15), perform integer traversal of static time delay τi0, and determine the optimal parameters γi,λi,ξi by using the GWO algorithm.

  • Step 3. Substitute the optimal parameters γi,λi,ξi and τi0 into Theorem 1 for solving P=(β1,β2,,βN,βN+1,βN+2)T.

  • Step 4. Determine the fitted and predicted values by using Theorem 2.

  • Step 5. The model's prediction accuracy can be compared with the benchmark models by using the critical indicators in Table 1.

Theorem 3.

Assume that sequences Yi(0)=(yi(0)(1),yi(0)(2),...,yi(0)(n))(i=1,2,...,N) are the multiple transformations of Xi(0), where yi(0)=ρixi(0). Then, the estimated and predicted values corresponding to Y1(0) and X1(0) satisfy yˆ1(0)(k)=ρ1xˆ1(0)(k). Then, the vector PY=(β1Y,β2Y,,βNY,βN+1Y,βN+2Y)T corresponding to Yi(0) satisfies

(16)

Proof. Assume that YY=(ρ1x1(1)(τ+2),ρ1x1(1)(τ+3),,ρ1x1(1)(n))T=ρ1Y.

The matrix Q can be expressed as

Due to ρ10, Q is an invertible matrix, thus the inverse matrix of Q can be calculated:

Based on Theorem 1, we can obtain

(17)

Substituting Q1 and P=(β1,β2,,βN,βN+1,βN+2)T into Eq. (17), the parameter vector corresponding to the multiple transformation can be described as

End of proof.

Theorem 4.

Assume that the sequences Yi(0) are multiple transformations of Xi(0). Then, the predicted value of Yi(0) is satisfied

(18)

Proof. Based on Theorem 1, we can obtain that the predicted values of the sequence Yi(0) are

(19)

Taking Eq. (16) into Eq. (19), it can be calculated

Thus, we can conclude that

To demonstrate the robustness of the new model, two simulation experiments are designed with two types of random data, including stochastic fluctuating and exponential data. Based on the proposed model, the output data X1 can be obtained by X1=X2+X3+ε, where X2,X3 represents the input data. Besides, εN(0,σ2) is a random error, which is generated by the normal distribution with mean 0 and variance σ2=0.01,0.8,1.8. For the experimental simulation, we randomly generate 20 datasets with the characteristics of random fluctuations and exponential laws. The first 18 observations are used to establish the prediction models, and the rest 2 samples for testing the forecasting accuracy. In addition, to illustrate the simulation results, the experiments are repeated 200 times. Figure 2 shows the boxplots of the fitted and predicted errors in 200 times with two prediction models (AMTGM (Wan et al., 2022; TDAGMLuo et al., 2021), and the average MAPE values are tabulated in Tables 2 and 3.

Figure 2
Boxplots of MAPE values for different variances.The image contains six sets of boxplots, each set showing the Mean Absolute Percentage Error (MAPE) values for three different models: NTVDGM, AMTGM, and TDAGM. The boxplots are divided into two columns: the left column shows the fitted MAPE values, while the right column shows the predicted MAPE values. Each row corresponds to a different variance level: 0.01, 0.8, and 1.8. The boxplots illustrate the distribution of MAPE values across 200 experiments for each model and variance level. The models are compared based on their performance in fitting and predicting data with varying levels of variance. All values are approximated.

Boxplots for the MAPE values in fitted and predicted periods generated by 200 experiments with different variances σ2=0.01,0.8,1.8. Source: Author’s own creation/work

Figure 2
Boxplots of MAPE values for different variances.The image contains six sets of boxplots, each set showing the Mean Absolute Percentage Error (MAPE) values for three different models: NTVDGM, AMTGM, and TDAGM. The boxplots are divided into two columns: the left column shows the fitted MAPE values, while the right column shows the predicted MAPE values. Each row corresponds to a different variance level: 0.01, 0.8, and 1.8. The boxplots illustrate the distribution of MAPE values across 200 experiments for each model and variance level. The models are compared based on their performance in fitting and predicting data with varying levels of variance. All values are approximated.

Boxplots for the MAPE values in fitted and predicted periods generated by 200 experiments with different variances σ2=0.01,0.8,1.8. Source: Author’s own creation/work

Close modal
Table 2

The average MAPE values are generated by exponential data with 200 experiments

Variance0.010.81.8
MAPE-fittedMAPE-predMAPE-fittedMAPE-predMAPE-fittedMAPE-pred
NTVDGM0.7240.2791.7050.8760.5892.114
AMTGM1.4121.5104.6492.5365.2583.275
TDAGM3.4113.3577.2124.6288.5387.446

Note(s): The optimum value of the difference variance for six predictions is in italic

Table 3

The average MAPE values are generated by random fluctuation data with 200 experiments

Variance0.010.81.8
MAPE-fittedMAPE-predMAPE-fittedMAPE-predMAPE-fittedMAPE-pred
NTVDGM3.7174.1598.33310.1347.69910.432
AMTGM1.6524.3119.83714.8629.54115.227
TDAGM7.6299.46511.02615.9389.54020.939

Note(s): The optimum value of the difference variance for six predictions is in italic

As noted in Figure 2(a) and Table 2, for the exponential data, the MAPE values of NTVDGM(1,N) and are superior to the compared models with the difference variance. In the predicted stage, the results of the three models stay within 5%, and when the variance increases to 1.8, the MAPE values generated by NTVDGM(1,N), AMTGM, and TDAGM remain in the range of [3.40%,3.43%], [2.85%,13.42%], and [8.41%,8.59%], indicating that the proposed model still maintains a stable prediction accuracy when the variance increases. Overall, as the variance increases, the fitting and prediction accuracy of NTVDGM(1,N) demonstrates significant superiority over the two comparative models. This finding suggests that when the data follow a quasi-exponential distribution pattern, the proposed model in this study exhibits enhanced fitting performance, improved predictive capability, and greater stability compared to the other two grey time-delay models.

Moreover, from the perspective of the random fluctuation data, as depicted in Figure 2(b) and Table 3, the predicted values generated by NTVDGM(1,N) are predominant among the projection models, indicating that the proposed model yields remarkable forecasting ability with the fluctuation data. As the variance increases, NTVDGM(1,N) maintains stable fitting accuracies of approximately 3.7%, 8.2%, and 7.6%, while its prediction accuracies stabilize at around 4.1%, 10.1%, and 10.4%, respectively. In contrast, the two comparative models exhibit significantly lower fitting and prediction accuracies under fluctuating data conditions. Furthermore, their error distributions are broader across all three variance levels, particularly at a variance of 0.01, where the dispersion is most pronounced. These results indicate that the computational stability of the comparative models is inferior to that of the proposed model. Consequently, the robustness of the proposed model is proven with the different data features.

The NTVDGM(1,N) model is implemented to forecast the output of the high-tech industry in Guangdong province. In order to evaluate its performance, the model is compared with several other existing models. To assess model performance, we conduct a comparative analysis with several established benchmark models. This evaluation is particularly relevant given China's 15th Five-Year Plan, which emphasizes the strategic importance of high-tech industries as a cornerstone of national economic development. Accurate forecasting of high-tech industry trends becomes crucial, as it provides policymakers with evidence-based insights for strategic decision-making. Consistent with economic theory, industrial output demonstrates direct dependence on input factors. Empirical studies consistently identify Research and Development (R&D) intensity as a primary determinant of high-tech industry performance. Accordingly, our analysis incorporates R&D expenditure and R&D personnel as key input variables, while accounting for time-delay effects in the input-output relationship.

To assess the practicability of NTVDGM(1,N), five existing prediction models are chosen as the compared models for the performance comparison. It is worth noting that the grey models consist of two grey time-delay models (TDAGM(1,N,t); AMTGM(1,N)), and one grey power model (GM(1,N) power model), the non-grey models include distributed-lag model (DLM) and Back Propagation Neural Network (BPNN). The experiment utilizes the data from 2006 to 2020 to construct the prediction models, while data from 2021 to 2022 are employed for projection. The errors generated by NTVDGM(1,N) and the above compared models are measured by four error indices. The results are displayed in Table 4 and Figure 3.

Table 4

Simulation and prediction results of the proposed model and the compared models

YearActualNTVDGM(1,N)TDAGM(1,N,t)AMTGM(1,N)GM(1,N) powerDLMBPNN
valueFitted valueAPE (%)Fitted valueAPE (%)Fitted valueAPE (%)Fitted valueAPE (%)Fitted valueAPE (%)Fitted valueAPE (%)
200612794.1312794.130.0012794.130.0010987.4914.12
200714582.913584.366.8515069.763.3410853.0825.58
200816070.816172.710.6316070.800.0016497.032.6528912.4079.9114302.1211.01
200916758.0516758.050.0018497.8410.3816314.992.6418103.918.0332124.6191.7015640.296.67
201020952.820933.100.0920556.221.8921316.551.7420045.564.3334470.5364.5218377.2612.29
201123227.623088.360.6022936.831.2523476.341.0722352.643.7735063.8050.9621693.886.60
201225046.626287.114.9525252.250.8225426.531.5225037.890.0338165.0952.3824747.531.19
201327871.127419.711.6227875.660.0227574.481.0627953.640.3041905.4850.3525229.099.48
201430328.930391.440.2130632.441.0030308.370.0730844.441.7047222.5055.7026905.3711.29
201533308.133048.710.7833856.881.6533016.130.8834056.852.2547869.5243.7230857.487.36
201637765.236289.683.9137454.460.8237018.811.9837447.360.8452481.2438.9736156.644.26
201742,25641144.272.6341056.672.8442521.510.6340613.693.8958095.5537.4840949.553.09
201846,74746036.481.5244512.434.7846530.150.4644184.555.4859814.3127.9544191.885.47
201946,72350929.159.0047877.372.4748716.574.2747606.511.8953697.3914.9345899.991.76
202050,18550902.971.4351839.493.3048889.052.5852180.193.9860285.4720.1357539.8014.66
MAPE (%)2.432.581.573.0348.368.99
RMSE1475.831010.28765.751118.8613528.372839.34
MAE944.97742.91545.19853.2013297.522333.02
202153,91554674.681.4157220.706.1349223.028.7057888.177.3756593.734.9768053.2726.22
202256,28558738.374.3663221.2112.3250798.459.7564290.7114.2265273.8915.9769183.3722.92
MAPE (%)2.889.239.2310.8010.4724.57
RMSE1816.065433.175104.756319.716632.3413532.53
MAE1606.535120.965089.265989.445833.8113518.32
Figure 3
A box-and-whisker plot comparing the errors of APE values generated by six prediction models.A box-and-whisker plot comparing the errors of APE values generated by six prediction models. The plot features six vertical box plots labeled as NTVDGM, TDAGM, AMTGM, GM(1,N) power, DLM, and BPNN. The x-axis represents the prediction models, while the y-axis represents the APE percentage ranging from 0 to 90. Each box plot shows the distribution of APE values for the respective model. The NTVDGM, TDAGM, AMTGM, and GM(1,N) power models have lower APE values with minimal spread, indicating better performance. The DLM model shows a significantly higher median and wider spread, suggesting higher errors. The BPNN model has a moderate spread with a higher median compared to the first four models. Outliers are present in all models, with DLM having the most. All values are approximated.

The errors of APE values generated by the six prediction models. Source: Author’s own creation/work

Figure 3
A box-and-whisker plot comparing the errors of APE values generated by six prediction models.A box-and-whisker plot comparing the errors of APE values generated by six prediction models. The plot features six vertical box plots labeled as NTVDGM, TDAGM, AMTGM, GM(1,N) power, DLM, and BPNN. The x-axis represents the prediction models, while the y-axis represents the APE percentage ranging from 0 to 90. Each box plot shows the distribution of APE values for the respective model. The NTVDGM, TDAGM, AMTGM, and GM(1,N) power models have lower APE values with minimal spread, indicating better performance. The DLM model shows a significantly higher median and wider spread, suggesting higher errors. The BPNN model has a moderate spread with a higher median compared to the first four models. Outliers are present in all models, with DLM having the most. All values are approximated.

The errors of APE values generated by the six prediction models. Source: Author’s own creation/work

Close modal

As shown in Table 4, in the training stage, the MAPE values of NTVDGM(1,N), TDAGM(1,N,t), AMTGM(1,N), and GM(1,N) power models are approximately 3%, indicating that NTVDGM(1,N) and the three benchmark grey models achieve satisfactory computational accuracy during the fitted stage. Among them, NTVDGM(1,N) demonstrates the best fitting performance, with MAPE, RMSE, and MAE values of 1.57%, 765.75, and 545.19, respectively. However, the fitting accuracy of NTVDGM(1,N) is slightly inferior to that of AMTGM(1,N). In the testing stage, the MAPE, RMSE, and MAE values are 2.88%, 1816.06, and 1606.53, respectively, which are the lowest values among the six forecasting models, demonstrating superior predictive accuracy. Overall, NTVDGM(1,N), TDAGM(1,N,t), AMTGM(1,N), and GM(1,N) power models demonstrate close agreement with the original series during the fitting phase. However, in the prediction stage, all three comparative grey models exhibit varying degrees of deviation, while NTVDGM(1,N) maintains the closest proximity to the original series, indicating their superior modeling performance.

Compared with the five benchmark models, during the training stage, AMTGM(1,N) achieves the lowest MAPE (1.57%) among all six forecasting models. In contrast, DLM shows the poorest performance with maximum values of MAPE (48.36%), RMSE (13528.37), and MAE (13,297.52) among all prediction models, indicating its limited adaptability to raw data and constrained predictive capability for complex nonlinear systems. The three grey prediction models demonstrate superior fitting accuracy compared to the two non-grey models, suggesting that grey forecasting methods are more suitable for predicting operating income in Guangdong's high-tech industry. In addition, during the testing stage, AMTGM(1,N), TDAGM(1,N,t), and the GM(1,N) power models all exhibit prediction errors exceeding 8% for both 2021 and 2022 years, showing significant deviations from NTVDGM(1,N). While the non-grey compared models performed reasonably well in the fitting stage, their predictive performance deteriorated markedly during the testing stage, with MAPE values reaching 24.57%.

Additionally, as evidenced by Table 2 and Figure 3, NTVDGM(1,N) exhibits the most stable APE distribution with minimal fluctuations. The APE values of NTVDGM(1,N) range within [0.09%, 9.00%]. In comparison, the TDAGM(1,N,t), AMTGM(1,N), and GM(1,N) power models, demonstrate APE distributions of [0.02%, 12.32%], [0.07%, 9.75%], and [0.03%, 14.22%] respectively, while models and show wider ranges of [4.97%, 91.70%] and [1.19%, 26.22%]. Comparative analysis of APE values across all six forecasting models indicates that NTVDGM(1,N) maintains the highest computational stability.

In summary, the model incorporating both time-varying characteristics of the delay parameter and the cumulative time-delay effect demonstrates optimal modeling performance, accurately capturing the development trend of operating income in Guangdong's high-tech industry.

In light of the system's inherent time-varying delay and nonlinearity, this paper develops an advanced nonlinear grey time-delay multivariable model with time-varying lag parameters and accumulative delay effect (NTVDGM(1,N)). What's more, the properties of NTVDGM(1,N) are discussed in the multiple transformations, indicating that the multiple transformations do not influence the modeling performance.

  1. The NTVDGM(1,N) model can significantly explore the dynamic time-delay, different delay effects, and fluctuation features within the system. The newly developed time delay function has the capability to accurately represent both static and dynamic time delays over time, which is more flexible than the existing grey time-delay model with fixed time delay parameters. Moreover, a new dynamic time-delay function is constructed, which adaptively adjusts its parameters according to different observed lag effects. Besides, the power exponent has been integrated into NTVDGM(1,N) enhancing the practicality of addressing the complexity associated with the time series' multiple characteristics. Furthermore, the variable lag parameters, shape parameter, and power exponent are calculated using the GWO algorithm, resulting in further improvement in the forecasting accuracy of the newly-designed approach.

  2. The predominant performance derives from the newly-designed model, which takes into account the cumulative time-delay effect and time-varying characteristics of the time delay parameters simultaneously. The time delay function with time-varying parameters can explore the dynamic time-delay effect, which can broaden the model's application. Additionally, the power exponent is incorporated into NTVDGM(1,N) for capturing the nonlinear influence of the input factors.

  3. The simulation experiments are used to analyze the robustness and stability of NTVDGM(1,N). The results generated by simulation experiments highlight the excellent reliability of the proposed model.

  4. The empirical analysis of high-tech industry's output in Guangdong Province demonstrates that NTVDGM(1,N) incorporates dynamic variations in time-lag parameters and system nonlinearity, outperforms both grey and non-grey comparative models in terms of prediction accuracy and stability. These results indicate that the proposed NTVDGM(1,N) is more suitable for forecasting analysis of high-tech industries under time-delay effects.

The NTVDGM(1,N) model can provide satisfactory prediction performance. We will optimize this model to further broaden its application in the future. Further discussions on the complex interactions in the system would be considered in the prediction model, such as the interaction effect between inputs and outputs. The time-delay function will be optimized with a simple form to reduce the computational complexity of the model and improve the stability of the optimal values of the parameters.

Ávila
,
M.M.
,
Durán
,
M.L.
,
Caballero
,
D.
,
Antequera
,
T.
,
Palacios-Pérez
,
T.
,
Cernadas
,
E.
and
Fernández-Delgado
,
M.
(
2019
), “
Magnetic Resonance Imaging, texture analysis and regression techniques to non-destructively predict the quality characteristics of meat pieces
”,
Engineering Applications of Artificial Intelligence
, Vol. 
82
, pp. 
110
-
125
, doi: .
Bonyadi
,
N.A.
and
Jalilian
,
A.S.H.
(
2022
), “
The role of dynamic capability on a business model in knowledge-intensive firms
”,
Kybernetes
, Vol. 
51
No. 
12
, pp. 
3449
-
3469
, doi: .
Deng
,
J.
(
1982
), “
Control problems of grey systems
”,
System and Control Letters
, Vol. 
5
, pp. 
288
-
294
.
Ding
,
S.
,
Dang
,
Y.
and
Xu
,
N.
(
2018
), “
Construction and optimization of a multi-variables discrete grey power model
”,
Systems Engineering and Electronics
, Vol. 
40
No. 
6
, pp. 
1302
-
1309
.
Ding
,
Q.
,
Xiao
,
X.
and
Kong
,
D.
(
2023
), “
Estimating energy-related CO2 emissions using a novel multivariable fuzzy grey model with time-delay and interaction effect characteristics
”,
Energy
, Vol. 
263
, 126005, doi: .
Duan
,
H.
,
Liu
,
Y.
and
Wang
,
G.
(
2022
), “
A novel dynamic time-delay grey model of energy prices and its application in crude oil price forecasting
”,
Energy
, Vol. 
251
, 123968, doi: .
Duman
,
G.M.
,
Kongar
,
E.
and
Gupta
,
S.M.
(
2019
), “
Estimation of electronic waste using optimized multivariate grey models
”,
Waste Management
, Vol. 
95
, pp. 
241
-
249
, doi: .
Kiran
,
M.
,
Shanmugam
,
P.V.
,
Mishra
,
A.
,
Mehendale
,
A.
and
Nadheera Sherin
,
H.R.
(
2021
), “
A multivariate discrete grey model for estimating the waste from mobile phones, televisions, and personal computers in India
”,
Journal of Cleaner Production
, Vol. 
293
, 126185, doi: .
Kouziokas
,
G.N.
(
2020
), “
SVM kernel based on particle swarm optimized vector and Bayesian optimized SVM in atmospheric particulate matter forecasting
”,
Applied Soft Computing
, Vol. 
93
, 106410, doi: .
Li
,
Q.
,
Wang
,
N.
and
Yi
,
D.
(
2018
),
Numerical Analysis
, (5th ed.) ,
Tsinghua University Press
,
Beijing
.
Li
,
H.
,
Wang
,
S.
,
Shi
,
H.
,
Su
,
C.
and
Li
,
P.
(
2023a
), “
Two-Dimensional iterative learning robust asynchronous switching predictive control for multiphase batch processes with time-varying delays
”,
IEEE Transactions on Systems, Man, and Cybernetics: Systems
, Vol. 
53
No. 
10
, pp. 
6488
-
6502
, doi: .
Li
,
X.
,
Wu
,
X.
and
Zhao
,
Y.
(
2023b
), “
Research and application of multi-variable grey optimization model with interactive effects in marine emerging industries prediction
”,
Technological Forecasting and Social Change
, Vol. 
187
, 122203, doi: .
Li
,
X.
,
Zhou
,
S.
,
Zhao
,
Y.
and
Yang
,
B.
(
2023c
), “
Marine and land economy–energy–environment systems forecasting by novel structural-adaptive fractional time-delay nonlinear systematic grey model
”,
Engineering Applications of Artificial Intelligence
, Vol. 
126
, 106777, doi: .
Li
,
H.
,
Wang
,
S.
,
Shi
,
H.
,
Wang
,
L.
,
Su
,
C.
and
Li
,
P.
(
2024a
), “
Robust asynchronous fuzzy predictive fault-tolerant tracking control for nonlinear multi-phase batch processes with time-varying reference trajectories
”,
Engineering Applications of Artificial Intelligence
, Vol. 
133
, 108415, doi: .
Li
,
X.
,
Yang
,
J.
,
Zhao
,
Y.
,
Zhou
,
S.
and
Wu
,
Y.
(
2024b
), “
Prediction and assessment of marine fisheries carbon sink in China based on a novel nonlinear grey Bernoulli model with multiple optimizations
”,
The Science of the Total Environment
, Vol. 
914
, pp. 
914169769
-
169769
, doi: .
Luo
,
D.
,
An
,
Y.
and
Wang
,
X.
(
2021
), “
Time-delayed accumulative TDAGM (1,N,t) model and its application in grain production
”,
Control and Decision
, Vol. 
36
No. 
8
, pp. 
2002
-
2012
.
Ma
,
X.
,
Liu
,
Z.
and
Wang
,
Y.
(
2019
), “
Application of a novel nonlinear multivariate grey Bernoulli model to predict the tourist income of China
”,
Journal of Computational and Applied Mathematics
, Vol. 
347
, pp. 
84
-
94
, doi: .
Ma
,
X.
,
Lu
,
H.
,
Ma
,
M.
,
Wu
,
L.
and
Cai
,
Y.
(
2023
), “
Urban natural gas consumption forecasting by novel wavelet-kernelized grey system model
”,
Engineering Applications of Artificial Intelligence
, Vol. 
119
, 105773, doi: .
Mao
,
S.
,
Kang
,
Y.
,
Zhang
,
Y.
,
Xiao
,
X.
and
Zhu
,
H.
(
2020
), “
Fractional grey model based on non-singular exponential kernel and its application in the prediction of electronic waste precious metal content
”,
ISA Transactions
, Vol. 
107
, pp. 
12
-
26
, doi: .
Mirjalili
,
S.
,
Mirjalili
,
S.M.
and
Lewis
,
A.
(
2014
), “
Grey wolf optimizer
”,
Advances in Engineering Software
, Vol. 
69
, pp. 
46
-
61
, doi: .
Moonchai
,
S.
and
Chutsagulprom
,
N.
(
2020
), “
Short-term forecasting of renewable energy consumption: augmentation of a modified grey model with a Kalman filter
”,
Applied Soft Computing
, Vol. 
87
, 105994, doi: .
Pan
,
H.
,
Chen
,
S.
and
Xiong
,
H.
(
2023
), “
A high-dimensional feature selection method based on modified Gray Wolf Optimization
”,
Applied Soft Computing
, Vol. 
135
, 110031, doi: .
Pang
,
H.
(
2019
),
Econometrics
, (4rd ed.) ,
Science Press
,
Beijing
.
Rajakumar
,
R.
,
Sekaran
,
K.
,
Hsu
,
C.-H.
and
Kadry
,
S.
(
2021
), “
Accelerated grey wolf optimization for global optimization problems
”,
Technological Forecasting and Social Change
, Vol. 
169
, 120824, doi: .
Shi
,
J.
,
Xiong
,
P.
,
Yang
,
Y.
and
Quan
,
B.
(
2020
), “
Forecasting smog in Beijing using a novel time-lag GM(1,N) model based on interval grey number sequences
”,
Grey Systems: Theory and Application
, Vol. 
ahead-of-print
No. 
ahead-of-print
, pp. 
754
-
778
, doi: .
Wan
,
G.
,
Li
,
X.
,
Yin
,
K.
,
Zhao
,
Y.
and
Yang
,
B.
(
2022
), “
Application of a novel time-delay grey model based on mixed-frequency data to forecast the energy consumption in China
”,
Energy Reports
, Vol. 
8
, pp. 
4776
-
4786
, doi: .
Wang
,
Z.
and
Lv
,
Y.
(
2021
), “
A non-linear systematic grey model for forecasting the industrial economy-energy-environment system
”,
Technological Forecasting and Social Change
, Vol. 
167
, 120707, doi: .
Wang
,
H.
and
Zhang
,
Z.
(
2022
), “
Forecasting Chinese provincial carbon emissions using a novel grey prediction model considering spatial correlation
”,
Expert Systems with Applications
, Vol. 
209
, 118261, doi: .
Wang
,
Y.
,
Zhang
,
Y.
,
Nie
,
R.
,
Chi
,
P.
,
He
,
X.
and
Zhang
,
L.
(
2022
), “
A novel fractional grey forecasting model with variable weighted buffer operator and its application in forecasting China's crude oil consumption
”,
Petroleum
, Vol. 
8
No. 
2
, pp. 
139
-
157
, doi: .
Xie
,
N.
and
Liu
,
S.
(
2008
), “
Research on the discrete grey model of multivariables and its properties
”,
Systems Engineering - Theory and Practice
, Vol. 
6
, pp. 
143
-
150
, doi: .
Xu
,
H.
and
Li
,
N.
(
2023
), “
Forecasting China's marine scientific research and education, marine industrial structure upgrading and marine economy growth based on the AWBO-MGM (1,m) model
”,
Marine Economics and Management
, Vol. 
6
No. 
1
, pp. 
1
-
22
, doi: .
Yan
,
S.
,
Su
,
Q.
,
Gong
,
Z.
,
Zeng
,
X.
and
Herrera Viedma
,
E.
(
2023
), “
Online public opinion prediction based on rolling fractional grey model with new information priority
”,
Information Fusion
, Vol. 
91
, pp. 
277
-
298
, doi: .
Ye
,
L.
,
Xie
,
N.
and
Hu
,
A.
(
2021a
), “
A novel time-delay multivariate grey model for impact analysis of CO2 emissions from China's transportation sectors
”,
Applied Mathematical Modelling
, Vol. 
91
, pp. 
493
-
507
, doi: .
Ye
,
L.
,
Xie
,
N.
and
Luo
,
D.
(
2021b
), “
Construction of accumulative time-delay nonlinear ATNDGM(1, N) model and its application
”,
Systems Engineering - Theory and Practice
, Vol. 
41
No. 
9
, pp. 
2414
-
2427
, doi: .
Zeng
,
B.
,
Duan
,
H.
and
Zhou
,
Y.
(
2019
), “
A new multivariable grey prediction model with structure compatibility
”,
Applied Mathematical Modelling
, Vol. 
75
, pp. 
385
-
397
, doi: .
Zhang
,
D.
and
Luo
,
D.
(
2023
), “
Research on the prediction model of agricultural drought hazard considering the time-delayed cumulative effect and system development characteristics
”,
Science of The Total Environment
, Vol. 
882
, 163523, doi: .
Zhang
,
M.
,
Guo
,
H.
,
Sun
,
M.
,
Liu
,
S.
and
Forrest
,
J.
(
2022
), “
A novel flexible grey multivariable model and its application in forecasting energy consumption in China
”,
Energy
, Vol. 
239
, 122441, doi: .
Zhou
,
H.
,
Dang
,
Y.
,
Yang
,
D.
,
Wang
,
J.
and
Yang
,
Y.
(
2023
), “
An improved grey multivariable time-delay prediction model with application to the value of high-tech industry
”,
Expert Systems with Applications
, Vol. 
213
, 119061, doi: .
Zhou
,
H.
,
Yang
,
Y.
and
Geng
,
S.
(
2024
), “
Forecasting the output of high-tech industry in China: a novel nonlinear grey time-delay multivariable model with variable lag parameters
”,
Expert Systems with Applications
, Vol. 
257
, 125054, doi: .
Arora
,
S.
and
Taylor
,
J.W.
(
2019
), “
Rule-based autoregressive moving average models for forecasting load on special days: a case study for France
”,
European Journal of Operational Research
, Vol. 
266
No. 
1
, pp. 
259
-
268
, doi: .
Asokan
,
A.
and
Anitha
,
J.
(
2020
), “
Adaptive Cuckoo Search based optimal bilateral filtering for denoising of satellite images
”,
ISA Transactions
, Vol. 
100
, pp. 
308
-
321
, doi: .
Chan
,
J.C.C.
(
2021
), “
Minnesota-type adaptive hierarchical priors for large Bayesian VARs
”,
International Journal of Forecasting
, Vol. 
37
No. 
3
, pp. 
1212
-
1226
, doi: .
De Araújo Morais
,
L.R.
and
Da Silva Gomes
,
G.S.
(
2022
), “
Forecasting daily Covid-19 cases in the world with a hybrid ARIMA and neural network model
”,
Applied Soft Computing
, Vol. 
126
, 109315, doi: .
Published in Marine Economics and Management. Published by Emerald Publishing Limited. This article is published under the Creative Commons Attribution (CC BY 4.0) licence. Anyone may reproduce, distribute, translate and create derivative works of this article (for both commercial and non-commercial purposes), subject to full attribution to the original publication and authors. The full terms of this licence may be seen at Link to the terms of the CC BY 4.0 licence.

or Create an Account

Close Modal
Close Modal