ORIGINAL RESEARCH article

Front. Environ. Sci., 10 November 2022

Sec. Environmental Informatics and Remote Sensing

Volume 10 - 2022 | https://doi.org/10.3389/fenvs.2022.1014856

A quantitative model based on grey theory for sea surface temperature prediction

  • 1. School of Physics and Electronic Technology, Liaoning Normal University, Dalian, China

  • 2. School of Geography, Liaoning Normal University, Dalian, China

  • 3. Institute of Geographic Sciences and Natural Resources Research, Chinese Academy of Sciences, Beijing, China

  • 4. School of Foreign Languages, Liaoning Normal University, Dalian, China

Abstract

In order to predict sea surface temperature (SST), combined with the genetic algorithm and the least-squares method, a GM(1,1|sin) power model prediction method based on similarity deviation is proposed. We first combined the data of two consecutive years into a new time series, analyzed the similarity of the data of the previous year, and obtained the most similar year and the corresponding new time series. Then, we established a GM(1,1|sin) power model to predict SST. In model validation, we predicted the monthly average SST from 2016 to 2020 with the data from 1985 to 2015, 2016, 2017, 2018, and 2019. The validation results showed that the maximum mean relative error (MRE) was 13.28%, the minimum MRE was 5.54%, and the average MRE and the root mean square error (RMSE) were 9.81% and 1.0627, respectively. All of evaluation metrics of Lin’s concordance correlation coefficient (LCCC) and the ratio of performance to deviation (RPD) were excellent. We iteratively predicted the monthly average SST from 2016 to 2020 with the data from 1985 to 2015, the maximum MRE was 13.91%, the minimum was 7.80%, and the average MRE, RMSE, LCCC and RPD are 11.07% 1.0603, 0.9894, and 7.497, respectively. Compared with GM(1,1), GM(1,1|sin + cos), and GM(1,1|sin) models, the proposed model outperformed these models with at least 50% in the MRE. It proves that the proposed model can be regarded as a better solution to predicting SST.

1 Introduction

SST prediction is closely related to the daily life of human beings, and understanding the changes of SST in advance plays an important leading role in the school of fish swimming, deep sea exploration, cold wave warning, and even military defense ().SST is one of the most important parameters in the study of global oceanic and atmospheric interactions. The prediction of SST is to predict the sea temperature field, especially the SST changes with time. Accurate prediction of SST can provide effective support for coping with marine disasters such as storms and typhoons and prevent red tide ().

Today, more than a dozen countries have made outstanding contributions to sea water temperature prediction services. Among them, the United States, the former Soviet Union, and Japan publicly provide the largest number of sea water temperature analysis and prediction products (). Sea water temperature prediction in China began in the early 1960s. At first, Shandong Ocean University, Shandong Marine and Fishery Research Institute, and Yantai Meteorological Institute cooperated to explore the single-station water temperature prediction method in the offshore area of Yantai. With the need of the development of marine economy, the daily average SST prediction of a single station in coastal cities and the water temperature prediction of coastal bathing beaches have been successively carried out ().

There have been many studies on the prediction of sea water temperature in recent years. reconstructed the time series in phase space and used a fuzzy neural network. The relative error of the predicted value was controlled within 10%, and the fitting correlation coefficient was 0.98. using a threshold autoregressive model selected sea surface temperature data from 1963 to May 1994 in Dalian’s Tiger Beach. The number of threshold intervals and the search range of threshold values were found by global optimization combined with a genetic algorithm. The predicted value in May 1995 was 0.58°C which was different from the measured value. selected CTD data of the East China Sea and the non-stationary time series were stabilized by the EMD method, and the correlation coefficient reached 0.94. selected remote-sensing data from the AVHRR satellite and adopted the periodic trend decomposition method of locally weighted regression, combined with the neural network, and the root mean square error reached 0.79°C. proposed a HWT prediction method based on a recursive neural network (RNN). The correlation coefficient range of prediction was from 0.9936 to 0.959, and the root mean square error range was from 0.5076°C to 1.3238°C. established a multi-variable artificial neural network model. The RMSE achieved based on the training results of SST was 0.348°C. designed a recursive unit (GRU) neural network algorithm on the basis of gating for medium and long-term SST prediction, and its average absolute error was within the range of 0–2.5°C. using EEMD obtained the eigenmode function, which solved the problem of the high signal-to-noise ratio of the results of the EMD algorithm and further improved the prediction accuracy, and the agreement degree between the predicted value and the measured value reached 99.61%. used the multi-scale fusion method to predict the daily mean temperature of sea water with a root mean square error of 0.996°C and also made a prediction on an hourly scale with a root mean square error of 1.06°C. used the deep neural network based on long and short memory to achieve a root mean square error of 0.5°C in 1 month and 0.66°C in 12 months. used the CMIP5 model to predict the next 100 years, and the results showed that SST would increase significantly by 2100: SST would increase by about 1.55°C per decade, while seasonal SST would increase by 1.03–1.95°C. used the CMIP6 model to calculate the temperature around the Korean Peninsula which will increase from 0.49°C to 0.59°C every 10 years.

How to improve the accuracy of grey prediction theory in the oscillation sequence has become a topic for mathematicians. The research achievements that have made breakthroughs are mainly divided into two aspects in recent years: on the one hand, the processing of the original sequence is improved. carried out translation transformation and geometric average transformation operation on the original data sequence. proposed the grey interval GM(1,1) model by the upper bound sequence and the lower bound sequence of the original sequence was taken. used a new-structure grey Verhulst model for predicting China’s tight gas production and the comprehensive error was 2.07%. carried out accelerated translation transformation and weighted mean generation transformation of the original sequence. proposed to carry out accelerated exponential transformation and geometric average generation transformation of the original sequence. On the other hand, the bleaching equation in grey prediction theory is improved, used a fractional discrete GM(1,1) power model based on the GM(1,1) power model. carried out the power function optimization of the GM(1,1) power model, and five derived models are proposed, including the oscillating GM(1,1) power model with time-varying parameters and considering the system delay. The residual error is also predicted by using the Fourier series, which improves the grey prediction theory to certain extent. introduced a new action quantity k2d and r-order was introduced into the traditional three-parameter discrete grey forecasting model. The results show that the comprehensive mean relative percentage error of the new model was 0.4765%. In this study, the GM(1,1|sin) power model is selected to estimate the SST.

To minimize the impact of the cold snap on the SST prediction, we used data from 1985 to 2020 to predict the SST for the next 50 years. However, these 35 years of data are not sufficient to predict temperature trends over the next 50 years. In view of the current situation, we designed a grey prediction model to obtain more reliable data to successfully overcome the problem resulting from insufficient data. Grey prediction theory is a kind of the dynamic model which uses discrete data to establish a differential equation based on the concepts of correlation space, smooth discrete functions, and so on. The equation is named as the grey model (GM) that generates discrete random numbers into numbers whose randomness is significantly weakened and more regular, so that it is convenient to study and describe the process of its change. The GM has a strict theoretical foundation and advantage of practicality. Therefore, the results of the grey prediction model are relatively stable, which is not only applicable to the prediction of large data amount, but also accurate when the data amount is small ().The GM is a powerful method for the problems characterized by samples with uncertainty. By identifying different degrees of development trends among system factors, the GM generates strong regularity of data sequences and then establishes the corresponding differential equation model to predict the future trend. The GM holds that the behaviors of systems are in order for the purpose of the implementation of a certain function, although they seem hazy and complex (). The traditional GM(1,1) prediction curve is approximate to a straight line, so it can only be predicted for some monotonic increasing or decreasing sequences. The prediction of the sea water temperature with strong fluctuation of vibration is not suitable for GM(1,1) (). established a GM(1,1|sin) power model based on the GM(1,1|sin) model to solve the compound oscillation sequence with different periods.

The purpose of this study was to use the experience predict method to establish a prediction model of SST which can be easily implemented and applied, a GM(1,1|sin) power model prediction method based on similarity deviation was proposed. We first combined the data of two consecutive years into a new time series, analyzed the similarity of the data of the previous year, and obtained the most similar year and the corresponding new time series. Then, we established a GM(1,1|sin) power model to predict SST. Based on the MATLAB simulation, this method used data from 1985 to 2015, a total of 372 monthly average SST to predict data of 2016–2020, and compared with the measured data, then predicted the SST of the 50 years after 2020, and drew conclusions.

2. Materials and methods

2.1 Data sources

The SST series is a kind of a time series. A time series refers to the sequence formed by arranging the values of a variable at different times in time sequence, and its time scale can be a day, month, year, hour, etc. The time series model is a mathematical model established by using the time series. It is mainly used for short-term prediction of the future and belongs to the trend predicting method. In reality, the vast majority of phenomena are rapidly changing. With the passage of time, the internal and external influencing factors change greatly, which reduces the prediction accuracy gradually. The method of time series analysis and prediction predicts the future according to the development trends and change rules of past and present, which can only make effective predictions in a relatively short period of time.

The SST data used in this study are reanalysis data and measured data from the National Data Center for Marine Science. The data are from the Northwest Pacific Ocean Reanalysis Product (CORA V1.0). The product elements include sea surface height, temperature, salinity, and currents. The sea area ranges from 99°E to 150°E and 10°S to 52°N, the spatial horizontal grid resolution is 0.5° × 0.5°, and the number of the vertical layer is 35. The length of time is 60 years from January 1958 to December 2018, and the time resolution is the monthly average of the past years. The spatial horizontal grid resolution of the measured data is 0.125° × 0.125°, which is the same as that of the reanalyzed data. The time span is 5 years from 2016 to 2020, and the time resolution is the daily average. If there is a null value in the data set, the cubic spline interpolation method is used for complement.

The product was developed based on the ocean reanalysis system of the Northwest Pacific Ocean, and the ocean dynamic model of the system was the Princeton Ocean Model with the Generalized Coordinate System (POMGCS). The meteorological driving field is the NCEP meteorological reanalysis field. The ocean data assimilation method used is the multi-grid three-dimensional variational ocean data assimilation method. The assimilated ocean observations include in situ temperature and salinity observations, satellite remote sensing sea surface height anomaly (SSHA), and sea surface temperature (Reynolds SST) data. National Marine Science data in the heart of the reanalysis data format for.nc database files is more than 67 Giga bytes, at the same time, the original file format in the actual use process is relatively complex, so it must be prepared in advance according to the requirements that will be appropriate for waters of the sea surface temperature extracted and stored as .mat format, and the read load can be used on MATLAB. This study takes the Bohai Sea as the research object and obtains the surface temperature of the Bohai Sea on a certain day, as shown in Figure 1.

FIGURE 1

As shown in Figure 2A, the observation of the SST series shows that the Bohai Sea area presents a single peak shape in the process of changing with month, that is, the maximum and minimum temperature values only appear once in every 12 months in a year. Along with the number of days in a month to promote the process of SST rendering multiple peak shapes, as shown in Figure 2B, that is, the trend of rising and falling will appear multiple times in a month. It can be seen that the variation characteristics of SST in different time scales are also different. If we observe SST on a daily scale, we can find that the change in SST is very dissimilar. If we observe SST on a monthly scale, it can be found that the SST has high similarity and obvious periodicity in different years. The similarity and periodicity of SST are conducive to the prediction of future temperature.

FIGURE 2

2.2 The evaluation metrics of sea surface temperature

In the process of predicting the future monthly average SST, we can make use of the similarity of the data over the past years to conduct the appropriate comparison. Therefore, the metrics of similarity assessment directly determine the accuracy of SST prediction results. There are many metrics to evaluate the similarity between two samples. The similarity deviation is introduced in this study. The mean relative error (MRE), the posterior difference ratio (PDR), and the probability of small error (PSE) are the metrics to evaluate whether the prediction sequence is suitable for the real sequence and sufficient to predict in the future.

If we have two samples, A(1) is the value of the first sample, B(1) is the value of the second sample, and the total number of both samples is N. Then, the mean relative error is as follows:

The residuals between two samples and the mean values of the residuals :

Posterior difference ratio:where

The is the variance of sample one, and is the variance of the residual between the first and second samples.

The calculation formula of the probability of small error (PSE) is defined in Eq. 7. Table 1 shows the relationship between the aforementioned parameters and the grey model accuracy.

TABLE 1

LevelMREPDRPSE
Ⅰ level<0.01<0.35>0.95
Ⅱ level<0.05<0.50<0.80
Ⅲ level<0.10<0.65<0.70
Ⅳ level>0.20>0.80<0.60

Predictive model test criteria.

In addition, the root mean square error (RMSE) and MRE are selected as the evaluation parameters of prediction accuracy. The formula for RMSE is as follows:

Similarity deviation is the parameter to reflect the difference between the “shape” and “value” of two samples. The similarity deviation of these two samples can be defined as SAB.where

In Eq. 10, reflects the coefficient of “value.” In Eq. 11, reflects the coefficient of “shape.” The default similarity deviation is the average value of the two metrics. The smaller the similarity deviation is, the higher the similarity between the two samples is. The monthly average temperature of the first two quarters of 3 years at a point in the Bohai Sea is defined as three samples. Based on the first sample as the benchmark, Table 2 shows the data points of the last two samples, as well as the calculated “value” coefficient, “shape” coefficient, and similar deviation. The results show that the second sample is more similar to the first sample. In other words, the use of this similarity criterion can provide some reference in the subsequent prediction.

TABLE 2

SampleI
123456
4.01811.86643.60236.999911.797517.3094
5.36153.60114.30847.951313.027018.10101.1260.3090.718
5.61743.55023.91097.956211.270019.23231.1660.7450.956

Result of similarity deviation.

The Lin’s concordance correlation coefficient (LCCC) was used to evaluate the prediction model performance, because it measures the “agreement” between predicted and measured values (Zhao et al., 2021a,b).where and are the means for the real and predicted values, and and are the corresponding variances.

Prediction accuracy was also assessed using the ratio of performance to deviation (RPD), which is calculated as the ratio of standard deviation (SD) to RMSE.

These two indexes divide the accuracy of the prediction model into four levels, as shown in Table 3.

TABLE 3

EvaluationLCCCRPD
Excellent>0.9>2.5
Very good0.8–0.92–2.5
Good0.65–0.81.8–2
Poor<0.65<1.8

Result of similarity deviation.

2.3 The SST prediction model

2.3.1 Data preprocessing

is defined as the sea surface temperature (SST) sequence.

The original data were cumulated to get the cumulative sequence, so as to weaken the volatility and randomness of the original sequence. The new data sequence is defined as .

According to the smoothness ratio test theory, the grade ratio and smoothness ratio test of SST can be defined as

When , if and , the data follow the exponential law and meet the smoothness requirements, so the grey prediction model for the sequence can be established (). Table 4 shows the grade ratio and smoothness ratio of the monthly average SST in the recent 5 years. According to the corresponding data, it can be found that when , the maximum grade ratio of the monthly average SST series of 2020 is 1.613, and the maximum smoothness ratio is 0.380, which meets the data index law and smoothness requirements.

TABLE 4

Month 20162017201820192020
21.4640.3171.8580.4621.3780.2741.7380.4241.8100.448
31.7820.4391.6330.3881.7540.4301.7340.4231.3850.278
41.7630.4331.6120.3791.8150.4491.6380.3891.5370.349
51.8000.4451.5730.3641.7200.4191.5940.3731.6130.380
61.6100.3791.4870.3281.6090.3791.5250.3441.5800.367
71.4770.3231.3950.2831.4290.3001.4080.2901.4500.310
81.3130.2391.2930.2261.3300.2481.2860.2221.3220.244
91.1930.1621.1850.1561.2070.1711.1940.1631.2090.173
101.1370.1211.1120.1001.1430.1251.1250.1111.1400.123
111.0850.0781.0660.0621.0890.0821.0850.0781.0750.070
121.0560.0531.0530.0501.0700.0651.0520.0491.0540.051

Original sequence grade ratio and smoothness ratio test.

2.3.2 Modeling

For the cumulative sequence set up GM(1,1|sin) power model, the corresponding bleaching equation is definewhere is called the development coefficient, is called the grey action, and and are constant. When , this model is converted into the traditional GM(1,1) model, when the , this model is converted into the GM(1,1|sin) model. After the values of and are determined, the column matrix composed of , , and is denoted as .

According to the cumulative sequence , the mean generator matrix and the constant term vector are defined as follows:

The least-squares method is used to obtain the grey parameter shown in .

By putting the grey parameter into the linear equation, we obtain the following value:

The 3/8 Simpson integral formula turns the integral into an interval sum (), and then, an approximate solution is obtained. Then, the integral is converted into

The final solution of the bleaching equation is shown as follows:

is an approximate value obtained by the least-squares method, so is just an approximate result. In order to distinguish it from the cumulative sequence , it is written as . The function expression subtracts in order to restore the original sequence, and then the approximate original sequence was obtained.That is,

Given and , the least-squares method can be used to solve the remaining parameters, and then the prediction curve is obtained. The least-squares method is a given algorithm that takes the sum of squares of errors as the objective function to find its minimum value, and the result has a unique output value. In the whole process of solving the model, the two parameters and directly determine the quality of the predicted results, so the reasonable choice of these two parameters is particularly important in the whole solution. The MRE of the prediction model is taken as the objective function, and the global search is carried out by using the genetic algorithm to calculate the minimum value of the average relative error.

3 Results

3.1 The similarity deviation

With a similar predict method, the GM(1,1|sin) power model realizes the SST prediction. Therefore, this study puts forward a new method, the GM(1,1|sin) power model based on the similarity deviation. This method takes two successive years as an original sequence, combined with the similarity of the SST changes and determines the appropriate original sequence on the basis of similarity deviation, in order to solve the model parameters, then solves the grey prediction model to predict. The unknown SST of 2020 can be assumed and predicted. The data of two consecutive years (2019–2020) are first combined, and the similarity of SST changes can be used to calculate the similarity deviation between each year and 2019 according to Eq. 9. The year with the highest similarity to 2019 was found and combined with the data of the next year to form the original sequence composed of 24 data for the module to determine the model parameters, and then the prediction curve was obtained.

Table 5 shows the data of similarity deviation of each year and 2019. The changing trend is shown in Figure 3. According to the calculation results, 2016 is the year most similar to 2019.

TABLE 5

Years19851986198719881989
Similarity deviation2.7732.4812.7712.8012.266
Years19901991199219931994
Similarity deviation2.7272.5562.3152.4732.546
Years19951996199719981999
Similarity deviation2.3682.4462.2492.3892.453
Years20002001200220032004
Similarity deviation2.5012.3751.8402.3672.573
Years20052006200720082009
Similarity deviation2.6972.7602.4642.4731.367
Years20102011201220132014
Similarity deviation1.6631.3691.1603.1052.659
Years20152016201720182019
Similarity deviation1.4820.883*0.9841.1180.000

Similarity deviation between all years and 2019.

The bold values represent that the results of the model of our paper compared with other models. The results of our paper are the best.

FIGURE 3

3.2 Monthly scale prediction

By combining the monthly average SST from 2016 to 2017 and bringing them into the model, the bleaching equation and the evaluation metrics are obtained. The prediction model is shown in Eq. 26. The specific data are shown in Table 6 and Table 7. The average values of MRE, RMSE, LCCC, and RPD are 9.84%, 1.2363, 0.9870, and 6.0996, respectively. The evaluation metrics of LCCC and RPD were excellent based on Table 3.

TABLE 6

MonthReal valueGM(1,1|sin) power model prediction value =-0.251 =3.103
15.61743.8666
24.55023.8743
33.91093.4965
47.55626.1501
513.270012.5664
620.232319.4237
724.817223.8727
825.758724.2542
922.100220.7817
1017.954015.3944
1110.954010.7021
128.48318.3704

Prediction results in 2020.

TABLE 7

Evaluation metrics of SSTMRE (%)RMSELCCCRPD
Value9.841.23630.98706.0996

Evaluation metrics of SST prediction results in 2020.

In order to eliminate the particularity of some years, the GM(1,1|sin) power model based on similarity deviation is used to predict the monthly average SST in the recent 5 years. The specific steps will not be repeated. Table 8 shows each predict year and its corresponding model. The maximum MRE is 13.28%, and the minimum is 5.54%. The 5-year MRE is 9.81%. In addition, the maximum RMSE is 1.3285, and the minimum is 0.6522. The 5-year RMSE is 1.0627, which indicates that the 5-year forecast deviated from the real value by about 1°C. All of evaluation metrics of LCCC and RPD were excellent. The specific contrast between prediction and reality is shown in Figure 4.

TABLE 8

YearGM(1,1|sin) power modelMRE (%)RMSELCCCRPD
201611.840.87930.99368.6965
20178.521.32850.98355.3129
201813.281.21740.98856.8892
20195.540.65220.996111.0875
20209.841.23630.98706.0995

Prediction models in 5 years.

FIGURE 4

In the recent 5 years, there are 2 years in which the MRE is more than 10%, which are 2016 and 2018. Comparing the real value of these 2 years with the value of other years, the lowest temperature on record occurred in February 2018 and February 2016, and the highest temperature in 2016 occurred in July, and there was a sudden temperature change from June to August in 2018. Because the grey prediction model belongs to an autoregressive model, these abnormal temperature phenomena will directly affect the prediction results, resulting in deviation of the prediction results from the real value. In the other 3 years, due to the relatively stable temperature change, the prediction value all obtained good results. It can be seen that the factors affecting the quality of the grey prediction model are not only related to the established parameters in the model, but also related to the data of the original sequence.

3.3 Spatial distribution map of the monthly scale

We have verified the suitability of the prediction model based on the results of the monthly scale prediction. The Bohai Sea is the only inland sea in China, and it is connected to the Yellow Sea in the southeast. There are many factors influencing SST variation in this area near the land margin. In winter, due to the cold current, the temperature in the central and southeastern areas of the Bohai Sea was lower than that in other areas. Accordingly, the temperature in these areas was higher under the influence of the summer warm current.

The spatial distribution of the Bohai Sea area is represented according to the data of the monthly scale prediction in 2020, as shown in Figure 5. In January and February, the SST in the center was low and around the coastline was almost the same. The lowest was 3.88°C in the area connected with the Yellow Sea. In March, the SST was further reduced, and the temperature range was between 3.768 and 3.786. This is because the land temperature in January and February is the lowest in the year. The spatial distribution of the SST was roughly the same from April to December. The temperature rose first and then decreased, and the highest was about 25.5°C in August.

FIGURE 5

4 Discussions

4.1 Comparison of the prediction models

This study selects a central Bohai Sea area (120°E-120.125°E and 38.5°N-38.625°N) as the research object. The monthly average SST of 2020 is selected as an original sequence. Table 9 shows the established GM(1,1|sin) power model (; ).Using the least-squares method, , , and were determined. The MRE of the GM(1,1|sin) power model is 4.20%, which is better than that of the GM(1,1|sin) model with 12.8%. The RMSE, LCCC, and RPD of GM(1,1|sin) power model are 0.5783, 0.9972, and 13.1378, respectively, which are better than other model results. The traditional GM(1,1) model and GM(1,1|sin + cos) model fail to describe the trend of the SST. Figure 6 shows the SST by four different models.

TABLE 9

MonthReal value of SST (°C)GM(1,1) model prediction value (°C)GM(1,1|sin + cos) model prediction value (°C) p = -2.393 q = 0.237GM(1,1|sin) model prediction value (°C) p = -0.787GM(1,1|sin) power model prediction value (°C) =0.676 =1.094
15.61745.61745.61745.61745.6174
24.550211.13174.56924.75925.6896
33.910911.78129.88525.04144.2765
47.556212.468512.82075.29947.5564
513.270013.196015.256410.789313.7744
620.232313.965914.917218.711320.3134
724.817214.780717.249724.857625.2506
825.758715.643015.813226.087725.8409
922.100216.555714.282322.176222.0370
1017.954017.521613.684515.947216.5014
1110.954018.54389.0327911.614211.0398
128.483119.62576.9808912.30428.4228
MRE62.52%35.63%12.8%4.20%*
RMSE6.97865.30001.68480.5783
LCCC0.33650.65370.97570.9972
RPD0.52370.77624.485813.1378

Different models data of months average SST in 2020.

The bold values represent that the results of the model of our paper compared with other models. The results of our paper are the best.

FIGURE 6

The metrics of the four models were also obtained as shown in Table 10. The optimal model, GM(1,1|sin) power model achieves Ⅱ level accuracy standard. Each model corresponding to the relative error is shown in Figure 7A, the GM(11|sin) power model of relative error is shown in Figure 7B. The maximum and minimum relative error of GM(1,1) model are 2.01 and 0.00557, respectively. The GM(1,1|sin + cos) model corresponding relative error maximum value is 1.53, and the minimum value is 0.0042. The GM(1,1|sin) model corresponding relative error maximum value is 0.45, and the minimum value is 0.00163. The GM(1,1|sin) power model corresponding relative error maximum value is 0.25, and the minimum value is 0.00026. According to the relevant metrics PDR and PSE based on Table 1, the results proved that the GM(1,1|sin) power model can approximately reflect the monthly changes in SST.

TABLE 10

ModelMREPDRPSE
GM(1,1) model0.62520.81041
GM(1,1|sin + cos) model0.35630.39651
GM(1,1|sin)model0.12800.04701
GM(1,1|sin) power model*0.04200.00561

Evaluation metrics of grey models in SST.

FIGURE 7

4.2 Comparison of respective prediction and consecutive prediction

Considering that the monthly average SST from 2016 to 2020 is to be predicted, the prediction value of 2016 will be taken as the real value after obtained, and the data will be predicted for five consecutive years by using the cycle prediction method. The results of the monthly SST prediction from 2016 to 2020 with the data from 1985 to 2015, 2016, 2017, 2018, and 2019 are shown in Table 8. The maximum MRE is 13.28%, and the minimum is 5.54%. The 5-year MRE is 9.81%. The average RMSE, LCCC, and RPD are 1.0627, 0.9897, and 7.617, respectively. The annual prediction model is shown in Table 11. The maximum value of the MRE is 13.91%, the minimum is 7.80%, and the average value is 11.07%. The average RMSE, LCCC, and RPD are 1.0603, 0.9894, and 7.497, respectively. All of evaluation metrics of LCCC and RPD were excellent. It can be seen that the prediction performance of the GM(1,1|sin) power model is stable because the deviation between the predicted value and the real value is very close in the respective prediction and consecutive prediction. The specific contrast between prediction and real is shown in Figure 8.

TABLE 11

YearGM(1,1|sin) power modelMRE (%)RMSELCCCRPD
201611.840.87930.99258.6965
20179.921.53140.97684.6402
201813.910.93710.99288.4634
20197.800.88330.99368.1988
202011.901.07060.99087.4863

Adjusted prediction models in 5 years.

FIGURE 8

The overall prediction trend and real condition in recent 5 years are shown in Figure 9. The overall data trend of the two predictions is the same, and the MRE of all of the respective prediction is larger. The MRE of the prediction value is 9.84%. Using the same method to predict from 2016 to 2019, the MRE is 11.84%, 8.52%, 13.28%, and 5.54%. Compared with the real value, the prediction value obtained using this method in the continuous prediction of the past 5 years has a maximum MRE of 13.91%, a minimum of 7.80%, and an average of 11.07%. The average values of RMSE, LCCC, and RPD are 1.0603, 0.9894, and 7.497, respectively. The average monthly SST from 1985 to 2020 is selected to predict the SST in the next 50 years, and the obtained results are given in Figure 10. It also predicted the daily average SST in each month.

FIGURE 9

FIGURE 10

4.3 Limitations

Due to the limitation of time and data, there are still more work conducted in future study. For example,

  • (1) In the process of predicting sea surface temperature, only the temperature itself is considered for analysis and prediction. Actually, SST is related to a set of factors such as atmospheric temperature, sea water salinity, ocean current movement, and so on. These factors should be considered comprehensively in the model, and the factors should be weighted to rank the influencing factors to find out the physical theories that really affect the SST.

  • (2) Because the effective time interval of time series prediction is short, the longer the prediction time is, the greater the error will be. In addition, since many countries have been aware of the impact of global warming, they will take more green and sustainable measures to mitigate the adverse trend in the future. Therefore, the prediction results of this study after 50 years are believed to have certain deviation from the measured results.

  • (3) We used the cubic spline interpolation method to fill the null values, which will definitely cause deviation from the real value. In the selection of data points, only the data near the center of the Bohai Sea were collected, and the data from other areas were ignored. The model established on the data set may not be completely applicable universally.

5 Conclusion

In this study, based on the empirical prediction method, the grey prediction method is used to analyze the sea surface temperature changing trend and create the prediction model. The prediction models are validated by the mean relative error and the similarity deviation metrics. The main work and achievements of this study are summarized as follows:

  • (1) The study carried out the analysis by experience prediction methods, considering that the SST has certain periodicity and the sustainability of change, similarity, and correlation with other marine elements, to make a qualitative or quantitative prediction. The method is simple and easy to construct, the predict effect is satisfactory. The MRE of this model is 4.20% when describing the monthly average SST in 2020.

  • (2) According to the constructed grey prediction model, a validation experiment was conducted from January to December 2020. The experiment combined the similarity deviation in statistics, establishing a model by selecting appropriate similar years, and then predicting the target year. The MRE of the prediction value is 9.84%. Using the same method to predict from 2016 to 2019, the MRE values are 11.84%, 8.52%, 13.28%, and 5.54%. Compared with the real value, the prediction value obtained using this method in the continuous predict of the past 5 years has a maximum MRE of 13.91%, a minimum of 7.80%, and the average values of MRE RMSE, LCCC, and RPD are 11.07% 1.0603, 0.9894, and 7.497, respectively. It also predicted the daily average SST in each month of 2020. The MRE is between 1.49% and 9.89%. The lowest result appears in December and the highest occurs in March.

Statements

Data availability statement

We appreciate the data provided by National Science and Technology Resource Sharing Service Platform - National Marine Science Data Center (http://mds.nmdis.org.cn/ accessed on 10 January 2021). Information about the data accessed can be found in the article, further inquiries can be directed to the corresponding authors.

Author contributions

Conceptualization: JG, FM, and XL; methodology, data curation, formal analysis, investigation, and writing—original draft preparation: FM, ZQ, MG, and JC; and writing—review andediting and supervision: JG, LW-e, and FM. All authors have read and agreed to the published version of the manuscript.

Funding

This research was funded by the National Natural Science Foundation of China, grant number 41671158, supported by the Foundation of Liaoning Educational Committee (LJKZ0979), supported by the College Students’ Innovative Entrepreneurial Training Plan Program (S202110165051), and supported by Youth Innovation Promotion Association of Chinese Academy of Sciences to LW-e.

Acknowledgments

The authors would like to thank Liaoning Normal University for the laboratory facilities and the necessary technical support. They also thank the National Science and Technology Resource Sharing Service Platform (China National Marine Science Data Center) for data support.

Conflict of interest

The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.

Publisher’s note

All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors, and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.

References

Summary

Keywords

sea surface temperature, grey theory, GM(1,1|sin) power model, genetic algorithm, prediction

Citation

Meng F, Gu J, Wang L, Qin Z, Gao M, Chen J and Li X (2022) A quantitative model based on grey theory for sea surface temperature prediction. Front. Environ. Sci. 10:1014856. doi: 10.3389/fenvs.2022.1014856

Received

09 August 2022

Accepted

14 October 2022

Published

10 November 2022

Volume

10 - 2022

Edited by

Bing Xue, Institute for Advanced Sustainability Studies (IASS), Germany

Reviewed by

Dongrui Han, Shandong Academy of Agricultural Sciences, China

Bao-Jie He, Chongqing University, China

Updates

Copyright

*Correspondence: Jilin Gu, ; Ling-en Wang,

This article was submitted to Environmental Informatics and Remote Sensing, a section of the journal Frontiers in Environmental Science

Disclaimer

All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article or claim that may be made by its manufacturer is not guaranteed or endorsed by the publisher.

Outline

Figures

Cite article

Copy to clipboard


Export citation file


Share article

Article metrics