High-frequency CSI300 futures trading volume predicting through the neural network
Abstract
EconStor is a publication server for scholarly economic literature, provided as a non-commercial public service by the ZBW.
Full text
Xu, Xiaojie; Zhang, Yun Article High-frequency CSI300 futures trading volume predicting through the neural network Asian Journal of Economics and Banking (AJEB) Provided in Cooperation with: Ho Chi Minh University of Banking (HUB), Ho Chi Minh City Suggested Citation: Xu, Xiaojie; Zhang, Yun (2024) : High-frequency CSI300 futures trading volume predicting through the neural network, Asian Journal of Economics and Banking (AJEB), ISSN 2633-7991, Emerald, Leeds, Vol. 8, Iss. 1, pp. 26-53, https://doi.org/10.1108/AJEB-05-2022-0051 This Version is available at: https://hdl.handle.net/10419/334112 Standard-Nutzungsbedingungen: Die Dokumente auf EconStor dürfen zu eigenen wissenschaftlichen Zwecken und zum Privatgebrauch gespeichert und kopiert werden. Sie dürfen die Dokumente nicht für öffentliche oder kommerzielle Zwecke vervielfältigen, öffentlich ausstellen, öffentlich zugänglich machen, vertreiben oder anderweitig nutzen. Sofern die Verfasser die Dokumente unter Open-Content-Lizenzen (insbesondere CC-Lizenzen) zur Verfügung gestellt haben sollten, gelten abweichend von diesen Nutzungsbedingungen die in der dort genannten Lizenz gewährten Nutzungsrechte. Terms of use: Documents in EconStor may be saved and copied for your personal and scholarly purposes. You are not to copy documents for public or commercial purposes, to exhibit the documents publicly, to make them publicly available on the internet, or to distribute or otherwise use the documents in public. If the documents have been made available under an Open Content Licence (especially Creative Commons Licences), you may exercise further usage rights as specified in the indicated licence. https://creativecommons.org/licenses/by/4.0/
High-frequency CSI300 futures trading volume predicting through the neural network Xiaojie Xu and Yun Zhang North Carolina State University at Raleigh, Raleigh, North Carolina, USA Abstract Purpose –For policymakers and participants of financial markets, predictions of trading volumes of financial indices are important issues. This study aims to address such a prediction problem based on the CSI300 nearby futures by using high-frequency data recorded each minute from the launch date of the futures to roughly two years after constituent stocks of the futures all becoming shortable, a time period witnessing significantly increased trading activities. Design/methodology/approach –In order to answer questions as follows, this study adopts the neural network for modeling the irregular trading volume series of the CSI300 nearby futures: are the research able to utilize the lags of the trading volume series to make predictions; if this is the case, how far can the predictions go and how accurate can the predictions be; can this research use predictive information from trading volumes of the CSI300 spot and first distant futures for improving prediction accuracy and what is the corresponding magnitude; how sophisticated is the model; and how robust are its predictions? Findings –The results of this study show that a simple neural network model could be constructed with 10 hidden neurons to robustly predict the trading volume of the CSI300 nearby futures using 1–20 min ahead trading volume data. The model leads to the root mean square error of about 955 contracts. Utilizing additional predictive information from trading volumes of the CSI300 spot and first distant futures could further benefit prediction accuracy and the magnitude of improvements is about 1–2%. This benefit is particularly significant when the trading volume of the CSI300 nearby futures is close to be zero. Another benefit, at the cost of the model becoming slightly more sophisticated with more hidden neurons, is that predictions could be generated through 1–30 min ahead trading volume data. Originality/value –The results of this study could be used for multiple purposes, including designing financial index trading systems and platforms, monitoring systematic financial risks and building financial index price forecasting. Keywords CSI300 futures, Trading volume, Forecasting, Neural network Paper type Research paper 1. Introduction For policymakers and participants of financial markets, predictions of trading volumes of financial indices are important issues. This is because such predictions carry significant market implications for financial index prices and their movements (Wang et al., 2013,2019; Hou and Li, 2014;Sohn and Zhang, 2017;Susheng and Zhen, 2014;Yan and Hongbing, 2018; Ausloos et al., 2020), which are an essential part of various decisioning processes with AJEB 8,1 26 JEL Classification —C45, C53, G12, G15, G17 © Xiaojie Xu and Yun Zhang. Published in Asian Journal of Economics and Banking. Published by Emerald Publishing Limited. This article is published under the Creative Commons Attribution (CC BY 4.0) licence. Anyone may reproduce, distribute, translate and create derivative works of this article (for both commercial and non-commercial purposes), subject to full attribution to the original publication and authors. The full terms of this licence may be seen at http://creativecommons.org/licences/by/4.0/ legalcode Funding: No funds, grants or other support were received. Human participants or animal participants: This article does not contain any studies with human participants or animals performed by any of the authors. Conflict of interest: The authors have no relevant financial or non-financial interests to disclose. The current issue and full text archive of this journal is available on Emerald Insight at: https://www.emerald.com/insight/2615-9821.htm Received 20 May 2022 Revised 22 August 2022 24 November 2022 10 February 2023 3 April 2023 24 April 2023 Accepted 11 May 2023 Asian Journal of Economics and Banking Vol. 8 No. 1, 2024 pp. 26-53 Emerald Publishing Limited e-ISSN: 2633-7991 p-ISSN: 2615-9821 DOI 10.1108/AJEB-05-2022-0051 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
purposes of generating trading profits and preventing trading losses. Considering the tremendous importance and need to monitor ever-changing financial markets that ultimately determine the safety and soundness of the financial and economic environment of different countries and regions, it is of particular interest to regulators and traders to understand the issue of financial trading in the high-frequency domain (Xu and Zhang, 2023a). To fulfill this mission, the literature has witnessed a great amount of effort in constructing different types of models for making predictions. These modeling techniques have included traditional regression types of (time-series) econometric models and machine learning models. 1.1 Traditional regression and time series models Chen et al. (2011) have proposed a hierarchical model that has two different components, which combine an intraday approach and a daily approach, to predict the trading volumes of 30 DJIA stocks. Chen et al. (2011) have found that their proposed hierarchical method leads to higher prediction accuracy than any of the two individual approaches. Joseph et al. (2011) have used a simple linear regression model for predictions of abnormal activities of trading for 470 S&P 500 companies. Joseph et al. (2011) have determined that online search activities offer useful predictive information for the prediction horizon of one week. Brownlees et al. (2011) have compared a rolling average method and a multiplicative error model for prediction purposes of different exchange-traded fund volumes. Brownlees et al. (2011) have found that the multiplicative error model results in higher accuracy. Gharehchopogh et al. (2013) have applied a simple linear regression model for predicting the trading volume of the S&P 500 index by using predictive information from the price of the index. Ye et al. (2014) have compared static and dynamic versions of a volume-weighted approach for predicting SSE 50 stocks’intra-daily volumes. Ye et al. (2014) have suggested that the dynamic version leads to better predictions. Satish et al. (2014) have proposed combining an ARIMA model and a rolling average approach for predictions of trading volumes of 30 DJIA stocks. Bordino et al. (2014) have found that predictive information from Yahoo Finance could benefit predictions of trading volumes of NYSE and Nasdaq stocks on both daily and hourly frequency. Nasir et al. (2019) have used a hybrid framework of vector auto-regressive models and nonparametric methods for predictions of trading volumes of Bitcoin on a weekly basis through predictive information from Google searching activities. Nasir et al. (2019) have found that more searching activities are associated with higher trading volumes of Bitcoin. Kao et al. (2020) have combined a vector auto-regressive model with a smoothing approach for the purpose of assessing causality between trading volumes and financial returns. 1.2 Modern methods Chen et al. (2016) have compared a state-space model based upon Kalman filtering with a rolling average approach and a multiplicative error method for predictions of trading volumes of many stocks from different exchanges. Chen et al. (2016) have determined that the state-space technique leads to higher accuracy for intraday predictions. Ma and Li (2021) have proposed multi-state Kalman filtering for the purpose of generating higher prediction accuracy as compared to two-state filtering for trading volumes of nearly a thousand stocks in the USA. 1.3 Machine learning and deep learning techniques Kaastra and Boyd (1995) have explored the comparison between a neural network and an ARIMA model for predictions of trading volumes of different agricultural commodities’ futures contracts. Kaastra and Boyd (1995) have determined that the neural network generates higher prediction accuracy. Alvim et al. (2010) have examined comparisons among High-frequency CSI300 futures 27 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
a partial least squares, a support vector machine and a naive no-change model for predicting the trading volumes of nine Bovespa stocks. Alvim et al. (2010) have found that the naive model leads to the lowest prediction accuracy. Oliveira et al. (2017) have evaluated the usefulness of a support vector machine for predicting trading volumes of Dow Jones and S&P 500 on a daily basis by using predictive information from micro-blogging. Lu et al. (2020b) have combined feature extracting functions of a CNN model and predicting functions of an LSTM model for predicting prices of stocks by using predictive information from trading volumes and historical prices. Yan and Yang (2021) have utilized the same predictive information set as that of Lu et al. (2020b) in predicting prices of stocks through encoder/ decoder LSTM models. Zhao et al. (2021) have considered different machine learning models that include a support vector machine, a random forest and an LSTM, and a graph-based method for predicting trading volumes’movement patterns by using predictive information from prices of stocks. Zhao et al. (2021) have determined that the graph-based method leads to higher prediction accuracy. Separate from stock markets, Shen et al. (2021) have illustrated that an LSTM could be useful for predicting trading volumes of foreign business in different countries and regions. Zhang (2020) has demonstrated that a Levenberg–Marquardt trained neural network could be effectively utilized for predicting trading volumes of exports and imports. 1.4 Time series decomposition approaches Lu et al. (2020a) have explored various machine-learning techniques to predict trading volumes and prices of carbon emission rights. Lu et al. (2020a) have paid special attention to the use of ensemble mode decomposition and data smoothing methods. Xie et al. (2020) have investigated the usefulness of decomposition approaches for predictions of trading volumes of electricity. For financial indices, Liu et al. (2022) have shed light on how to decompose trading activities of stocks into shortand long-run components in assessing potential extreme trading information. Chac on et al. (2020) have incorporated ensemble mode decomposition and data smoothing techniques into an LSTM for predictions of prices of stocks. Chac on et al. (2020) have determined that such a framework could benefit from improving prediction accuracy. Regarding the case of the CSI300, recent work has generally focused on predictions of prices through the use of different time-series techniques (e.g. Wang and Chen, 2013;Xu, 2017,Xu, 2018,2019b;Zhang and Sun, 2017;Huang et al., 2018;Zhou et al., 2019a) and machine learning approaches (e.g. Sun et al., 2015;Yang and Cheng, 2015;Wang et al., 2016; Lu and Li, 2017;Yao et al., 2018;Ning, 2020;Long et al., 2019;Zhou et al., 2019b). Therefore, our present work targets at filling the research gap of trading volume predictions for the CSI300. Specifically, we address such a prediction problem based upon the CSI300 nearby futures by using high-frequency data recorded each minute from the launch date of the futures to roughly two years after constituent stocks of the futures all becoming shortable, a time period witnessing significantly increased trading activities. The stock market was established in China in the early 1990s. Since then, it has undergone through dramatic developments with economic growth. However, until March 2005, there was no financial index designed for the purpose of reflecting the overall market status. This situation was resolved on 04/08/2005 when the CSI300 was launched. This financial index includes 300 stocks traded in Shanghai and Shenzhen exchanges and reflects 70% of the total market capitalization. For further financial developments, the futures of the CSI300 index was launched on 04/16/2010. Since then, it has turned out to be the most actively traded financial contract in China. Two pilot programs by the China Securities Regulatory Commission are worth noting. First, on 12/05/2011, the CSI300’s underlying stocks available for shortable trading increased from 90 to 260. Second, on 01/31/2013, the CSI300’s underlying stocks AJEB 8,1 28 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
available for shortable trading increased from 260 to 300. Such programs have contributed to elevated trading volumes of the CSI300 and the nearby futures have attracted the most trading activities (Xu, 2019b). For understanding more institutional backgrounds of the CSI300, one could refer to Yang et al. (2012),Hou and Li (2013),Xu (2017,2018,2019b) and Xu and Zhang (2021c,2022b). To perform our prediction exercise, we adopt the neural network (denoted as NN) for modeling the irregular trading volume series of the CSI300 nearby futures. The NN has been found in the literature to have great prediction potential for financial and economic applications in terms of time-series data (e.g. Yang et al., 2008,2010;Wang and Yang, 2010; Cabrera et al., 2011;Zhang and Pan, 2014;Yang and Cheng, 2015;Kong and Zhu, 2018;Xu and Zhang, 2021b,2022d). We concentrated on answering the research questions as follows: are we able to utilize the lags of the trading volume series to make predictions; if this is the case, how far can the predictions go and how accurate can the predictions be; can we use predictive information from trading volumes of the CSI300 spot and first distant futures for improving prediction accuracy and what is the corresponding magnitude; how sophisticated is the model; and how robust are its predictions? Our results show that we could construct a rather simple neural network model with 10 hidden neurons to robustly predict the trading volume of the CSI300 nearby futures using 1–20 min ahead trading volume data. The model leads to the root mean square error of about 955 contracts. Utilizing additional predictive information from trading volumes of the CSI300 spot and first distant futures could further benefit prediction accuracy and the magnitude of improvements is about 1–2%. This benefit is particularly significant when the trading volume of the CSI300 nearby futures is close to be zero. Another benefit, at the cost of the model becoming slightly more sophisticated with more hidden neurons, is that predictions could be generated through 1–30 min ahead trading volume data. Our contributions to the literature are as follows. First, to our knowledge, the present work is the first one on predictions of trading volumes of the CSI300 nearby futures. Our results here would fill the research gap in terms of gaining understanding of the problem of trading volume predictions based on an important financial index. These would have implications from a practical standpoint for many economic agents, including policymakers, traders and investors and regulatory agencies. Specifically, the results would benefit from monitoring ever-changing trading activities for ensuring the safety and soundness of financial systems. Second, the current study is the first one that employs a powerful machine learning approach for the prediction purpose of the trading volume of the CSI300 nearby futures. Under the special and unique market structure, including a domestic individual investor having a high barrier to enter trading of the futures, as well as a foreign institutional investor with qualifications and relatively low participation ratios of institutional traders as compared to individual investors (Ng and Wu, 2007;Xu, 2017), we successfully illustrate that the NN could effectively make rather accurate and robust predictions of the trading volume, which shows the great potential of the NN under different market structures. Such results should be of practical use to financial indices traded in different countries that share a similar market structure with the CSI300, probably for a particular time period. Third, our analysis is the first one that adopts the high-frequency data recorded on a minute basis for financial trading volume predictions, although previous studies have investigated the issue for different financial indices using intraday trading data (Xu, 2018). A good understanding of high-frequency trading volume predictions could greatly benefit investors and policymakers in risk management and financial price index predictions as part of market evaluations. The remainder of this study is organized as follows. Section 2 describes data used. Section 3 discusses models. Section 4 presents results. Section 5 provides conclusions. High-frequency CSI300 futures 29 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
2. Data Our data are sourced from Wind Information Co., Ltd, which include trading volumes of the CSI300 spot and CSI300 futures [1]. The trading volumes are recorded on a minute basis, and for each trading day, the data span 9:16 a.m.–11:30 a.m. and 1:01 p.m.–3:15 p.m. The time period analyzed here ranges from 04/16/2010 (the date on which the futures was launched) to 11/14/2014. Thus, there are 299,970 observed trading volumes for each series. The data recorded on the 1-min basis not only reflect more trading activities than the data recorded on the 5or 10-min basis but also maintain sufficient economic significance so that investigations of the one-minute data carry necessary importance to traders and policymakers (Xu, 2018). Visualization of the trading volumes of the CSI300 spot, CSI300 nearby futures and CSI300 first distant futures [2] is provided in Figure 1. We could see from Figure 1 that these three trading volume series show obvious chaotic and noised patterns. Figure 1 also reveals that the nearby futures contract is most actively traded and its trading volumes have been expanding during the time period considered here. Summary statistics of the three trading volumes series are provided in Table 1. We could observe that the trading volumes are leptokurtic and skewed positively. 3. Models Two different types of nonlinear autoregressive neural network (denoted as ANN) models have been considered in this work. The first model is named a pure ANN or the CSI300 nearby futures own-lag only model. The model is denoted as follows: yðtÞ¼fðyðt1Þ;...;yðtdÞÞ;(1) where yis employed to reflect the trading volume of the CSI300 nearby futures, tis employed to reflect the time, dis employed to reflect the number of delays used by the model and fis employed to reflect the function form of the model. It is worth noting that the function form fis yet unknown in advance and the model estimated could be denoted as follows: yðtÞ¼ α 0þX k j¼1 α j f X d i¼1 βijyðtiÞþβ0j ! þ ε ðtÞ;(2) where kis employed to reflect the number of hidden layers used by the model with the transfer function being f ,β ij is employed to reflect the parameter that is associated with the connection’s weight between the input unit iand the hidden unit j, α j is employed to reflect the connection’s weight between the hidden unit jand the output unit, β 0j and α 0 are employed to reflect the constants that are associated with, respectively, the hidden unit jand the output unit and «is employed to reflect the error item. The second model is an ANN that has exogenous inputs (ANN–X). The model is denoted as follows: yðtÞ¼fðyðt1Þ;...;yðtdÞ;xðt1Þ;...;xðtdÞÞ ¼ α 0þX k j¼1 α j f X d i¼1 βijyðtiÞþβ0 ijxðtiÞ hi þβ0j ! þ ε ðtÞ;(3) where xis employed to reflect the trading volume of the CSI300 spot alone or xis employed to reflect the trading volumes of the CSI300 spot and CSI300 first distant futures together, and β0 ij is employed to reflect the parameter that is associated with the connection’s weight between the exogenous input unit iand the hidden unit j. The ANN–X model has included more predictive information as compared to the ANN model and thus could explore the potential AJEB 8,1 30 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
CSI300 spot CSI300 nearby futures contract CSI300 first distant futures contract Mean 275,070 1,514 292 Median 222,010 1,101 54 Maximum 9,550,879 31,586 30,849 Minimum 0 0 0 Standard deviation 230,860 1,473 718 Skewness 4.945 2.691 6.327 Kurtosis 79.902 16.450 77.031 Jarque–Bera pvalue <0.001 <0.001 <0.001 Source(s): Elaborated by the authors Figure 1. Trading volume data Table 1. Summary statistics of trading volume data High-frequency CSI300 futures 31 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
usefulness of the additional predictive information for improving prediction accuracy (Xu, 2019a,2020). We make use of the ANN models based upon a two-layer feed-forward network whose hidden layer adopts a transfer function in the form of a logistic sigmoid function as follows: f ðzÞ¼ 1 1þe−z(4) and whose output layer adopts a linear function. It should be noted that the output y(t) would be fed back via the delays to inputs of the neural network and the training process would take on the form of open loops in order to achieve efficiency purposes. In open loops, the true output will be utilized instead of the estimated output. To be more specific, adopting the open loop would help ensure inputs to neural networks being more accurate and that resultant neural networks have pure feed-forward architectures. In terms of numbers of hidden neurons, we have tested 5, 10, 15, 25, 35 and 50. Regarding the delays, we have examined 1, 2, 5, 10, 20 and 30. As a result, 36 testing pairs are considered. For model estimations, we have segmented the trading volume data by using 70% for training, 15% for validation and 15% for testing. For the consideration of robustness analysis, we have also examined the following alternative data segmentation ratios by reserving 15% of the trading volume series for testing: 60% for training and 25% for validation, 65% for training and 20% for validation, 75% for training and 10% for validation and 80% for training and5% for validation. One could employ different algorithms for training a machine learning model. For our case, we explored the following three algorithms: the Levenberg–Marquardt (Levenberg, 1944;Marquardt, 1963) algorithm, the scaled conjugate gradient (Møller, 1993) algorithm and the Bayesian regularization (MacKay, 1992;Foresee and Hagan, 1997) algorithm. These three algorithms have been demonstrated by previous studies in terms of their success in achieving relatively good accuracy under various circumstances (e.g. Doan and Liong, 2004;Kayri, 2016;Khan et al., 2019; Selvamuthu et al., 2019;Xu and Zhang, 2021a,d,2022a,c,2023b). Baghirli (2015) and Al Bataineh and Kaur (2018) have carried out targeted studies comparing these three algorithms. Table 2 contains the specification of each ANN and ANN–X model setting examined in the present work. Figure 2 visualizes the architecture of the final neural networks constructed in this study. Algorithm Levenberg–Marquardt scaled conjugate gradient Bayesian regularization Delay 1 2 5 10 20 30 Hidden neuron 5 10 15 25 35 50 Training vs validation vs testing 60% vs 25% vs 15% 65% vs 20% vs 15% 70% vs 15% vs 15% 75% vs 10% vs 15% 80% vs 5% vs 15% Source(s): Elaborated by the authors Table 2. Explored ANN and ANN–X model settings for the trading volume prediction of the CSI300 nearby futures AJEB 8,1 32 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
Figure 2. The three final neural network models’block diagram representations High-frequency CSI300 futures 33 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
show prediction results of 0 as visualized in Figure 6. This problem is somewhat remediated by further including the trading volumes of the CSI300 spot and CSI300 first distant futures in the ANN–X models that we turn to next. 4.2 ANN–X Similar to the analysis based on the ANN models, Figure 8 reports RMSEs for the ANN–X models when the trading volume of the CSI300 spot is further included as part of model training. We test the trading volume of the CSI300 spot prior to the trading volume of the CSI300 first distant futures due to the closer relation between the CSI300 nearby futures and CSI300 spot as compared to that between the CSI300 nearby futures and CSI300 first distant futures (Xu, 2019b). Balancing model prediction accuracy and model performance stabilities across different phases, we make the selection of the ANN–X model that has 15 hidden neurons and 30 delays, denoted as ANN–X-1. This model results in RMSEs of 949.37 for training, 946.50 for validation and 935.85 for testing. The summary of the ANN–X-1 model is included in Table 3. Predictions from the ANN–X-1 model are reported in Figure 9 and corresponding prediction errors are reported in Figure 10. We finally include the trading volumes of both the CSI300 spot and CSI300 first distant futures for model training and report RMSEs for ANN–XmodelsinFigure 11. Again, balancing model prediction accuracy and model performance stabilities across the three phases, we make the selection of the ANN–X model that has 35 hidden neurons and 30 delays, denoted as ANN– X-2. It results in RMSEs of 942.44 for training, 934.04 for validation and 940.20 for testing. The summary of the ANN–X-2 model is included in Table 3. Predictions from the ANN–X-2 model are visualized in Figure 12 and corresponding prediction errors are visualized in Figure 13. In Figure 14, we make comparisons of model performance based on the ANN-1, ANN–X-1 and ANN–X-2 models. It could be observed that including the trading volumes of the CSI300 spot and CSI300 first distant futures could help improve prediction accuracy by a robust modest magnitude of about 1–2%. This, however, significantly helpssomenear-zeropredictionsofthe trading volumes via the own-lag-only model, i.e. the ANN-1 model, as can be observed by making comparisons of prediction results shown in Figures 6, 9, and 12.Tobemorespecific,oneshould be able to observe, in Figure 6,thatthereexistcertainamountsofdatapointsthatshowpredicted trading volumes of zero or near zero while their associated observed trading volumes are not zero. These data points could be visually located in Figure 6, whose associated vertical axis (i.e. the predicted trading volume of the CSI300 nearby futures contract) values are zero or near zero while the corresponding horizontal axis (i.e. the observed trading volume of the CSI300 nearby futures contract) values are not zero. When we turn attention to results in Figures 9 and 12,wewould observe that such data points have been largely eliminated, particularly, for the results shown in Figure 12. Another benefit is that the prediction of the trading volume of the CSI300 nearby futures would be generated via 1–30 min ahead trading volume data with the incorporation of the additional series, considering that the ANN–X-1 and ANN–X-2modelsarebasedupon30delays and the 1-min data are employed for model building, although the complexity of the models would slightly increase as additional hidden neurons would be required. 4.3 Subperiod analysis To test whether prediction accuracy would be affected by the number of shortable stocks in the CSI300, we run the models, i.e. ANN-1, ANN–X-1 and ANN–X-2, on three subperiods with different numbers of shortable stocks and compare the results. The first subperiod is April 16, 2010–December 4, 2011, during which 90 stocks in the CSI300 are shortable. The second subperiod is December 5, 2011–January 30, 2013, during which 260 stocks are shortable. The third subperiod is January 31, 2013–November 14, 2014, during which all 300 stocks are shortable. The results are presented in Figure 15, together with those for the whole sample AJEB 8,1 40 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
860.00 880.00 900.00 920.00 940.00 960.00 980.00 1000.00 1020.00 1040.00 1 2 5 10 20 30 1 2 5 10 20 30 1 2 5 10 20 30 1 2 5 10 20 30 1 2 5 10 20 30 1 2 5 10 20 30 5 5 5 5 5 5 10 10 10 10 10 10 15 15 15 15 15 15 25 25 25 25 25 25 35 35 35 35 35 35 50 50 50 50 50 50 RMSE Top row: # of delays; boƩom row: # of hidden neurons Training ValidaƟon TesƟng Source(s): Elaborated by the authors Figure 8. ANN–X models (X 5the trading volume of the CSI300 spot): RMSEs High-frequency CSI300 futures 41 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
from April 16, 2010, to November 14, 2014. We could observe from Figure 15 that the trading volume of the CSI300 nearby futures of the first subperiod is most accurately predicted, followed by the second subperiod and then the third subperiod. This result is intuitive because with more stocks becoming shortable, the trading becomes more volatile and harder to predict. From Figure 15, we still observe that incorporating the spot and first distant futures improves predictions for the three subperiods. Specifically, this can be seen when comparing the result of ANN-1 with those of ANN–X-1 and ANN–X-2 for a given subperiod and subsample. Figure 9. ANN–X-1 predictions Figure 10. ANN–X-1 prediction errors AJEB 8,1 42 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
840.00 860.00 880.00 900.00 920.00 940.00 960.00 980.00 1000.00 1020.00 1040.00 1060.00 1 2 5 10 20 30 1 2 5 10 20 30 1 2 5 10 20 30 1 2 5 10 20 30 1 2 5 10 20 30 1 2 5 10 20 30 5 5 5 5 5 5 10 10 10 10 10 10 15 15 15 15 15 15 25 25 25 25 25 25 35 35 35 35 35 35 50 50 50 50 50 50 RMSE Top row: # of delays; boƩom row: # of hidden neurons Training ValidaƟon TesƟng Source(s): Elaborated by the authors Figure 11. ANN–X models (X 5trading volumes of the CSI300 spot and CSI300 first distant futures): RMSEs High-frequency CSI300 futures 43 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
4.4 Benchmark analysis We have performed benchmark analysis through comparisons of the ANN-1 model with the linear autoregressive (denoted as AR) model and the linear autoregressive integrated moving average (denoted as ARIMA) model. Considering that the ANN–X-1 and ANN–X-2 models include additional predictive information as compared to the AR and ARIMA models and the performance of ANN–X-1 and ANN–X-2 has been shown to be better than that of the ANN-1 model, we have not benchmarked them against the AR or ARIMA model. In determining the lag of the AR model, the Bayesian information criterion (Schwarz, 1978) has been employed. Figure 12. ANN–X-2 predictions Figure 13. ANN–X-2 prediction errors AJEB 8,1 44 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
The structure of the ARIMA model has also been determined through the Bayesian information criterion (Schwarz, 1978). We have applied the modified Diebold-Mariano (Diebold and Mariano, 2002) test (Harvey et al., 1997) for the purpose of making comparisons of model performance. The modified test mitigates some shortcomings of the original test, particularly the potential over-sized issues. The modified test is based upon d t as follows: dt¼error M1 t 2 error M2 t 2 ;(10) where error M1 tand error M2 tare employed to denote two error terms at time tthat are generated based upon model M 1 and model M 2 , respectively. Here, we would denote AR or ARIMA as model M 1 and ANN-1 as model M 2 . The test statistic for comparing model performance is denoted as MDM as follows: MDM ¼Tþ12hþT−1hðh1Þ T 1=2 T−1γ0þ2X h1 k¼1 γk !"# −1=2 d ;(11) where Tis employed to denote the length of the time period of the testing phase, his employed to denote the prediction horizon (h51 for our application), dis employed to denote the sample average of d t , γ0¼T−1X T t¼1 dt d 2(12) is employed to denote the variance of d t , and γk¼T−1X T t¼kþ1 dt d dt−k d (13) is employed to denote the kth auto-covariance of d t for k51, ...,h1 and h≥2. Under the null that two models being compared result in equal mean squared errors, the MDM test would follow the t–distribution whose degrees of freedom is T1. We report RMSEs stemming from the AR model and the ARIMA model in Table 3. We have found that the p values of the MDM tests are below 0.001. This result suggests that prediction accuracy stemming from the ANN-1 model is statistically significantly better than prediction accuracy 920.00 925.00 930.00 935.00 940.00 945.00 950.00 955.00 960.00 Training ValidaƟon TesƟng RMSE F1 (10 hidden neurons, 20 delays) F1 + Spot (15 hidden neurons, 30 delays) F1 + Spot + F2 (35 hidden neurons, 30 delays) Source(s): Elaborated by the authors Figure 14. Performance comparisons among the ANN-1 model (F1, 10 hidden neurons, 20 delays), the ANN–X-1 model (F1 þSpot, 15 hidden neurons, 30 delays), and the ANN– X-2 model (F1 þSpot þF2, 35 hidden neurons, 30 delays), where F1 is used to stand for the CSI300 nearby futures, Spot is used to stand for the CSI300 spot, and F2 is used to stand for the CSI300 first distant futures High-frequency CSI300 futures 45 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
0 200 400 600 800 1000 1200 1400 04/16/2010 - 11/14/2014 04/16/2010 - 12/04/2011 12/05/2011 - 01/30/2013 01/31/2013 - 11/14/2014 04/16/2010 - 11/14/2014 04/16/2010 - 12/04/2011 12/05/2011 - 01/30/2013 01/31/2013 - 11/14/2014 04/16/2010 - 11/14/2014 04/16/2010 - 12/04/2011 12/05/2011 - 01/30/2013 01/31/2013 - 11/14/2014 ANN-1 ANN-X-1 ANN-X-2 RMSE RMSE Training RMSE Validation RMSE Testing Source(s): Elaborated by the authors Figure 15. Subperiod analysis AJEB 8,1 46 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
stemming from the AR model and the ARIMA model. We have also conducted comparisons of performance based on the superior predictive ability (SPA) test (Hansen, 2005)asa robustness check and found that this test determines that the performance of the ANN-1 model is statistically significantly better than that of the AR model and the ARIMA model as well. We note that using the Akaike information criterion (Akaike, 1974) for determining the lag of the AR model and the structure of the ARIMA model does not affect this conclusion. It should be mentioned here that a certain model that is not performing as well as compared to another model would not necessarily mean that the particular model would not be able to contribute to prediction results. Many previous studies on prediction combinations actually have targeted at constructing different weights for different models’predictions with the purpose of potentially improving prediction accuracy. One interesting research field of prediction combinations is combining linear models and nonlinear models. Previous research, such as Stock and Watson (1998) and Blake and Kapetanios (1999), would have offered good examples in this research area. Hansen et al. (2011) have introduced the concept of the model confidence set (MCS), which is a useful technique to select optimal models with a given level of confidence. Following the same idea of comparing the ANN-1 model with the AR and ARIMA models, we have also compared the performance of the ANN–X-1 model with that of the AR–X-1 and ARIMA–X-1 models, where X refers to the trading volume of the CSI300 spot, and performance of the ANN–X-2 model with that of the AR–X-2 and ARIMA–X-2 models, where X refers to the trading volumes of the CSI300 spot and first distance futures. The RMSEs based upon the AR–X-1, ARIMA–X-1, AR–X-2 and ARIMA–X-2 models are reported in Table 3.Wehave found that the pvalues of the MDM tests are below 0.001 for comparisons between the ANN–X1 model and the AR–X-1 and ARIMA–X-1 models. This result suggests that the performance of ANN–X-1 is statistically significantly better than that of AR–X-1 and ARIMA–X-1. Similarly, we have found that the pvalues of the MDM tests are below 0.001 for comparisons between the ANN–X-2 model and the AR–X-2 and ARIMA–X-2 models. This result suggests that the performance of ANN–X-2 is statistically significantly better than that of AR–X-2 and ARIMA– X-2. As a robustness check, we have also conducted comparisons of performance based on the SPA test (Hansen, 2005) and found that it still holds that performance of ANN–X-1 is statistically significantly better than that of AR–X-1 and ARIMA–X-1 and performance of ANN–X-2 is statistically significantly better than that of AR–X-2 and ARIMA–X-2. 5. Conclusion For policymakers and participants of financial markets, predictions of trading volumes of financial indices are important issues. In this present work, we address such a prediction problem based on the CSI300 nearby futures by using high-frequency data recorded on a minute basis, which has never been explored in previous studies. We adopt the neural network for modeling the irregular trading volume series and have key empirical findings as follows. Our results show that we could construct a rather simple neural network model, trained via the Levenberg–Marquardt (Levenberg, 1944;Marquardt, 1963) algorithm, with 10 hidden neurons to robustly predict the trading volume of the CSI300 nearby futures using one to twenty minutes ahead trading volume data. The model robustly leads to the root mean square error of about 955 contracts across the three phases of training, validation and testing. Utilizing additional predictive information from trading volumes of the CSI300 spot and first distant futures could further benefit prediction accuracy and the magnitude of improvements is about 1–2%. This benefit is particularly significant when the trading volume of the CSI300 nearby futures is close to be zero. Another benefit, at the cost of the model becoming slightly more sophisticated with more hidden neurons, is that predictions could be generated through 1–30 min ahead trading volume data as the corresponding neural networks would use 30 High-frequency CSI300 futures 47 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
delays and 15 or 35 neurons, which also are trained via the Levenberg–Marquardt (Levenberg, 1944;Marquardt, 1963) algorithm. Our results could be used for multiple purposes, including designing financial index trading systems and platforms in terms of ongoing evaluating system/platform limits for processing trading activities, monitoring systematic financial risks in terms of ongoing detecting possible abnormal trading activities and building financial index price forecasting as suggested in the literature that one might make use of predictive information from the trading volume for helping improve the prediction accuracy of financial index prices. Our results here would be useful to policymakers from different countries for the purpose of designing another financial index or reforming an existing financial index. To be more specific, gaining good understanding of the trends of financial trading volumes would help the planning of a relatively new financial index from a thin market to a liquid and mature market. Although our present work focuses on relatively fundamental neural network models, developments in the machine learning field suggest that there exist more advanced models, such as the convolutional neural network and long short-term memory neural network, which have been seen in the literature for financial predictions. Explorations of more advanced models should be a worthwhile avenue for future studies on predicting financial trading volumes, including that of the CSI300 futures. Notes 1. It is possible that different platforms could generate slightly different trading volume data for each minute. It is worth noting that trading volumes of the CSI300 spot are always 0 from 9:16 a.m. to 9:29 a.m. on a trading day. 2. When different futures contracts are being investigated, the contract that has the closest settlement date is named the nearby contract. The first distant contract is the contract which settles right after the nearby contract. 3. We state “a relatively low complex model”from our empirical judgment that the ANN-1 model with 10 hidden neurons is not so complex for our case. References Akaike, H. (1974), “A new look at the statistical model identification”,IEEE Transactions on Automatic Control, Vol. 19 No. 6, pp. 716-723. Al Bataineh, A. and Kaur, D. (2018), “A comparative study of different curve fitting algorithms in artificial neural network using housing dataset”,NAECON 2018-IEEE National Aerospace and Electronics Conference, pp. 174-178, IEEE, doi: 10.1109/NAECON.2018.8556738. Alvim, L., dos Santos, C.N. and Milidiu, R.L. (2010), “Daily volume forecasting using high frequency predictors”,Proceedings of the 10th IASTED International Conference, p. 248. Ausloos, M., Zhang, Y. and Dhesi, G. (2020), “Stock index futures trading impact on spot price volatility. the CSI 300 studied with a Tgarch model”,Expert Systems with Applications, Vol. 160 No. 1, 113688, doi: 10.1016/j.eswa.2020.113688. Baghirli, O. (2015), “Comparison of Lavenberg-Marquardt, scaled conjugate gradient and Bayesian regularization backpropagation algorithms for multistep ahead wind speed forecasting using multilayer perceptron feedforward neural network”. Blake, A. and Kapetanios, G. (1999), “Forecast combination and leading indicators: combining artificial neural network and autoregressive forecasts”, Manuscript, National Institute of Economic and Social Research. Bordino, I., Kourtellis, N., Laptev, N. and Billawala, Y. (2014), “Stock trade volume prediction with yahoo finance user browsing behavior”,2014 IEEE 30th International Conference on Data Engineering, pp. 1168-1173, IEEE, doi: 10.1109/ICDE.2014.6816733. AJEB 8,1 48 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025
Brownlees, C.T., Cipollini, F. and Gallo, G.M. (2011), “Intra-daily volume modeling and prediction for algorithmic trading”,Journal of Financial Econometrics, Vol. 9 No. 3, pp. 489-518, doi: 10.1093/ jjfinec/nbq024. Cabrera, J., Wang, T. and Yang, J. (2011), “Linear and nonlinear predictablity of international securitized real estate returns: a reality check”,Journal of Real Estate Research, Vol. 33 No. 4, pp. 565-594, doi: 10.1080/10835547.2011.12091317. Chac on, H.D., Kesici, E. and Najafirad, P. (2020), “Improving financial time series prediction accuracy using ensemble empirical mode decomposition and recurrent neural networks”,IEEE Access, Vol. 8, pp. 117133-117145, doi: 10.1109/ACCESS.2020.2996981. Chen, S., Chen, R., Ardell, G. and Lin, B. (2011), “End-of-day stock trading volume prediction with a two-component hierarchical model”,The Journal of Trading, Vol. 6 No. 3, pp. 61-68, doi: 10.3905/ jot.2011.6.3.061. Chen, R., Feng, Y. and Palomar, D. (2016), “Forecasting intraday trading volume: a Kalman filter approach”, available at: SSRN 3101695. Diebold, F.X. and Mariano, R.S. (2002), “Comparing predictive accuracy”,Journal of Business and Economic Statistics, Vol. 20 No. 3, pp. 134-144, doi: 10.2307/1392185. Doan, C.D. and Liong, S.y. (2004), “Generalization for multilayer neural network bayesian regularization or early stopping”,Proceedings of Asia Pacific Association of Hydrology and Water Resources 2nd Conference, pp. 5-8. Foresee, F.D. and Hagan, M.T. (1997), “Gauss-newton approximation to bayesian learning”, Proceedings of International Conference on Neural Networks (ICNN’97), pp. 1930-1935, IEEE, doi: 10.1109/ICNN.1997.614194. Gharehchopogh, F.S., Bonab, T.H. and Khaze, S.R. (2013), “A linear regression approach to prediction of stock market trading volume: a case study”,International Journal of Managing Value and Supply Chains, Vol. 4 No. 3, p. 25, doi: 10.5121/ijmvsc.2013.4303. Hagan, M.T. and Menhaj, M.B. (1994), “Training feedforward networks with the marquardt algorithm”, IEEE Transactions on Neural Networks, Vol. 5 No. 6, pp. 989-993, doi: 10.1109/72.329697. Hansen, P.R. (2005), “A test for superior predictive ability”,Journal of Business and Economic Statistics, Vol. 23 No. 4, pp. 365-380, doi: 10.1198/073500105000000063. Hansen, P.R., Lunde, A. and Nason, J.M. (2011), “The model confidence set”,Econometrica, Vol. 79 No. 2, pp. 453-497, doi: 10.3982/ECTA5771. Harvey, D., Leybourne, S. and Newbold, P. (1997), “Testing the equality of prediction mean squared errors”,International Journal of Forecasting, Vol. 13 No. 2, pp. 281-291 No. 2, doi: 10.1016/S01692070(96)00719-4. Hou, Y. and Li, S. (2013), “Price discovery in Chinese stock index futures market: new evidence based on intraday data”,Asia-Pacific Financial Markets, Vol. 20 No. 1, pp. 49-70, doi: 10.1007/s10690-0129158-8. Hou, Y. and Li, S. (2014), “The impact of the csi 300 stock index futures: positive feedback trading and autocorrelation of stock returns”,International Review of Economics and Finance, Vol. 33, September 2014, pp. 319-337, doi: 10.1016/j.iref.2014.03.001. Huang, W., Lai, P.C. and Bessler, D.A. (2018), “On the changing structure among Chinese equity markets: Hong Kong, Shanghai, and Shenzhen”,European Journal of Operational Research, Vol. 264 No. 3, pp. 1020-1032, doi: 10.1016/j.ejor.2017.01.019. Joseph, K., Wintoki, M.B. and Zhang, Z. (2011), “Forecasting abnormal stock returns and trading volume using investor sentiment: evidence from online search”,International Journal of Forecasting, Vol. 27 No. 4, pp. 1116-1127, doi: 10.1016/j.ijforecast.2010.11.001. Kaastra, I. and Boyd, M.S. (1995), “Forecasting futures trading volume using neural networks”,The Journal of Futures Markets, Vol. 15 No. 8, p. 953, doi: 10.1002/fut.3990150806. High-frequency CSI300 futures 49 Downloaded from http://www.emerald.com/ajeb/article-pdf/8/1/26/9525990/ajeb-05-2022-0051.pdf by ZBW German National Library of Economics user on 16 December 2025