Purpose

This study aims to examine the efficacy of integrating machine learning (ML) architectures and feature selection protocols within a traditional asset pricing framework to enhance equity return predictability.

Design/methodology/approach

Leveraging a methodological pipeline that synergizes artificial neural networks (ANN) with sequential feature selection (SeFS) and Least Absolute Shrinkage and Selection Operator (LASSO) regularization, this study analyzed the momentum, value and quality risk premia across 949 conventional and 621 Islamic equities in the Indonesian market from 2016 to 2025. To isolate robust signals, this study further uses complete ensemble empirical mode decomposition with adaptive noise (CEEMDAN) for price denoising.

Findings

Empirical results indicate that momentum factors, particularly those with a one-month horizon, exhibit superior predictive power. Feature selection consistently identifies one-month momentum, earnings-to-price and gross profit-to-total assets as primary predictors. Notably, Islamic equities exhibit greater sensitivity to valuation anomalies, with EBIT/EV and gross profit-to-enterprise value providing additional predictive power. The ANN models achieve robust forecasting performance, with forecasting performance metrics ranging from 70% to 85%. The predictive outcomes exhibit significant invariance to the number of hidden layers, suggesting that factor risk premia possess an inherent structural stability that is not materially enhanced by increasing the complexity of the deep network.

Practical implications

The findings provide actionable insights for portfolio managers, Islamic fund managers and quantitative investors by demonstrating that ML models combined with factor investing strategies can substantially improve equity return forecasting in emerging markets. The study also highlights the relevance of short-term momentum and value-related factors for Islamic equities, supporting the development of more efficient Shariah-compliant investment strategies and AI-driven portfolio allocation systems.

Originality/value

This study advances the asset pricing and Islamic finance literature by integrating ANN, SeFS, LASSO and CEEMDAN within a unified equity return forecasting framework. It provides novel comparative evidence from 949 conventional and 621 Islamic Indonesian equities, highlighting the predictive dominance of short-term momentum and value-related factors. The findings also reveal the structural stability of factor risk premia across different neural network complexities in an emerging market context.

The efficient market hypothesis (EMH), which posits that no economic agent can consistently achieve superior returns relative to the market, remains a foundational theoretical construct in financial economics. Under the assumption of market efficiency, asset pricing frameworks have been developed to elucidate price dynamics and account for the heterogeneity in average returns across diverse asset classes (Cochrane, 2011). Within these pricing models, the expected return of an asset is defined by a linear combination of factor risk premiums, representing the compensation required by investors for exposure to specific systematic risks. Consequently, a vast body of literature has proposed an extensive array of factors, ranging from macroeconomic and fundamental variables to technical indicators and sentiment measures, to predict expected returns and investigate potential market inefficiencies (e.g. de Oliveira et al., 2013; Feng, Giglio, and Xiu, 2020; Hwang and Rubesam, 2019; Peng et al., 2021; Srijiranon et al., 2022). The EMH also evolved to AMF or adaptive markets hypothesis, whereby the markets are governed by competition, adaptation and natural selection. The implication is that the risk premium of an asset changes over time based on market conditions and the composition of market participants, emphasizing the importance of active risk management and dynamic adaptation in portfolio management (Lo, 2004).

Recent advancements in the field also highlight a transition toward advanced methodologies that transcend linear modeling to enhance the predictive accuracy of expected asset returns. The adoption of machine learning (ML) techniques facilitates the modeling of intricate, nonlinear relationships through optimized functions and hyperparameters. Accordingly, an expansive literature now examines the application of ML specifically for equity price forecasting. While certain studies prioritize technical indicator factors and investor sentiment factors (Peng et al., 2021; Zhen et al., 2025), others leverage broad sets of macroeconomic and fundamental factors (Gu, Kelly, and Xiu, 2020; Qiu, Song, and Akagi, 2016). These studies frequently use artificial neural networks (ANN) and support vector machines (SVM), often integrated with feature selection protocols – such as wrapper and embedded techniques – to ensure that model training is restricted to the most informative predictors.

Parallel to these developments, a substantial body of empirical research advocates for factor investing as a central component of active portfolio construction, highlighting three predominant investment styles: momentum, value and quality. For example, Fama and French (1996), and Lakonishok, Shleifer, and Vishny (1994) used value factors to capture value premiums within the stocks’ expected returns. While Jegadeesh and Titman (1993) used factors that capitalize on the persistence of historical price trends as momentum, Novy-Marx (2013) used quality factors that represent earnings power of the companies. Other studies focused on multi-factors to observe returns predictability (Dewandaru et al., 2015; Cakici et al., 2013; Asness, Moskowitz, and Pedersen, 2013).

However, the research gap exists whereby despite the progress in using ML for general return prediction, empirical research applying the ML techniques specific to the context of factor investing remains sparse. First, most research in factor investing used multilinear regression setup to capture risk premiums to predict stocks’ excess returns. The multilinear models are systematically deployed to generate consistent excess returns relative to established benchmarks. The efficacy of these active mandates is traditionally evaluated via the information ratio (IR), which scales the realized alpha (residual return) against the tracking error (residual volatility). Second, there is no research that used ML techniques specific to Islamic equities within the context of factor investing.

Our research objective is to bridge this gap by using ML as nonlinear techniques, integrated with feature selection methods, in the context of factor investing that captures momentum, value and quality risk premiums for both conventional as well as Islamic equities. There are two main research questions as the main incremental contributions of this research. The first is to investigate on how the predictive performance of factor investing specific to momentum, value and quality factors in the context of ML utilization. The second is to observe the differences in predictive performance of using ML in factor investing between those of conventional equities and Islamic equities.

Our methodological approach involves sequential feature selection (SeFS) and ANN for modeling, complemented by the Complete Ensemble Empirical Mode Decomposition with Adaptive Noise (CEEMDAN) technique to generate denoised price series with heightened predictability. The empirical analysis is conducted on a sample of 954 equities constituting the Jakarta Stock Exchange Composite Index (JKSE), as well as 621 equities of the Indonesia Sharia Stock Index (ISSE). In this case, our research also aims to compare the forecasting performance of our methods between those of conventional and Islamic equities. The data set comprises monthly observations of stock returns and factor variables from January 2015 through December 2025.

The structure of the paper is as follows: Section 2 provides a concise review of the literature; Section 3 describes the data and methodology; Section 4 elaborates on the empirical findings; and Section 5 concludes the study.

This section elaborates the literature review in order on returns predictability, factor investing and forecasting asset returns through the application of ML and decomposition techniques.

Under the assumption of market efficiency, asset pricing frameworks have been developed to elucidate price dynamics and account for the heterogeneity in average returns across diverse asset classes (Cochrane, 2011). Within these pricing models, the expected return of an asset is defined by a linear combination of factor risk premiums, representing the compensation required by investors for exposure to specific systematic risks. Consequently, a vast body of literature has proposed an extensive array of factors, ranging from macroeconomic and fundamental variables to technical indicators and sentiment measures, to predict expected returns and investigate potential market inefficiencies (e.g. de Oliveira et al., 2013; Feng, Giglio, and Xiu, 2020; Hwang and Rubesam, 2019; Peng et al., 2021; Srijiranon et al., 2022). The EMH also evolved to AMF or Adaptive Markets Hypothesis, whereby the markets are governed by competition, adaptation and natural selection. The implication is that the risk premium of an asset changes over time based on market conditions and the composition of market participants, emphasizing the importance of active risk management and dynamic adaptation in portfolio management (Lo, 2004).

Within the domain of active equity management, factor-based investment strategies are systematically deployed to generate consistent excess returns relative to established benchmarks. The efficacy of these active mandates is traditionally evaluated via the IR, which scales the realized alpha (residual return) against the tracking error (residual volatility). A substantial body of academic literature advocates extensive empirical evidence highlighting three predominant investment styles central to active portfolio construction.

The initial two styles consist of momentum and value investing, both of which are extensively documented in the literature (Dewandaru et al., 2015; Cakici et al., 2013; Asness, Moskowitz, and Pedersen, 2013). Momentum strategies use factors that capitalize on the persistence of historical price trends, where securities with high cumulative past returns tend to exhibit continued outperformance (Jegadeesh and Titman, 1993). On the other hand, value-oriented strategies use factors related to fundamental-to-price multiples to identify undervalued securities, predicated on the value effect (Fama and French, 1996; Lakonishok, Shleifer, and Vishny, 1994). The third style involves the integration of quality factors, encompassing dimensions such as accruals-based earnings quality (Sloan, 1996), financial robustness (Piotroski and So, 2012) and gross profitability (Novy-Marx, 2013).

Cochrane (2011) described the recent development of identified asset pricing determinants expanding to a vast array of features far exceeds the factors used in classical models of Fama and French (1993, 1996). Since estimating all possible combination of these factors or features is computationally unfeasible, feature selection is essential for identifying the subset of factors with optimal predictive performance. Furthermore, when using nonlinear models to identify complex patterns, a high volume of features can introduce significant noise. In this case, feature selection is a necessary preprocessing step to ensure only the most informative features are used during model training. Literature suggests that a vast array of factors can be condensed into a smaller group of relevant factors without compromising a model’s explanatory power.

Some research papers investigate macroeconomic and firm fundamental factors. For example, de Oliveira et al. (2013) evaluated 46 factors, including macroeconomic, firm fundamentals, historical prices and technical indicators, to predict the stock price direction of a Brazilian company. By applying a filter-based feature selection method using a correlation criterion, the authors reduced the original set to 18 key factors. Feng, Giglio, and Xiu (2020) evaluated 99 risk factors from companies listed on the NYSE, NASDAQ and AMEX using a Two-Pass Regression approach combined with Double Selection LASSO and Monte Carlo simulations. Their results indicated that most recently proposed factors are statistically redundant, with only a small number demonstrating significant explanatory power. Hwang and Rubesam (2019) evaluated 83 factors from the asset pricing literature using linear models and a Bayesian estimation approach for seemingly unrelated regressions. Testing these models on US stocks, they found that their method selected only ten factors as significant, further highlighting the high level of redundancy within the broader set of variables. Literature also revealed that the excess market return was the only factor that was consistently significant throughout the periods (for example, Nobi, Maeng, Ha, and Lee, 2013; Sensoy, Yuksel, and Erturk, 2013).

Other research focuses on sentiment and technical indicator factors. For example, Srijiranon et al. (2022) developed a hybrid computational framework for stock market forecasting by integrating Principal Component Analysis for factors reduction with long short-term memory (LSTM) architectures. Specific to technical indicator factors, Peng et al. (2021) investigated a set of 124 technical analysis indicators as factors, applying three feature selection algorithms to shrink the feature set. The findings revealed that the factors were not uniformly selected by the feature selection methods. Kumari and Swarnkar (2023) used multiple feature selection techniques, including FFS and LASSO, for 83 technical indicators using day-to-day stock data of six stock indices, before incorporating into ML techniques such as Support Vector Machine, K-Nearest Neighbour and ANN.

ML has become a prominent research focus in finance in recent years. These methods offer significant flexibility because they do not rely on restrictive assumptions regarding data distribution or functional forms. Instead, ML aims to identify non-intuitive patterns within the data to improve forecasting accuracy. These techniques encompass a variety of linear and nonlinear approaches, enabling the modelling of complex relationships through a streamlined set of functions and hyperparameters. Consequently, an extensive body of literature now explores the application of ML specifically for stock price prediction.

For example, Gu, Kelly, and Xiu (2020) compared several ML techniques, including ANNs, random forests and boosted regression trees against traditional linear models. Using a data set of nearly 30,000 financial assets from the NYSE and NASDAQ, the study measured risk premiums and found that ML methods significantly improved predictive performance. Nayak, Pai, and Pai (2016) used Boosted Decision Trees, Logistic Regression and SVMs to forecast trends in the Indian stock market. The research integrated historical price data with market sentiment analysis derived from social media.

Specific to the application of neural networks, Zhen et al. (2025) used LSTM-CNN-Attention model for stock price prediction of Chinese A-share market from 2018 to 2022, supported by the sentiment indicator extracted from the principal component. Srijiranon et al. (2022) used LSTM networks to refine market forecasting, operationalizing a dual-input approach that incorporates both historical price data and sentiment metrics. Qiu, Song, and Akagi (2016) used ANNs to forecast Nikkei 225 Index returns. Their model incorporated several macroeconomic variables as inputs, while integrating genetic algorithms with the neural networks to enhance overall predictive accuracy. Moghaddam, Moghaddam, and Esfandyari (2016) used ANNs to forecast daily returns for the NASDAQ stock exchange. Their model incorporated historical prices and the specific day of the week as key input variables. On the other hand, many recent studies have explored the use of deep neural networks due to their ability to extract abstract data representations through increasing the number of hidden layers. Comprehensive reviews of research in finance applying these deep learning models with more hidden layers were documented in systematic survey papers by Ozbayoglu et al. (2020) and Sezer, Gudelek, and Ozbayoglu (2020).

The most recent study used Support Vector Regression (SVR) with lagged inflation and output gap in a Taylor-rule framework (1964–2024) to forecast 1-year-ahead U.S. excess stock returns, explaining about 40% of their variation and revealing nonlinear, time-varying links between monetary policy conditions and equity risk premia (Roumani, AlSalman, and Murphy, 2026). Feng, Shi, and Kutan (2026) used dynamic automated ML on high-frequency data (2015–2024) for six major and emerging markets to show that local volatility indexes are the strongest predictors of local stock market volatility, VIX has weaker spillover power in China and India than in developed markets, COVID-19 reduced US but increased Chinese market predictability, and cross-market volatility spillovers are asymmetric and much stronger during turbulence.

As financial time series are inherently noisy and nonstationary, predicting their behavior remains a significant challenge for researchers. Due to the complex and chaotic nature of stock prices, individual ML algorithms often fail to produce stable results. To address this, data decomposition techniques, such as empirical mode decomposition (EMD), are integrated with ML to create hybrid models that better capture these complexities. EMD is a flexible, data-driven method designed to analyze nonlinear and nonstationary signals. Research indicates that EMD often outperforms traditional Wavelet and Fourier transforms in processing complex time-series data (Huang et al., 1999). By extracting intrinsic mode functions (IMFs), this technique allows researchers to isolate meaningful economic trends from residual noise within timeseries of prices.

Previous research adopted EMD to various forecasting areas in financial markets. For instance, Metwally et al. (2025) forecasted the KSE-100 index by combining SVM with CEEMDAN for data decomposition, finding that this hybrid approach significantly improved accuracy metrics. Similarly, Yang and Dai (2012) predicted the Shanghai and Shenzhen 300 index by applying EMD prior to an SVM model; the resulting high- and low-frequency IMFs and residual components enabled the SVM to achieve superior performance rankings. In addition, Lin et al. (2012) used Least Squares Support Vector Regression (LSSVR) to forecast foreign exchange rates by modeling each IMF and residual component individually, demonstrating that this approach outperformed traditional forecasting methods.

Islamic asset classes exhibit a distinct risk-return profile shaped by specific Shariah compliance requirements. In the equity sector, Shariah screening involves two primary stages: a qualitative assessment that excludes firms involved in prohibited industries (such as alcohol, gambling and conventional finance) and a quantitative assessment that restricts interest-bearing debt (Derigs and Marzban, 2008). As a result, Islamic equities typically maintain lower financial leverage, potentially reducing leverage effect during economic crises (Hamada, 1972). In addition, these screening processes also limit the number of eligible stocks and lead to higher concentration in certain sectors. Consequently, investing in Islamic equities may generate a distinctive risk-return profile relative to their conventional equities.

This section elaborates methodology used for estimation, as well as the samples and variables representing the factors.

To improve the predictive performance of traditional factor-based investing, our forecasting framework consists of two stages. First, we use ML techniques for feature selection, which include Sequential Forward Selection (SFS) and Least Absolute Shrinkage and Selection Operator (LASSO). The aim of this stage is to select the main predictors that have strong predictive power within our forecasting model. Second, we use ML technique for forecasting equity returns, which is ANNs. Our forecasting model uses the predictors that are selected in the first stage. In addition, our research also performs our two-stage forecasting framework on the denoised equity returns to observe whether the forecasting performance is improved. In this case, we use Ensemble Empirical Mode Decomposition with Adaptive Noise (CEEMDAN) to denoise the price returns of each equity.

3.1.1 Feature selection techniques.

Our research uses two approaches for selecting factors to be used for prediction. The first one is Information Coefficient (IC), whereby the fundamental law of active investment management defines IC as skills to produce alpha or value added in the active portfolio (Grinold, 1989; Qian and Hua, 2003). The IC is obtained from computing correlation between cross-sectional factors at time t and cross-sectional stock excess returns at time t + 1.

The second approach is feature selection, whereby literature has classified feature selection methods into filter methods, wrapper methods and embedded methods. Peng et al. (2021) argue that filter methods, such as those based on covariance or mutual information, often fail to account for complex, nonlinear dependencies between variables. In addition, wrapper methods produce more efficient and streamlined classification models than filter-based approaches (John, Kohavi, and Pfleger, 1994). Given that our research uses ANNs, which are specifically designed to capture nonlinear patterns and abstract feature relationships, filter methods were excluded in favor of wrapper and embedded selection techniques. Hence, our research uses Sequential Forward Selection (SFS) and LASSO for feature selection. Recent research used these selection techniques, combined with ANN for prediction (for instance, Peng et al., 2021; Kumari and Swarnkar, 2023).

Sequential Forward Selection (SFS) is an iterative search algorithm that builds a feature subset by progressively adding variables that maximize classification performance. Starting with an empty set, the process follows four primary steps:

  1. initializing an empty subset S;

  2. identifying and selecting the individual feature that yields the optimal evaluation score;

  3. recursively adding subsequent features that provide the greatest incremental improvement; and

  4. terminating the process once a predefined number of features d is reached or when further additions no longer enhance model accuracy.

On the other hand, LASSO is a regularization technique that incorporates a penalty term into the linear regression objective function (Tibshirani, 1996). The resulting coefficients are determined by solving the following constrained optimization problem:

(1)

This is similar to:

(2)

The parameter λ serves as a regularization constant that dictates the extent of coefficient shrinkage. As increases, it penalizes the magnitude of the regression coefficients, effectively forcing nonessential variables to zero to generate a sparse model. Typically, the optimal value for λ is determined through K-fold cross-validation to minimize out-of-sample error. In classification tasks, regularization is applied to the logistic regression likelihood function in a similar manner. This process results in LASSO logistic regression, where the model coefficients are determined by solving the following optimization problem:

(3)

whereby the likelihood function optimized to obtain the beta coefficients is:

(4)

3.1.2 Artificial neural networks.

Our research uses ANN as the ML technique for prediction. Henrique, Sobreiro, and Kimura (2019) conducted a mapping of 57 high-impact journal papers focused on ML applications in financial stock price prediction. Their findings identified ANNs and SVMs as the most prominent techniques used in the field. In addition, Nazário et al. (2017) reviewed 85 articles published between 1959 and 2016, classifying them by market types, methodologies, risk considerations and operational tools. This analysis highlighted the widespread use of ANNs, primarily citing their consistent performance when dealing with smaller data sets. There are some advantages of ANN as compared other ML techniques such as LSTM networks, SVR, LightGBM, XGboost and Random Forest (Ouf, El Hawary, Aboutabl, and Adel, 2025). ANN features less complexity without heavy sequence machinery that is hard to train and prone to overfitting noisy temporal patterns of LSTM, handling large data sets and high-dimensional feature spaces better without the kernel and support-vector scaling issues of SVR, as well as learning smooth and continuous response surfaces without piecewise-constant functions of tree-based models (XGboost, Random Forest). ANN is expressed as:

(5)

where y is a vector of stock returns as a dependent variable, X is a matrix of factors as observed independent variables, w is a vector of parameters to be estimated. W1, …, WN are the parameters for each hidden layer, wo are the parameters of the output layer, N is the number of hidden layers, and ψi are activation functions, with the incorporation of nonlinearity. Our research uses ANN with more hidden layers.

3.1.3 Ensemble empirical mode decomposition with adaptive noise (CEEMDAN).

EMD produces IMFs with decreasing frequency and energy as their order increases, with the first IMF containing the majority of the noise (Li et al., 2024; Bao et al., 2010; Rilling et al., 2003). Hence, the first IMF is eliminated and then sum up the remaining IMF components to reconstruct denoised series of prices data with higher predictability.

A significant limitation of the standard EMD technique is mode mixing, which can obscure the underlying signal (Vapnik and Vapnik, 1998). While Ensemble Empirical Mode Decomposition (EEMD) was developed to mitigate this by introducing Gaussian noise, it often fails to fully eliminate the added noise during reconstruction, leading to significant reconstruction errors (Vapnik and Vapnik, 1998). To address both the mode mixing in EMD and the residual noise issues in EEMD, Huang et al. (1998) introduced CEEMDAN, a more robust variation that ensures a more precise and efficient signal decomposition.

The first step for CEEMDAN is that the signal augmented with Gaussian noise is processed using standard EMD. This procedure yields the primary IMF component, which is characterized by the following mathematical expression:

(6)

The first residual component is calculated as:

(7)

The next step is to decompose residue r1t+γ1Ewit to compute the second mode as:

(8)

The last flat residue component can be computed by repeating all the steps for every IMF as follows:

(9)

3.2.1 Momentum, quality, and value factors.

On momentum investing, the methodology uses price momentum, operationalized as the compounded return realized over a retrospective horizon denoted as J (Leivo and Pätäri, 2011). This parameter J represents the formation period, which is the temporal window used to quantify historical performance (Rey and Schmid, 2007). Such metrics constitute a standard approach within the literature, with common empirical applications using lag specifications ranging from one to twelve months to capture various look-back sensitivities.

Within the framework of quality investing, the analysis centers on franchise value, financial robustness and accruals-based earnings quality. To quantify franchise value, the methodology uses profitability metrics such as return on assets, return on equity (ROE) and return on capital (ROC) (Greenblatt, 2010), alongside the ratio of gross profits to total assets as an alternative proxy for economic rents (Novy-Marx, 2013). Finally, earnings quality is evaluated using Sloan’s (1996) accrual measures, which serve to identify potential earnings manipulation or capital overinvestment.

For value investing, Gray and Carlisle (2013) provide a comprehensive synthesis of variables validated by extensive empirical research to capture value and quality premiums. The primary metric identified is earnings yield, the reciprocal of the price-to-earnings (P/E) ratio. A secondary measure is enterprise yield (EBITDA/EV), colloquially termed the acquirer’s multiple. A tertiary iteration substitutes EBIT for EBITDA, a core component of the magic formula framework that integrates value metrics with quality indicators (Greenblatt, 2010); alternative specifications further refine this yield by using free cash flow or gross profit as the numerator. Finally, the book-to-market ratio is included as a foundational benchmark for identifying valuation anomalies (Fama and French, 1992).

3.2.2 Data and approach.

Our research uses samples of Indonesia equities that are constituents of Indonesian conventional stock market index and Indonesian Islamic stock market index. We use 949 equities that are the constituents of JKSE, which serves as the Indonesian conventional stock market index. On the other hand, we use 621 equities that are the constituents of ISSE, which serves as the Indonesian Islamic stock market index. The samples are monthly data of stock returns and factor variables, ranging from January 2015 until December 2025. In terms of using feature selection techniques and ML methods to minimize overfitting in our forecasting, we use a training-testing proportion of 75% to 25% on the data set. The training data set is used for feature selection and ANN estimation. In addition, Heaton, Polson, and Witte (2017) documented that increasing the number of hidden layers in an ANN enables the algorithm to better capture stylized facts within financial data. In the context of factor models, this structure allows an ANN to generalize complex cross-interactions, effectively functioning as a hierarchical nonlinear factor model. Therefore, our research uses ANN with 3, 5 and 7 hidden layers, using the Sigmoid function σ (⋅) as the activation function and running 400 training epochs, with the aim to minimize overfitting in testing data set.

Peng et al. (2021) and Kumari and Swarnkar (2023) suggest the following metrics to measure the prediction performance:

In this context, TP and TN represent the counts of correctly identified positive and negative cases, respectively. Conversely, FP denotes the number of incorrect positive predictions (Type I error), while FN refers to the number of missed positive cases (Type II error). Precision serves as a performance metric that accounts for Type I errors, specifically identifying instances where assets were predicted to appreciate but actually declined. In contrast, recall captures the impact of Type II errors, representing profitable opportunities that the model failed to identify. Peng et al. (2021) argue that precision is a critical metric for risk-averse investors focused on minimizing capital losses, whereas recall is more relevant for those with a higher risk tolerance who seek to maximize the capture of potential market gains.

Table 1 details the descriptive statistics for momentum, quality and value factors across the constituent equities of the Indonesian equity market. Analysis of momentum indicators reveals positive mean returns that scale monotonically with the look-back horizon, providing empirical evidence of sustained price persistence over the preceding ten-year period. This longitudinal trend is corroborated by quality metrics, specifically ROE, which exhibit robust average values, suggesting that the observed market momentum is fundamentally underpinned by strong corporate profitability. Concurrently, the sample is characterized by compressed valuation multiples; the preponderance of value factors, most notably the book-to-market ratio, report low mean values, indicating a market environment defined by attractive fundamental pricing.

Table 1.

Descriptive statistics of factors

FactorsMeanSDMin.Max.SkewnessKurtosis
Momentum factorsMomentum 1-month0.0170.223−1.0009.3687.924173.089
Momentum 3-month0.0560.522−1.00033.71715.707547.630
Momentum 6-month0.1130.963−1.00063.82821.558922.480
Momentum 9-month0.1681.544−1.000117.50029.3191513.085
Momentum 12-month0.2222.112−1.000139.00027.0441144.394
Value factorsEarnings/Price−0.0580.969−47.13644.916−10.203833.265
Book/Market0.7697.122−555.26787.523−37.2182135.139
EBIT/EV0.0390.937−59.38848.017−15.0881836.146
EBITDA/EV0.1063.334−81.820511.570135.35820923.280
FCF/EV−0.0343.625−336.788118.389−57.2835287.391
Gross profits/EV0.30921.810−441.1783258.031145.69921825.637
Quality factorsROE0.18632.539−3568.2063250.842−3.3789705.057
ROA0.03631.790−1391.1513612.44367.0798552.381
ROC0.0362.149−98.474179.21131.9473001.708
Gross profits/TA0.1302.880−341.6529.833−98.08510389.037
Earnings quality0.00931.664−1302.1143623.81670.2808793.741

Given the substantial cross-sectional volatility evidenced by the high standard deviations of the factors, we use a z-score transformation across the entire factor suite. This normalization procedure adheres to rigorous factor investing conventions, mitigating the influence of outliers and ensuring a standardized scale. Furthermore, this adjustment is essential to enhance the efficiency and consistency of the estimators within our predictive modeling framework.

Our research executes a feature selection protocol on the training data set, encompassing the momentum, quality and value factor dimensions. Table 2 details the calculated ICs for the factor suite across both conventional and Islamic equity universes. The computation of ICs aims to quantify the cross-sectional predictive efficacy of each factor. This metric serves as a direct proxy for forecasting skill, consistent with the theoretical framework of the Fundamental Law of Active Management. The results indicate that momentum factors possess superior predictive efficacy relative to quality and value dimensions, with one-month momentum exhibiting the highest information content. A comparative analysis reveals that conventional and Islamic equities share broadly similar IC magnitudes across most factors, suggesting a high degree of commonality in factor risk premia between the two segments. However, idiosyncratic differences emerge within the value and quality clusters: the earnings-to-price ratio demonstrates higher predictive power for conventional stocks, whereas the book-to-market ratio serves as a more robust signal for Islamic equities. Regarding quality metrics, gross profit-to-total assets yields higher predictive utility within the conventional equity sample.

Table 2.

Information coefficients (ICs) of factors

FactorsICs in conventional equitiesICs in Islamic equities
Momentum factorsMomentum 1-month0.5020.508
Momentum 3-month0.2700.272
Momentum 6-month0.1930.193
Momentum 9-month0.1580.158
Momentum 12-month0.1330.127
Value factorsEarnings/Price0.0620.045
Book/Market0.0170.028
EBIT/EV0.0440.044
EBITDA/EV0.0420.041
FCF/EV0.0110.014
Gross profits/EV0.0390.035
Quality factorsROE−0.0070.006
ROA0.0440.041
ROC0.0140.014
Gross profits/TA0.0640.047
Earnings quality0.0210.013

For the feature selection methodology, we implement a dual-optimization approach using SeFS and LASSO regression. These methodologies are used to isolate the most parsimonious and statistically significant predictors from the factor pool, thereby mitigating the curse of dimensionality and enhancing the generalization capabilities of the neural network architectures. Table 3 delineates the subset of factors identified through SeFS and LASSO regularization for both conventional and Islamic equity cohorts. Within the conventional universe, a consensus emerges across both selection methodologies, pinpointing one-month momentum, earnings-to-price and gross profit-to-total assets as the primary predictors. These empirical results suggest that the one-month look-back period serves as the most potent proxy for capturing the momentum risk premium. While the earnings-to-price ratio, a fundamental metric in value-based strategies, is confirmed as a significant driver of the value premium, our findings highlight the gross profit ratio as the preeminent signal for quality. This observation aligns with the research demonstrated by Novy-Marx (2013), which posits that the gross profit-to-total assets ratio is a powerful metric that captures the “franchise value” or economic profitability of a firm, often outperforming traditional accounting measures like ROE or earnings. Specifically, it underscores a company’s capacity to generate high-margin operational returns relative to its asset base, a quintessential characteristic of sustained competitive advantage and brand equity.

Table 3.

Feature selection of factors

Conventional equitiesIslamic equities
FactorsSequential feature selectionLASSOLASSO coefficientsSequential feature selectionLASSOLASSO coefficients
Momentum factorsMomentum 1-month6.2356.600
Momentum 3-month
Momentum 6-month
Momentum 9-month0.002
Momentum 12-month
Value factorsEarnings/Price0.0980.076
Book/Market0.030
EBIT/EV0.0580.035
EBITDA/EV
FCF/EV
Gross profits/EV0.013
Quality factorsROE
ROA0.1670.087
ROC
Gross profits/TA0.0680.094
Earnings quality

The convergence of these three primary predictors is similarly evidenced within the Islamic equity universe. However, a distinct divergence occurs in Islamic equities, where both feature selection methods also incorporate EBIT-to-Enterprise Value (EBIT/EV) and gross profit-to-Enterprise Value (GP/EV) as significant value-based predictors. This suggests that Islamic equities may exhibit a more pronounced sensitivity to valuation anomalies, as a broader array of value metrics demonstrates the requisite predictive efficacy to capture the value premium within this specific asset class. This is understandable since Islamic equities are extracted from Shari’ah screening. The screening criteria impose a certain limit of interest-based debt in the balance sheets of the companies using debt-to-equity ratio. Consequently, Islamic equities are exposed to lesser financial leverage effect relative to that of conventional equities. Specifically, the equities with smaller leverage portray lower market to book values, along with lesser volatility in earnings and excess returns, which is in line with the main characteristics of undervalued stocks.

Our study uses the selected factors in our forecasting model. Table 4 details the empirical results of the ANN estimation, using a standard 75 / 25 training-to-testing data set partition. The predictive framework was evaluated across multiple specifications, incorporating various feature selection regimes and varying depths of hidden layers to assess the efficacy of deep learning architectures. The model demonstrates predictive power, with forecasting performance metrics – including accuracy, precision, recall and the F-score – consistently oscillating within the 70% to 85% range. Specifically, Table 4 shows that the Recall and F-score demonstrate the forecasting performance within the 70% to 80% range, whereas the Accuracy and Precision reveals the performance between 80% to 85% range. Interestingly, the results exhibit relative insensitivity to the specific feature selection methodology or the expansion of hidden layers. This lack of significant variance suggests that the factor risk premia derived from classical asset pricing models possess an inherent structural stability that is not materially enhanced by increasing model complexity or architectural depth. A cross-segment comparison reveals that Islamic equities yield marginally superior forecasting performance relative to conventional equities. This discrepancy is likely rooted in the fundamental composition of the Shariah-compliant universe; as identified in our feature selection analysis, Islamic equities exhibit a more pronounced orientation toward value factors. Consequently, the increased stability and persistence of the value premium over the long-term horizon facilitate more robust empirical estimation and higher predictive accuracy within this asset class.

Table 4.

Neural network forecasting

Stock universeFeature selection methodsNo. of hidden layersAccuracy (%)Precision (%)Recall (%)F-score (%)
Conventional equitiesSeFS384.7082.9875.2978.95
584.7782.7375.8879.15
784.7782.5476.1479.21
LASSO384.9683.9574.7979.11
584.7683.5874.6078.84
784.5282.7974.9078.65
Islamic equitiesSeFS385.2084.4978.1381.19
585.1584.7577.6581.04
784.9084.3877.3880.73
LASSO385.1584.2478.2881.15
585.1384.3978.0181.08
784.9383.9877.9780.86

To further refine the robustness of our results, we apply our methodology to equity prices processed via CEEMDAN. Following the literature (Li et al., 2024; Bao et al., 2010; Rilling et al., 2003), which identifies the first IMF as containing the preponderance of stochastic noise, we reconstruct the price series by aggregating the remaining IMF components. This denoising procedure aims to isolate the underlying signal, thereby enhancing the predictability of the data. Figure 1 presents the sample result of an Indonesia-listed equity in terms of decomposing equity prices, along with all the related IMFs.

Figure 1.
A stock-price decomposition for A D R O compares denoised and original prices and displays six intrinsic mode functions, I M F 1 through I M F 6.The title of the main plot is Stock Ticker: A D R O. The main plot compares Denoised Prices and Original Prices across about 150 observations. The vertical axis ranges from 0 to about 4,500. The two price traces closely follow one another. They begin near 600, rise to about 1,000, fall towards 500, and then increase to around 1,700. Later values fluctuate between about 1,000 and 2,400 before rising sharply above 3,000. The highest peak is near 4,000. The series then declines, rises again to about 3,800, and ends near 2,000. Six component plots appear on the right. I M F 1 contains rapid fluctuations around zero, with amplitudes increasing in the later observations and extending from roughly minus 500 to 450. I M F 2 also fluctuates around zero, with larger later oscillations extending from about minus 250 to 350. I M F 3 shows slower oscillations and several larger peaks and troughs, ranging from about minus 300 to 350. I M F 4 contains broader oscillations, ranging from approximately minus 500 to 500. I M F 5 shows long wave-like movements, ranging from about minus 800 to 750. I M F 6 rises gradually from about 1,000 to approximately 2,600 and then levels off near the end.

Decomposition of equity prices (CEEMDAN)

Figure 1.
A stock-price decomposition for A D R O compares denoised and original prices and displays six intrinsic mode functions, I M F 1 through I M F 6.The title of the main plot is Stock Ticker: A D R O. The main plot compares Denoised Prices and Original Prices across about 150 observations. The vertical axis ranges from 0 to about 4,500. The two price traces closely follow one another. They begin near 600, rise to about 1,000, fall towards 500, and then increase to around 1,700. Later values fluctuate between about 1,000 and 2,400 before rising sharply above 3,000. The highest peak is near 4,000. The series then declines, rises again to about 3,800, and ends near 2,000. Six component plots appear on the right. I M F 1 contains rapid fluctuations around zero, with amplitudes increasing in the later observations and extending from roughly minus 500 to 450. I M F 2 also fluctuates around zero, with larger later oscillations extending from about minus 250 to 350. I M F 3 shows slower oscillations and several larger peaks and troughs, ranging from about minus 300 to 350. I M F 4 contains broader oscillations, ranging from approximately minus 500 to 500. I M F 5 shows long wave-like movements, ranging from about minus 800 to 750. I M F 6 rises gradually from about 1,000 to approximately 2,600 and then levels off near the end.

Decomposition of equity prices (CEEMDAN)

Close Figure 1.

Table 5 delineates the ICs, while Table 6 summarizes the feature selection outcomes for the denoised series. Consistent with our prior findings, momentum factors maintain superior predictive efficacy, with the one-month and three-month horizons demonstrating the highest information content. As evidenced in Table 6, short-term momentum factors (1, 3 and 6-month horizons) are consistently identified as primary predictors for both conventional and Islamic cohorts. This suggests that the denoised returns exhibit a smoother risk-return profile, which is more effectively captured by short-range momentum signals. Furthermore, the persistent selection of multiple value factors within the Islamic equity sample reinforces our earlier conclusion that Islamic stocks are fundamentally characterized by a pronounced and predictable value risk premium orientation.

Table 5.

Information coefficients (ICs) of factors with denoised equity prices

FactorsICs in conventional equitiesICs in Islamic equities
Momentum factorsMomentum 1-month0.3320.340
Momentum 3-month0.3380.349
Momentum 6-month0.1940.203
Momentum 9-month0.1550.161
Momentum 12-month0.1240.125
Value factorsEarnings/Price0.0090.012
Book/Market0.0090.027
EBIT/EV0.0180.035
EBITDA/EV0.0200.037
FCF/EV0.0120.024
Gross profits/EV0.0200.025
Quality factorsROE0.0010.009
ROA0.0110.020
ROC0.0070.016
Gross profits/TA0.0070.008
Earnings quality0.0010.002
Table 6.

Feature selection of factors with denoised equity prices

Conventional equitiesIslamic equities
FactorsSequential feature selectionLASSOLASSO coefficientsSequential feature selectionLASSOLASSO coefficients
Momentum factorsMomentum 1-month1.153Yes1.192
Momentum 3-month1.543Yes1.518
Momentum 6-month−0.119Yes−0.077
Momentum 9-monthYes
Momentum 12-monthYes
Value factorsEarnings/Price
Book/Market0.069
EBIT/EV0.021
EBITDA/EV0.001
FCF/EV0.009
Gross profits/EV0.003
Quality factorsROE
ROA
ROC
Gross profits/TA
Earnings quality

Table 7 provides the empirical findings of the ANN estimation when applied to the denoised equity price series. The model demonstrates significant predictive capacity, with performance indicators – encompassing accuracy, precision, recall and F-score – yielding consistent values between 65% and 75%. Consistent with our preceding analysis, these predictive outcomes exhibit substantial invariance to changes in the feature selection protocol or the number of hidden layers. Notably, the forecasting performance observed with denoised prices is marginally inferior to that achieved with the original price series. The denoised price series mostly contain the trend components, which are captured by all momentum factors presented in Table 6. Nonetheless, removing the short-term components reduce the predictive power of both value and quality factors, thereby decreasing the overall forecasting performance for the denoised series. This suggests that the signal decomposition process may have inadvertently discarded idiosyncratic information related to quality and value risk premiums embedded within the raw data, thereby reinforcing the conclusion that market dynamics are efficiently captured by factor risk premiums within a classical asset pricing framework. Furthermore, the results corroborate our earlier findings that Islamic equities facilitate slightly more accurate forecasting compared to conventional equities.

Table 7.

Neural network forecasting with denoised equity prices

Stock universeFeature selection methodsNo. of hidden layersAccuracy (%)Precision (%)Recall (%)F-score (%)
Conventional equitiesSeFS371.9171.9266.0968.88
572.0172.0966.1068.96
771.8371.5766.5668.97
LASSO371.9171.9266.0968.88
572.0172.0966.1068.96
771.8371.5766.5668.97
Islamic equitiesSeFS372.6173.3766.4869.76
572.6273.3066.6469.81
772.4173.4565.6869.35
LASSO372.5573.0666.5369.65
572.6372.8767.1769.91
772.4972.3967.6869.95

This study investigates the efficacy of integrating ML architectures and sophisticated feature selection protocols within a traditional asset pricing framework to enhance the predictability of equity returns. While linear factor models remain the cornerstone of financial theory, we contend that nonlinear specifications are essential to capture the latent structural complexities of modern markets. We propose a methodological pipeline that synergizes ANN and SFS with established risk premia – specifically momentum, value and quality. To isolate robust signals, we further augment the data using CEEMDAN to generate denoised price series. The empirical robustness of this approach is tested on a decadal data set (2016–2025) comprising 949 constituents of the JKSE and 621 constituents of the ISSI, providing a unique comparative lens between conventional and Islamic equity universes.

Our empirical results, anchored by ICs, reveal that momentum factors exhibit superior predictive efficacy, with one-month momentum emerging as the most potent signal. The convergence of SeFS and LASSO regularization identifies a parsimonious set of primary predictors: one-month momentum, earnings-to-price and gross profit-to-total assets. Notably, the prominence of the gross profit ratio validates the “franchise value” hypothesis, signaling economic profitability more effectively than traditional accounting metrics. While these predictors are consistent across both cohorts, Islamic equities display a distinct sensitivity to a broader array of valuation anomalies, incorporating EBIT/EV and GP/EV as significant drivers. This suggests that Shariah-compliant assets are characterized by a more multifaceted and pronounced value risk premium.

The predictive performance of our ANN models remains robust, with classification metrics – including accuracy, precision, recall and F-score – consistently ranging between 70% and 85%. Interestingly, the models exhibit a high degree of invariance to architectural depth (the number of hidden layers) or specific feature selection methods, suggesting that the underlying factor risk premia possess an inherent structural stability that transcends algorithmic complexity. In other words, for practical usage to gain structural stability, investors or traders who use nonlinear ML techniques using both training and test data sets for forecasting asset returns should use of prominent risk factors that are derived from vast theoretical and empirical studies. On the other hand, a cross-segment analysis indicates that Islamic equities yield marginally superior forecasting results. This outperformance is likely attributable to their fundamental orientation toward value factors, which appear to offer greater persistence and stability for empirical estimation within this asset class.

Finally, our robustness tests using CEEMDAN-denoised returns corroborate the dominance of short-term momentum signals (1, 3 and 6-month horizons), which effectively capture the smoothed risk-return profiles of the reconstructed series. Although the denoised models maintain significant predictive capacity (65%–75%), they marginally underperform relative to the original price series. This performance delta implies that the decomposition process may discard idiosyncratic but meaningful information, reinforcing the conclusion that classical asset pricing factors, even in their raw state, efficiently capture the primary drivers of market dynamics. Collectively, these findings offer critical insights for investors and scholars navigating the intersection of factor investing and deep learning in emerging markets for both conventional and Islamic equities. The first implication for emerging market investors is that investors can enhance predictive performance of traditional-based factor investing by using ML techniques to capture nonlinearity in relationships. Second, emerging market investors who use ML can gain benefits of structural stability in ML architecture by using prominent risk premiums as a set of predictors, which are derived from theoretical and empirical evidence with respect to asset pricing models. Third, Shari’ah-compliant investors can gain similar forecasting performance relative to that of conventional investors in the context of using ML in factor investing, with Islamic factor investing is driven more by value premiums.

Asness
,
C.S.
,
Moskowitz
,
T.J.
and
Pedersen
,
L.H.
(
2013
), “
Value and momentum everywhere
”,
The Journal of Finance
, Vol.
68
No.
3
, pp.
929
-
985
.
Bao
,
F.
,
Wang
,
X.
,
Tao
,
Z.
,
Wang
,
Q.
and
Du
,
S.
(
2010
), “
EMD-based extraction of modulated cavitation noise
”,
Mechanical Systems and Signal Processing
, Vol.
24
No.
7
, pp.
2124
-
2136
.
Cakici
,
N.
,
Fabozzi
,
F.J.
and
Tan
,
S.
(
2013
), “
Size, value, and momentum in emerging market stock returns
”,
Emerging Markets Review
, Vol.
16
, pp.
46
-
65
.
Cochrane
,
J.H.
(
2011
), “
Presidential address: discount rates
”,
The Journal of Finance
, Vol.
66
No.
4
, pp.
1047
-
1108
.
de Oliveira
,
F.A.
,
Nobre
,
C.N.
and
Zárate
,
L.E.
(
2013
), “
Applying artificial neural networks to prediction of stock price and improvement of the directional prediction index–case study of petr4, petrobras, Brazil
”,
Expert Systems with Applications
, Vol.
40
No.
18
, pp.
7596
-
7606
.
Derigs
,
U.
and
Marzban
,
S.
(
2008
), “
Review and analysis of current shari’ah-compliant equity screening practices
”,
International Journal of Islamic and Middle Eastern Finance and Management
, Vol.
1
No.
4
, pp.
285
-
303
.
Dewandaru
,
G.
,
Masih
,
R.
,
Bacha
,
O.I.
and
Masih
,
M.M.
(
2015
), “
Combining momentum, value, and quality for the islamic equity portfolio: multi-style rotation strategies using augmented Black-Litterman factor model
”,
Pacific-Basin Finance Journal
, Vol.
34
, pp.
205
-
232
.
Fama
,
E.F.
and
French
,
K.R.
(
1992
), “
The Cross-Section of expected stock returns
”,
Journal of Finance
, Vol.
47
No.
2
, pp.
427
-
465
.
Fama
,
E.F.
and
French
,
K.R.
(
1993
), “
Common risk factors in the returns on stocks and bonds
”,
Journal of Financial Economics
, Vol.
33
No.
1
, pp.
3
-
56
.
Fama
,
E.F.
and
French
,
K.R.
(
1996
), “
Multifactor explanations of asset pricing anomalies
”,
The Journal of Finance
, Vol.
51
No.
1
, pp.
55
-
84
.
Feng
,
G.
,
Shi
,
J.
and
Kutan
,
A.M.
(
2026
), “
Your fear is (partly) mine: the role of non-VIX volatility in forecasting regional stock market volatility using interpretable machine learning
”,
Journal of International Money and Finance
, Vol.
160
, p.
103467
.
Feng
,
L.
,
Giglio
,
S.
and
Xiu
,
D.
(
2020
), “
Taming the factor zoo: a test of new factors
”,
The Journal of Finance
, Vol.
75
No.
3
, pp.
1327
-
1370
.
Gray
,
W.
and
Carlisle
,
T.
(
2013
),
Quantitative Value: A Practitioner’s Guide to Automating Intelligent Investment and Eliminating Behavioral Errors
,
Wiley Finance Series
.
Greenblatt
,
J.
(
2010
),
The Little Book That Still Beats the Market
,
John Wiley and Sons
,
NJ
.
Grinold
,
R.
(
1989
), “
The fundamental law of active management
”,
The Journal of Portfolio Management
, Vol.
15
No.
3
, pp.
30
-
37
.
Gu
,
S.
,
Kelly
,
B.
and
Xiu
,
D.
(
2020
), “
Empirical asset pricing via machine learning
”,
The Review of Financial Studies
, Vol.
33
No.
5
, pp.
2223
-
2273
.
Hamada
,
R.S.
(
1972
), “
The effect of the firm’s capital structure on the systematic risk of common stocks
”,
The Journal of Finance
, Vol.
27
No.
2
, pp.
435
-
452
.
Heaton
,
J.
,
Polson
,
N.
and
Witte
,
J.H.
(
2017
), “
Deep learning for finance: deep portfolios
”,
Applied Stochastic Models in Business and Industry
, Vol.
33
No.
1
, pp.
3
-
12
.
Henrique
,
B.M.
,
Sobreiro
,
V.A.
and
Kimura
,
H.
(
2019
), “
Literature review: machine learning techniques applied to financial market prediction
”,
Expert Systems with Applications
, Vol.
124
.
Huang
,
N.E.
,
Shen
,
Z.
and
Long
,
S.R.
(
1999
), “
A new view of nonlinear water waves: the Hilbert spectrum
”,
Annual Review of Fluid Mechanics
, Vol.
31
No.
1
, pp.
417
-
457
.
Huang
,
N.E.
,
Shen
,
Z.
,
Long
,
S.R.
,
Wu
,
M.C.
,
Shih
,
H.H.
,
Zheng
,
Q.
,
Yen
,
N.C.
,
Tung
,
C.C.
and
Liu
,
H.H.
(
1998
), “
The empirical mode decomposition and the Hilbert spectrum for nonlinear and non-stationary time series analysis
”,
Proceedings of the Royal Society of London. Series A: Mathematical, Physical and Engineering Sciences
, Vol.
454
No.
1971
, pp.
903
-
995
.
Hwang
,
S.
and
Rubesam
,
A.
(
2019
), “
Searching the factor zoo
”,
IÉSEG working paper series, 2018-ACF-03
,
available at:
Link to Searching the factor zooLink to the cited article.
Jegadeesh
,
N.
and
Titman
,
S.
(
1993
), “
Returns to buying winners and selling losers: implications for stock market efficiency
”,
The Journal of Finance
, Vol.
48
No.
1
, p.
65
.
John
,
G.H.
,
Kohavi
,
R.
and
Pfleger
,
K.
(
1994
), “Irrelevant features and the subset selection problem”, In
Machine Learning Proceedings 1994
,
Elsevier
, pp.
121
-
129
.
Kumari
,
B.
and
Swarnkar
,
T.
(
2023
), “
Forecasting daily stock movement using a hybrid normalization based intersection feature selection and ANN
”,
Procedia Computer Science
, Vol.
218
, pp.
1424
-
1433
.
Lakonishok
,
J.
,
Shleifer
,
A.
and
Vishny
,
R.R.W.
(
1994
), “
Contrarian investment, extrapolation, and risk
”,
The Journal of Finance
, Vol.
49
No.
5
, pp.
1541
-
1578
.
Leivo
,
T.H.
and
Pätäri
,
E.J.
(
2011
), “
Enhancement of value portfolio performance using momentum and the long-short strategy: the Finnish evidence
”,
Journal of Asset Management
, Vol.
11
No.
6
, pp.
401
-
416
.
Li
,
H.
,
Mei
,
Y.
,
Hao
,
X.
and
Chen
,
Z.
(
2024
), “
Out-of-sample equity premium predictability: an EMD-denoising based model
”,
Pacific-Basin Finance Journal
, Vol.
88
, p.
102536
.
Lin
,
C.-S.
,
Chiu
,
S.-H.
and
Lin
,
T.-Y.
(
2012
), “
Empirical mode decomposition–based least squares support vector regression for foreign exchange rate forecasting
”,
Economic Modelling
, Vol.
29
No.
6
, pp.
2583
-
2590
.
Lo
,
A.A.
(
2004
), “
The adaptive markets hypothesis
”,
The Journal of Portfolio Management
, Vol.
30
No.
5
, pp.
15
-
29
.
Metwally
,
D.S.
,
Ali
,
M.
,
Alghamdi
,
S.M.
and
Khand
,
D.M.
(
2025
), “
A novel hybrid model to forecast the stock price based on CEEMDAN and support vector regression
”,
Journal of Radiation Research and Applied Sciences
, Vol.
18
No.
2
, p.
101385
.
Moghaddam
,
A.H.
,
Moghaddam
,
M.H.
and
Esfandyari
,
M.
(
2016
), “
Stock market index prediction using artificial neural network
”,
Journal of Economics, Finance and Administrative Science
, Vol.
21
No.
41
, pp.
89
-
93
.
Nayak
,
A.
,
Pai
,
M.M.
and
Pai
,
R.M.
(
2016
), “
Prediction models for Indian stock market
”,
Procedia Computer Science
, Vol.
89
, pp.
441
-
449
.
Nazário
,
R.T.F.
,
e.
,
Silva
,
J.L.
,
Sobreiro
,
V.A.
and
Kimura
,
H.
(
2017
), “
A literature review of technical analysis on stock markets
”,
The Quarterly Review of Economics and Finance
, Vol.
66
, pp.
115
-
126
.
Nobi
,
A.
,
Maeng
,
S.E.
,
Ha
,
G.G.
and
Lee
,
J.W.
(
2013
), “
Random matrix theory and cross-correlations in global financial indices and local stock market indices
”,
Journal of the Korean Physical Society
, Vol.
62
No.
4
, pp.
569
-
574
.
Novy-Marx
,
R.
(
2013
), “
The quality dimension of value investing
”,
Ivey Energy Policy and Management Centre
, pp.
1
-
54
.
Ouf
,
S.
,
El Hawary
,
M.
,
Aboutabl
,
A.
and
Adel
,
S.
(
2025
), “
A comparative analysis of the performance of machine learning and deep learning techniques in predicting stock prices
”,
Journal of Computer and Communications
, Vol.
13
No.
4
, pp.
180
-
196
.
Ozbayoglu
,
A.M.
,
Gudelek
,
M.U.
and
Sezer
,
O.B.
(
2020
), “
Deep learning for financial applications: a survey
”,
Applied Soft Computing
, Vol.
93
, p.
106384
.
Peng
,
Y.
,
Albuquerque
,
P.H.M.
,
Kimura
,
H.
and
Portela
,
C.A.
(
2021
), “
Feature selection and deep neural networks for stock price direction forecasting using technical analysis indicators
”,
Machine Learning with Applications
, Vol.
5
, p.
100060
.
Piotroski
,
J.D.
and
So
,
E.C.
(
2012
), “
Identifying expectation errors in value/glamour strategies: a fundamental analysis approach
”,
Review of Financial Studies
, Vol.
25
No.
9
, pp.
2841
-
2875
.
Qian
,
E.
and
Hua
,
R.
(
2003
),
The Information Ratio of Active Management
,
Putnam Investments
.
Qiu
,
M.
,
Song
,
Y.
and
Akagi
,
F.
(
2016
), “
Application of artificial neural network for the prediction of stock market returns: the case of the Japanese stock market
”,
Chaos, Solitons and Fractals
, Vol.
85
, pp.
1
-
7
.
Rey
,
D.M.
and
Schmid
,
M.M.
(
2007
), “
Feasible momentum strategies: evidence from the Swiss stock market
”,
Financial Markets and Portfolio Management
, Vol.
21
No.
3
, doi: .
Rilling
,
G.
,
Flandrin
,
P.
and
Goncalves
,
P.
(
2003
), “
On empirical mode decomposition and its algorithms
”,
IEEE-EURASIP Workshop on Nonlinear Signal and Image Processing
, Vol.
3
, pp.
8
-
11
.
Roumani
,
Y.F.
,
AlSalman
,
Z.
and
Murphy
,
A.
(
2026
), “
Effective machine learning estimates of stock market returns using Taylor-rule inputs
”,
Economics Letters
, Vol.
259
, p.
112805
.
Sensoy
,
A.
,
Yuksel
,
S.
and
Erturk
,
M.
(
2013
), “
Analysis of cross-correlations between financial markets after the 2008 crisis
”,
Physica A: Statistical Mechanics and Its Applications
, Vol.
392
No.
20
, pp.
5027
-
5045
.
Sezer
,
O.B.
,
Gudelek
,
M.U.
and
Ozbayoglu
,
A.M.
(
2020
), “
Financial time series forecasting with deep learning: a systematic literature review: 2005–2019
”,
Applied Soft Computing
, Vol.
90
, p.
106181
.
Sloan
,
R.G.
(
1996
), “
Do stock prices fully reflect information in accruals and cash flows about future earnings?
”,
The Accounting Review
, Vol.
71
No.
3
, pp.
289
-
315
.
Srijiranon
,
K.
,
Lertratanakham
,
Y.
and
Tanantong
,
Y.
(
2022
), “
A hybrid framework using PCA, EMD and LSTM methods for stock market price prediction with sentiment analysis
”,
Applied Sciences
, Vol.
12
No.
21
, p.
10823
.
Tibshirani
,
R.
(
1996
), “
Regression shrinkage and selection via the lasso
”,
Journal of the Royal Statistical Society Series B: Statistical Methodology
, Vol.
58
No.
1
, pp.
267
-
288
.
Vapnik
,
V.
and
Vapnik
,
V.
(
1998
),
Statistical Learning Theory
, Vol.
1
Wiley
,
New York, NY
, p.
624
.
Yang
,
J-h.
and
Dai
,
X-Z
(
2012
), “
Prediction of Shanghai and Shenzhen 300 index based on EMD-SVM model
”, In
Paper presented at the 2012 International Conference on Systems and Informatics (ICSAI2012)
.
Zhen
,
K.
,
Xie
,
D.
and
Hu
,
X.
(
2025
), “
A multi-feature selection fused with investor sentiment for stock price prediction
”,
Expert Systems with Applications
, Vol.
278
, p.
127381
.
Published by Emerald Publishing Limited. This article is published under the Creative Commons Attribution (CC BY 4.0) licence. Anyone may reproduce, distribute, translate and create derivative works of this article (for both commercial and non-commercial purposes), subject to full attribution to the original publication and authors. The full terms of this licence maybe seen at Link to the terms of the CC BY 4.0 licenceLink to the terms of the CC BY 4.0 licence.

or Create an Account

Close subscription notice
Close access options