Bayesian nonparametric measurement of factor betas and clustering with application to hedge fund returns
Abstract
EconStor is a publication server for scholarly economic literature, provided as a non-commercial public service by the ZBW.
Full text
Garay, Urbi; ter Horst, Enrique; Molina, German; Rodriguez, Abel Article Bayesian nonparametric measurement of factor betas and clustering with application to hedge fund returns Econometrics Provided in Cooperation with: MDPI – Multidisciplinary Digital Publishing Institute, Basel Suggested Citation: Garay, Urbi; ter Horst, Enrique; Molina, German; Rodriguez, Abel (2016) : Bayesian nonparametric measurement of factor betas and clustering with application to hedge fund returns, Econometrics, ISSN 2225-1146, MDPI, Basel, Vol. 4, Iss. 1, pp. 1-23, https://doi.org/10.3390/econometrics4010013 This Version is available at: https://hdl.handle.net/10419/171866 Standard-Nutzungsbedingungen: Die Dokumente auf EconStor dürfen zu eigenen wissenschaftlichen Zwecken und zum Privatgebrauch gespeichert und kopiert werden. Sie dürfen die Dokumente nicht für öffentliche oder kommerzielle Zwecke vervielfältigen, öffentlich ausstellen, öffentlich zugänglich machen, vertreiben oder anderweitig nutzen. Sofern die Verfasser die Dokumente unter Open-Content-Lizenzen (insbesondere CC-Lizenzen) zur Verfügung gestellt haben sollten, gelten abweichend von diesen Nutzungsbedingungen die in der dort genannten Lizenz gewährten Nutzungsrechte. Terms of use: Documents in EconStor may be saved and copied for your personal and scholarly purposes. You are not to copy documents for public or commercial purposes, to exhibit the documents publicly, to make them publicly available on the internet, or to distribute or otherwise use the documents in public. If the documents have been made available under an Open Content Licence (especially Creative Commons Licences), you may exercise further usage rights as specified in the indicated licence. http://creativecommons.org/licenses/by/4.0/
Article Bayesian Nonparametric Measurement of Factor Betas and Clustering with Application to Hedge Fund Returns Urbi Garay 1, Enrique ter Horst 1,2,∗, German Molina 3and Abel Rodriguez 4 1IESA, Caracas 1010, Venezuela; [email protected] (U.G.) 2CESA, Bogota , Colombia 3Idalion Capital Group, London W1J 8NR, United Kingdom; [email protected] 4Baskin School, University of California at Santa Cruz, Santa Cruz 95064, USA; [email protected] *Correspondence: enriqueter[email protected]; Tel.: +57-310-813-0446 Academic Editors: Roberto Casarin, Francesco Ravazzolo, Herman K. van Dijk and Nalan Basturk Received: 8 June 2015; Accepted: 28 January 2016; Published: 8 March 2016 Abstract: We define a dynamic and self-adjusting mixture of Gaussian Graphical Models to cluster financial returns, and provide a new method for extraction of nonparametric estimates of dynamic alphas (excess return) and betas (to a choice set of explanatory factors) in a multivariate setting. This approach, as well as the outputs, has a dynamic, nonstationary and nonparametric form, which circumvents the problem of model risk and parametric assumptions that the Kalman filter and other widely used approaches rely on. The by-product of clusters, used for shrinkage and information borrowing, can be of use to determine relationships around specific events. This approach exhibits a smaller Root Mean Squared Error than traditionally used benchmarks in financial settings, which we illustrate through simulation. As an illustration, we use hedge fund index data, and find that our estimated alphas are, on average, 0.13% per month higher (1.6% per year) than alphas estimated through Ordinary Least Squares. The approach exhibits fast adaptation to abrupt changes in the parameters, as seen in our estimated alphas and betas, which exhibit high volatility, especially in periods which can be identified as times of stressful market events, a reflection of the dynamic positioning of hedge fund portfolio managers. Keywords: nonparametric clustering; Bayesian; cluster; nonparametric alpha and beta; hedge fund performance JEL: C2, C13, C14 1. Introduction We use a novel dynamic and arbitrary mixture of Gaussian Graphical Models (GGM) to describe sets of returns with complex relationships in a dynamic, nonstationary and nonparametric form and, at the same time, extract new nonparametric, nonstationary and dynamic alphas (excess returns) and betas (exposures to risk factors) of those returns. The procedure we follow builds on the formal mathematical model devised by [1]. However, we adjust their approach to financial settings, by constructing an approach that will extract and construct implied nonparametric alphas and betas, which are key parameters of interest in financial applications. A simulation study will offer an initial insight to the level of outperformance of our approach against some traditionally used approaches. Additionally, we apply this procedure to hedge fund index data for the period January 1994 to June 2009, and find that our estimated alphas or excess returns are, on average, 0.75% per month. These alphas are 0.13% per month higher than those estimated applying OLS for the same period and, thus, Econometrics 2016,4, 13; doi:10.3390/econometrics4010013 www.mdpi.com/journal/econometrics
Econometrics 2016,4, 13 2 of 23 our procedure uncovers that average hedge fund alphas could be underestimated when measured using OLS. Our paper contributes to the existing literature mainly in two ways: first, it devises a new methodology to extract market dynamic alphas and betas, which is less bound to widely-extended assumptions on the parametric structure of returns. This approach is especially beneficial when the generating process is very dynamic with varying numbers of clusters (and series within clusters). Second, we provide information about clusters that can be used by allocators/decision-makers outside of the more traditional expected return/risk-reward settings (for example, through relation-in-distress measurements, which could allow allocators to reduce exposures in combinations of assets/funds that tend to cluster when volatilities rise or when returns drop, or increase exposures in combinations that tend to de-cluster when returns drop, increasing the level of expected “idiosyncracy-under-stress” when it’s most needed in the portfolio). Our approach allows for the joint modelling of multiple series both as dependent and independent sets of variables. Our methodology finds a natural application in problems where a set of variables is driven by another set of variables in a difficult-to-parametrize, non-stationary fashion. Our application, hedge fund returns, can be seen as jointly driven by a set of factors, as opposed to fixed asset benchmarks, with exposures to those varying over time. Our results show much higher alphas to those considered through more traditional methods, indicating sufficiently relevant differences to consider this approach as one of non-marginal impact. These contrast with the alphas outlined in recent literature (see Titman and Tiu [2], Mamaysky et al. [3], Ferson and Schadt [4], Patton and Ramadorai [5], for a recent review). Furthermore, our estimated alphas and betas exhibit very high volatility, particularly in periods which can be identified as times of stressful market events. This is in line with the existence of many dynamic drivers in this particular application, such as dynamic market conditions, dynamic internal fund allocations, shift in portfolio manager styles and exposures, as well as dynamics of liquidity parameters of the individual funds. These findings question the constant alpha and beta assumption implicit in some studies conducted to measure hedge fund performance, as well as in the dynamic (yet parametric) approaches presented in others like [3]. Ours is more in line with a dynamic mixture model, with the mixtures viewed as the flexible, accomodating distributions to the different approaches to portfolio management and exposures of funds over time. The question of investment performance measurement (in both absolute and relative terms) has received increasing attention by both academicians and practitioners. The hedge fund industry, the focus of our application, has grown rapidly during the past two decades, and institutional investors such as endowments and pension funds have gradually increased their allocations to hedge funds. According to Hedge Fund Research, the total number of funds around the world grew to around ten thousand and total assets under management (AUM) reached around 2 trillion US dollars by the end of 2013. The hedge fund industry was severely affected by the global financial crisis of 2008–2009, with the total number of funds and AUM dropping by around a third during the crisis, although they mostly recovered by the end of 2010. The analysis of their returns, however, is nontrivial as the underlying processes are both complex and dynamic. Hedge fund managers change their allocations, positions, and their styles over time. Risk or leverage limits are often imposed. Market events change the focus of specific managers and new managers come and go with different styles. Additionally, the alpha of different styles changes over time, with periods where some styles outperform/underperform others over arbitrary periods before they revert. Even within a hedge fund group/style, there will be much different approaches that will generate very different return series (for example, depending on the frequency at which they operate). As such, any modelling approach must be sufficiently flexible to account for these (and other) changes in the nature of the return process over time. We offer a flexible, nonstationary, nonparametric approach that borrows information across dynamic clusters, but, more importantly, makes neither parametric assumptions on the nature of the (dynamic) alphas and betas nor on their autocorrelation structure. This approach,
Econometrics 2016,4, 13 3 of 23 although we focus on the Hedge Fund return application, is applicable to any set of return series where the dynamics are too complex and unknown to be modelled using parametric assumptions, yet a dynamic model is needed for all the key elements of the series (alphas, betas and clusters). It is important to mention that the focus of this paper is mainly descriptive, rather than inferential. Due to the high variability and the dynamic nature of the series we model, and the cluster dependence on the drivers of the series, although inference is technically possible, our focus is on providing a better understanding of the returns. A more accurate understanding of alphas, betas and clustering styles will provide decision-makers with new sources of information. This information, although usable directly in the allocation process (through the alphas and betas), finds also a natural space indirectly through cluster analysis (for example, putting limits on combined weights assigned to series that have a high probability of clustering in distress periods). The measurement of alpha, in the case hedge funds returns, as well as other complex return series, is complicated by their dynamic use of strategies that include long and short positions, as well as derivatives that result in very dynamic factor exposures against fixed benchmarks. This in turn generates non-linear returns that require tailored benchmarks. Furthermore, and according to [6], risk exposures of hedge funds have also declined in response to the rising dominance of institutional investors replacing family offices and private individuals as the primary source of investor capital. These same authors argue that the rapid growth of hedge funds has been responsible for the decline in their performance between the 1990s and mid-2000s and that perhaps “all the low-hanging fruits have been picked.” This view questions the use of a constant, mean-reverting level for the alphas, and calls for non-stationary models for key paramers, such as the one proposed in this paper. Additionally, it is also questionable that betas to key factors can be stationary, especially as funds deploy new ways to exploit competitive advantages (for example, as funds/sectors move to trading in higher frequencies, betas of those funds to factors defined in lower frequencies may diminish). When empirically tested, many of the approaches assume that the coefficients of the regressions are constant over time (or come from a constant distribution). If, in fact, these coefficients are time-varying and non-stationary, then the estimated parameters using these models would be unreliable. In the context of our application, as an example, [7] propose to measure the conditional performance of hedge fund indices using a stochastic discount factor approach, which imposes fewer limitations on the behavior of underlying returns. They find that estimated alphas, which are found to be positive, are similar when hedge fund performance is measured assuming either of the following four cases: (i) an absolute return approach (e.g., Alpha = Ri−Rf); (ii) a single-factor, fixed-exposure approach (e.g., CAPM); (iii) a single-factor, linear time-varying exposure approach (e.g., Merton’s model); and (iv) an extension of Merton’s approach consisting of a multi-factor, linear time-varying exposure approach. This finding of similar estimated alphas regardless of the model used leads them to conclude that better models are needed to measure hedge fund performance. Our approach finds very different (across funds and over time) alphas and betas, which seems more in line with expectations of highly dynamic funds. In a related prior approach, for a similar application, [3] address the issue of time variation in mutual fund factor loadings and develop a Kalman filter model to test for market-timing ability; showing that even though the Kalman filter model does not appear to exhibit market-timing ability at the daily frequency, it does so at the monthly frequency. While we focus our application on monthly returns, the flexibility of the design accomodates via a non-parametric approach to estimate dependence between alphas, making it less reliant on a strong parametric link between them over time, like the one imposed by the Kalman filter. Although a special case of our model and with another application in mind, [8] use a regime-switching beta model to measure dynamic risk exposures of hedge funds to various risk factors during different market volatility conditions. For interesting overlaps of our work with networks and graphical models please see [9,10]. Kosowski et al.[11] use Bayesian measures to estimate hedge fund alpha at the individual hedge fund level. These measures are based on the robust boostrap approach suggested by [12], and the
Econometrics 2016,4, 13 4 of 23 Bayesian framework of [13]. They argue, in the same vein as [14–16], that hedge fund performance measures do not follow parametric normal distributions because these funds hold derivatives such as options, because of the dynamic nature of their trading strategies, and also due to small sample problems. These features of hedge fund performance help explain their finding that Bayesian nonparametric measures yield superior performance predictability relative to alphas estimated when specific parametric models are assumed. The Bayesian paradigm provides in the case aforementioned, as well as ours, a natural, flexible tool to the fast estimation of the quantities of interest. Our approach is especially amenable for cases where small sample sizes also hinder many alternative approaches due to the (random) large dimensionality of the problem. The implementation of a Bayesian approach for measuring hedge fund performance should not be surprising, because Bayesian measures have been traditionally used to help overcome the small-sample problem typical of hedge fund returns1, making this also an area where our approach may become amenable. Using a robust boostrap procedure, [11] also find that hedge fund performance at the top cannot be explained solely by luck, that performance persists at annual horizons, and that OLS alphas of top hedge funds tend to be incorrectly estimated. It is worth noticing that the model used in [11] is not only parametric in nature, and therefore exposed to model risk, but also stationary. On the contrary, the alphas, as well as the betas derived in our paper, are nonstationary and dynamic. To the best of our knowledge there is no previous work that employs a formal mathematical model to describe in a dynamic, nonstationary and nonparametric way multivariate returns against multivariate factors, and extract at the same time some new nonparametric, nonstationary and dynamic alphas and betas, under a clustering scheme for shrinking, and in a fully Bayesian approach. Gaussian graphical models (GGMs), also called covariance selection models [22], are popular tools for modeling dependence across observables. GGMs assume that the returns generated by different asset classes and/or investment vehicles follow a joint multivariate Gaussian distribution, and explore the pattern of partial correlations to understand how the different outcomes influence each other. Because observations are assumed to follow a (arbitrary mixture of) multivariate Gaussian distribution, absence of partial correlation corresponds to a zero in the appropriate entry of the precision (inverse covariance) matrix and indicates that the variables are conditionally independent. The conditional independence patterns inferred in this way can be represented using a graph where nodes correspond to variables and an edge is present between two variables if they are conditionally dependent, and absent otherwise. GGMs are a natural alternative to endogenous factor models as described in [23,24], which explain the joint multivariate outcome as a linear combination of a small number of unknown factors determined directly from the data being modeled2. However, GGMs are particularly appealing over these endogenous factor models because they allow researchers to asses conditional rather than marginal independence, which in turn makes it straightforward to distinguish between direct and indirect interactions between the variables. GGMs have been successfully applied in finance and econometrics (for example, see [1,28–31]), where they have been shown to provide interpretability and enhanced predictive performance. However, our extraction of cluster-implied alphas and betas adds a new layer of outputs to the existing literature, since they allow direct use of these parametric measures derived from a non-parametric setting. 1The works in [17–20] had already applied Bayesian methodologies to measure mutual fund alphas. The consequence of using small samples is an even greater problem for hedge fund performance measurement when compared to mutual fund performance measurement due to higher variability and generally lower sample sizes due to a shorter lifespan. Furthermore, the use of short sales and derivatives by mutual funds, which are tightly regulated investment vehicles, is minimal, although this has been changing recently as reported by [21], whereas they are essential for many hedge fund strategies. 2Note that these endogenous factor models differ from more traditional exogenous factor models such as the three-, fourand fivefactor models of [25–27], which employ factors determined independently from the data being analyzed.
Econometrics 2016,4, 13 5 of 23 The paper is organized as follows. Section 2presents the statistical methods we employ. Section 3 offers sensitivity analysis and simulation exercises to compare our model to traditionally used models that pursue similar features of the data. In Section 4we describe the data used in the study, and the application through hedge fund returns. Section 5shows the main results obtained, while Section 6 offers the results of estimating time-varying alphas and betas, as well as cluster analysis, using the infinite hidden Markov model (iHMM-GGM) proposed by [1]. Finally, we outline our conclusions and potential extensions in Section 7. 2. Modeling Dependence among Asset Classes and Investment Vehicles In this section we describe the statistical methodology. We will start by reviewing basic notions of Gaussian Graphical Models and their link with sparse (regularized) regression. Then, we will move to describe hidden Markov models that employ Gaussian Graphical Models to describe state-specific distributions. 2.1. Gaussian Graphical Models Let zt= (z1t, . . . , zqt)0be a vector of observed returns, where zit represents the return of fund i over period t, for t=1, . . . , T. A Gaussian graphical model for the sequence z1, . . . , zTassumes that observations are independent and identically distributed from a multivariate normal distribution, zt∼Nqµ,K−1, where µis the vector of mean returns, Gis a graph3describing the conditional independence structure among the entries of zt, and K=K(G)is a precision matrix such that [K(G)]i,j=0, if and only if the edge connecting node iand node jis missing from G. Bayesian inference for GGMs proceeds by placing priors on the unknown parameters (µ,K,G). A popular approach uses conjugate priors and factorizes the joint prior p(µ,K,G)as p(µ,K,G) = p(µ|K)p(K|G)p(G) where p(µ|K)is a multivariate Gaussian distribution Nµ0,(n0K)−1,p(K|G)corresponds to a G-Wishart prior [32], and p(G)is a uniform distribution on the space of graphs. It is worthwhile to note that Kis a function of the graph Gsince Gwill define the structure of zeroes Kwill have. This prior will model the parameters as independently as possible conditional on the graph G. For small values of q, the posterior distribution of all model parameters can be computed in closed-form and inference is straightforward. For moderate to large values of q, the number of possible graphs is typically too large for explicit enumeration and Markov chain Monte Carlo algorithms that efficiently explore the space of graphs are typically required (for details and examples see [33–35], among others). GGMs can be used for sparse regression [29]. Consider a standard multivariate regression model, which assumes that the conditional distribution of the response ytgiven a vector of predictors xtis Gaussian. In our manuscript ytwill be the cross-sectional vector of hedge fund returns at time tand xtthe vector of market risk factors. An equivalent model is implied by a joint Gaussian distribution on the vector zt= (y0 t,x0 t)0. Indeed, if zt∼N(µ,K−1)where µ= µy µx!K= Kyy Kyx Kxy Kxx! then standard results for the multivariate Gaussian distribution imply that yt|xt∼Nµy−(Kyy)−1Kyx {xt−µx},(Kyy)−1(1) 3The graph G is assumed to be decomposable as explained in Rodriguez et al. [1] (p. 986).
Econometrics 2016,4, 13 6 of 23 Estimates of this joint model can be used to construct estimates of the regression function. Indeed, if maximum likelihood methods are used to estimate µand K, then the estimates of the intercept and slopes from (1) will be identical to those obtained through ordinary least squares (OLS). For example, in the Capital Asset Pricing Model (CAPM) the intercept αcorresponds to µy+ (Kxx)−1Kxyµxwhile the market risk parameter βcorresponds to −(Kxx)−1Kyx [36]. On the other hand, placing a prior on (µ,K)is equivalent to placing a prior distribution on the regression coefficients of a linear regression model together with a prior on the marginal distribution of xt. This procedure, although slightly more convoluted than simply placing a prior on the regression coefficients, can be easily generalizable to mixtures of Gaussian graphical models to generate nonparametric regression procedures. We used the CAPM in the previous analysis because the CAPM is a linear model and the IHMM-GGM model that we are using is built as an approximation of infinite linear OLS models. Thus, we can compare our model to those models that are more commonly used in finance (linear models such as the CAPM and OLS multivariate regressions). Our approach is less constrained because it is nonparametric, dynamic and nonstationary, whereas the CAPM is parametric, nondynamic, and stationary. 2.2. Dynamic Mixtures of GGMs Our discussion in the introduction suggests that hedge fund returns should not be assumed to be normally distributed, making the use of standard Gaussian graphical models potentially inappropriate for our application. To alleviate this issue we follow [1] and construct a hidden Markov model where the market fluctuates among an unknown number of clusters, and where hedge fund returns were generated from a collection cluster-specific GGMs. This structure for the hidden Markov model implies a nonparametric mixture model for the stationary (marginal) distribution on the market returns, which ensures that our model is capable of capturing all the stylized features of hedge fund returns that were discussed in Section 1. More specifically, introduce cluster variables ξ1, . . . , ξTwith ξt∈ {1, 2, 3, . . .}and let zt|µξt,Kξt,Gξt∼Nµξt,Kξt(Gξt)−1, i.e., conditional on the market being in cluster l(which corresponds to ξt=l), the observations are generated from a GGM with mean vector µland precision matrix Klsuch that its conditional independence structure is given by graph Gl. Therefore if Glhas off-diagonal zeroes this corresponds to independence between specific asset returns. If we look at the analysis unconditional of ξtthen the same interpretation of independence does not carry through anymore. To account for the sequential nature of the data we allow the cluster of the system to evolve as a Markov chain with a potentially infinite number of clusters, Pr(ξt=l|ξt−1=k) = πk,l where the vector πk= (πk,1,πk,2,πk,3, . . .)(containing the transition probabilities from cluster k to all other clusters) is given a Dirichlet process prior [37,38], πk|α,v∼DP(α,v), independently for each k, where v= (v1,v2, . . .)and vl=vl∏ k<l {1−vk},vl∼Beta(1, β).
Econometrics 2016,4, 13 7 of 23 The Dirichlet process is an extension of the Dirichlet distribution to countably infinite spaces. Indeed, the model just described can be obtained as the limit of a finite hidden Markov model, (πk,1, . . . , πk,L)|α,v∼Dirichlet(αv1, . . . , αvL) (v1, . . . , vL)|β∼Dirichlet β L, . . . , β L as L→∞. Therefore, this specification implies that E(πk) = v,i.e.,vrepresent a common mean for the vector of transition probability across clusters, while αcontrols how much each πkdeviates from this common mean. The model we just described (called an infinite hidden Markov model, or iHMM, see [39,40]) admits a priori an unlimited number of clusters. However, the prior structure is such that a posteriori only a few distinct regimes are actually employed to fit the data. Hence, this infinite HMM (iHMM-GGM) allows us to automatically and parsimoniously estimate the number of clusters. The iHMM specification above can be used to generate dynamic nonparametric estimates of the parameters of the CAPM. Indeed, conditioning on the cluster ξtwe can write an expression that is analogous to (1), yt|xt∼Nµy ξt−(Kyy ξt)−1Kyx ξtnxt−µx ξto,(Kyy ξt)−1. In other words, the model clusters time periods according to the cluster to which they belong and assigns the same CAPM parameters to periods in the same cluster. Uncertainty on the cluster indicators ξ1, . . . , ξTcan be incorporated through model average, leading to flexible estimates. Posterior computation for the iHMM involves the use of simulation-based Markov chain Monte Carlo algorithms. Given an initial guess for the model parameters, these iterative algorithms sequentially generate random realizations of blocks of parameters conditionally on the current values of the rest. After the algorithm has converged, these random draws can be used to produce approximate estimates of all parameters of interest, including point and interval estimates. For details, see [1]. The software we employed to fit our models is available from the authors by request. Last but not least, finite mixture problems have been used in different scientific communities, such as unsupervised clustering in neural network applications, latent class analysis in the social sciences, and regime switching models in economics, among others [41]. A finite mixture model has a finite number of clusters or states and therefore is a special case of our iHMM-GGM, which has infinite. Moreover, setting two clusters for ξt=1, 2 and noticing that our iHMM-GGM is also dynamic in nature, we get back a classical and well known regime switching model. For more on such conections between regime switching models and finite mixture models, please see [41] (Chapter 10). 3. Sensitivity Analysis and Simulation Exercises 3.1. Sensitivity An important question to ask is whether the prior distribution on L(prior number of clusters) has any impact on its posterior distribution. The ideal scenario is when the posterior of Lgiven the data does not change with different prior distributions of L, therefore making the estimation and algorithm robust to choices of prior. The prior on L implied by the Dirichlet process is given by [1]: p(L|α0,n) = S(n,L)n!αL 0 Γ(α0) Γ(α0+n)(2)
Econometrics 2016,4, 13 8 of 23 where S(n,L)denotes the Stirling number of the first kind [42]. We can conclude that the mean number of non-empty mixture components grows with α0, the concentration parameter the Dirichlet process DP(α0,M)which has a prior Gamma(a,b)[1]. We tested the effect of the priors with a representative range of key parameters aand b. As detailed in Equation (2), the key parameters for the implied prior on Lare the hyperparameters a and b. For these two hyperparameters, we define what could be a priori the boundaries of reasonable a priori values for the applications in this paper. Figure 1includes the distribution of three of those priors that represent a reasonable range for the problem at hand: •Gamma(a=1, b=5). This is a strong, very informative, prior that maps into a distribution of the clusters with a prior mean of clusters = 2 and a standard deviation of the number of clusters = 2 (prior two moments included in each of the graphs in Figure 1). This would be appropriate if we believe that there is a very limited number of clusters in the data. •Gamma(a=1, b=1). This is a medium-intensity prior that implies a larger number of clusters on average, and a heavier tail. We chose this prior for our analysis due to the nature of our data, where we anticipated a large number of clusters, but were uncertain as to how many. •Gamma(a=5, b=5). This is a low-intensity, more vague, prior that implies a much larger number of clusters. 0 5 10 15 20 25 30 35 0.00 0.05 0.10 0.15 0.20 0.25 0.30 Cluster Prior Distribution Under Different Hyperpriors Prior Number of Clusters Prior Density π(a=1,b=1) µ(L) = 5.37 σ(L) = 17.39 π(a=1,b=5) µ(L) = 2.05 σ(L) = 1.92 π(a=5,b=1) µ(L) = 17.81 σ(L) = 43.83 Figure 1. The implied prior on Lgiven a diferent set of hyperparameters values for aand b. This figure includes the distribution of three of those priors that represent a reasonable range for the problem at hand. Figure 2includes the posterior distribution of the number of clusters for each of the priors. As we can see, the choice of priors has non-negligible effects on the posterior mean of the clusters, and a small, effectively irrelevant, effect on its posterior standard deviation. This sensitivity analysis, tested on extreme priors, underlines the minor effect they have in our particular application. We do not expect this to be the case in all datasets, and in fact, special attention should always be put in this issue for this particular model. However, we feel comfortable with the relatively minor impact of the key hyperparameters on the posterior distribution for similar applications to the ones explored in this paper.
Econometrics 2016,4, 13 15 of 23 Figure 6presents the Heatmap displaying the probability that two observations of edges of nine hedge fund strategy returns, the average for all strategies, and the seven factors from Fung and Hsieh’s [52] model belong to the same cluster between January 1994 and June 2009. The figure shows very high probabilities (more than 80%) that the following edges of hedge fund strategies belong to the same cluster during most of the months of the period January 1994 to June 2009, particularly during the bull market periods of 1994–1999 and 2003–2007: •Short-bias and Event-Driven •Short-bias and Fixed-Income •Short-bias and Global Macro •Short-bias and Managed Futures •Emerging Markets and the bond market factor •Emerging Markets and the commodity lookback straddle •Event-Driven and the stock market factor •Event-Driven and the firm size factor •Event-Driven and the bond market factor •Fixed-Income and Managed Futures •Managed Futures and the firm size factor Edge Time 1-2 1-3 1-4 1-5 1-6 1-7 1-8 1-9 1-10 2-3 2-4 2-5 2-6 2-7 2-8 2-9 2-10 3-4 3-5 3-6 3-7 3-8 3-9 3-10 4-5 4-6 4-7 4-8 4-9 4-10 5-6 5-7 5-8 5-9 5-10 6-7 6-8 6-9 6-10 7-8 7-9 7-10 8-9 8-10 9-10 1994 1995 1996 1997 1998 1999 2000 2001 2002 2003 2004 2005 2006 2007 2008 Figure 6. (HeatmapEdgeChronoCSFBwithoutfactors file, 4 September 2011): Heatmap displaying the probability that two observations of nine hedge fund returns, the average for all strategies belong to the same cluster between January 1994 and June 2009. Source: Datastream, David Hsieh’s website, and own calculations. The notation used on the x-axis is as follows. CSFB/TREMONT Hedge Fund Strategies: 1: Convertible Arbitrage, 2: Short Bias, 3: Emerging Markets, 4: Equity Market Neutral, 5: Event-Driven, 6: Fixed-Income, 7: Global Macro, 8: Hedge Funds (all strategies), 9: Long-Short Equity, and 10: Managed Futures.
Econometrics 2016,4, 13 16 of 23 Furthermore, and as expected, the probability that the Equity Market Neutral strategy and any other strategy belongs to the same cluster (except for the edge Equity Market Neutral and Short Bias during the second half of the sample period) is fairly low. Fung and Hsieh [52] had already identified two breakpoints in factor exposures by hedge funds (September 1998, the collapse of LTCM, and March 2000, the beginning of the end of the Internet and technology bubble of the 1990s). This is consistent with hedge funds changing their strategies through time, this is, with alphas and betas being time-varying. Agarwal et al. [21] and Fung et al. [53] also use these breakpoints in their studies. Interestingly, [12] also report a structural break in the series of hedge fund returns in December of 2000. The figure also illustrates that the global financial crisis of late 2008 (August through December), and the crisis experienced by financial markets in 1998 as a result of the Russian default and the LTCM debacle (September through November of that year) were two extraordinary events, because the probability that any two hedge fund strategies belonged to the same cluster was high, and almost the same for all the possible edges of hedge fund strategies during those two sub-periods, as can be seen by the lines of yellow/orange colors during those periods. This outlines the dangers from a risk management point of view of the presence of apparently uncorrelated groups of strategies or portfolios, that become highly correlated at periods of stress, due presumably to synchronized unwinding of positions across different strategies. These results also confirm the findings by [52], who had already reported a structural break in hedge fund index returns in September of 1998. In a related article, [54] examine the extent to which hedge fund styles suffer from contagion. To that end, they use parametric and semi-parametric analysis and monthly hedge fund index data for the period 1990 to 2008. Contagion is defined as “correlation over and above what one would expect from economic fundamentals” (based on [55]). They find that hedge fund returns that fall in the bottom 10% of a hedge fund style’s monthly returns, cluster across styles. They also suggest that liquidity shocks to a number of contagion channel variables help explain hedge fund contagion. 6. Time-Varying Alphas and Betas Estimated Using iHMM-GGM In this section we ran the model of [1] explained in Section 2for 120,000 iterations and having descarted the first 20,000. The pseudo-code for computing the alpha and beta for each t = 1,...,186 consists in the following steps: 1. For i = 1,...,120,000. 2. Generate a value Lifor the number of clusters in iteration i. 3. Conditional on Li, label each data vector ztwith a specific cluster ξi tso that zξi t∼N(µξi t,K−1 ξi t )for t = 1,...,186. 4. Construct αξi t=µy ξi t + (Kyy ξt)−1Kyx ξi t µx ξi t and βξi t=−(Kyy ξt)−1Kyx ξi t for t = 1,...,186 from the distribution yt|xt∼Nµy ξt−(Kyy ξt)−1Kyx ξtnxt−µx ξto,(Kyy ξt)−1. 5. For a given t, average over all clusters ξi tthe values of αξi tand βξi tas αt=1/100, 000 ∑100,000 i=1αξi t and βt=1/100, 000 ∑100,000 i=1βξi t Figure 7present the estimated mean alphas for each strategy from highest to lowest (in percent per month) and for the whole sample period: (1) Global Macro (1.02%); (2) Long/Short Equity (0.83%); (3) Event Driven (0.79%); (4) Emerging Markets (0.70%); (5) Managed Futures (0.58%); (6) Convertible Arbitrage (0.57%); (7) Equity Market Neutral (0.52%); (8) Fixed-Income (0.35%); and (9) Short-Bias (−0.02%). Our results can be compared to those reported in a recent paper by [56], who also used a similar sample period (January 1995 to December 2009 versus our sample period of January 1994 to June 2009), and also worked with Fung and Hsieh’s seven factor model, but based their study on data from individual hedge funds available in TASS (we used CSFB/Tremont hedge fund indices) to estimate alpha applying OLS regressions. Ibbotson et al. [56] found a positive alpha for all strategies, although
Econometrics 2016,4, 13 17 of 23 it was only statistically significant for the following strategies: equity market neutral (annualized alpha equal to 2.38%), event driven (annualized alpha equal to 3.73%), fixed income (annualized alpha equal to 2.39%), and long/short equity (annualized alpha equal to 5.16%). The average hedge fund had a statistically significant alpha of 3.01% per year. In our study, average hedge fund alpha was substantially higher, around 0.75% per month or 9.4% when annualized. CSFB mean returns year Mean alphas (in monthly percentage) 1995 2000 2005 2010 -0.2 0.0 0.2 0.4 0.6 Arbitrage Short Bias Emerging Mkt Neutral Event Fixed Income Macro Hedge Fund Long Short Equity Futures Figure 7. (mu file, 4 September 2011): Estimated mean alphas (in percent per month) of the nine CSFB/TREMONT hedge fund strategies and the average for all strategies (January 1994 to June 2009). Source: Datastream, David Hsieh’s website, and own calculations. We also computed the difference between our estimated iHMM-GGM time-varying alphas for the nine CSFB/TREMONT hedge fund strategies and the average for all strategies, and the alphas estimated using OLS regressions during the sample period. We found that for all hedge fund strategies, except for Short-Bias (−0.6% per month), Managed Futures (−0.08% per month), and Fixed-Income (−0.04% per month), the difference between our estimated iHMM-GGM time-varying alphas and the estimated alphas using OLS regressions was positive on average. The highest difference was for Emerging Markets and Long/Short, at around 0.25% per month for each. The average hedge fund had an alpha that was around 0.13% per month (1.6% per year) higher when estimated using our iHMM-GGM procedure and the estimated alphas using OLS regressions. This implies that, over the time period we are using, estimations of alpha using OLS regressions such as those conducted in most of the prior research could underestimate the true alpha generated by the average hedge fund belonging to these three strategies. In Table 5we report the standard deviations of the 186 time-varying alphas for each of the nine strategies and for the average of all strategies. It is important to notice that computing the standard deviation over the 186 timevarying alphas only makes sense if we assume that these 186 alphas are each independent and identically distributed, something that we have seen is not the case since they are likely to be time-varying and not identically distributed (given the fact that they come estimated from our iHMM-GMM model). However, we still decided to report the standard deviations to have an idea of the variability of the alphas over all the 186 periods. Global macro strikes out as the strategy having the most volatile alpha, around twice the volatility of that for managed futures, the second most volatile strategy in terms of alpha. This finding may be explained, in part, by the effect on the performance of global macro that had extraordinary events such as those of September 1998, in which Long Term Capital Management, a hedge fund
Econometrics 2016,4, 13 18 of 23 belonging to this strategy and one of the largest hedge funds at the time, collapsed. As expected, the volatility of alpha for equity market neutral and long-short hedge funds is relatively low. The average performance of these strategies, especially in the case of equity market neutral, is more predictable and stable. The volatilities of alpha for event driven, convertible arbitrage and fixed-income strategies are close to the average volatility of all strategies. The volatility of alpha for emerging markets, a strategy that is in many respects similar to global macro, and short-bias hedge funds were around the lowest. Table 5. Alpha volatilites over all the 186 periods for all the 10 strategies (standard deviations, monthly percentage). Source: Datastream, David Hsieh’s website, and own calculations. CA SB EM EMN ED FI GM HF LS F 0.322 0.150 0.240 0.118 0.446 0.486 1.010 0.507 0.229 0.502 Figure 8shows the estimated iHMM-GGM monthly alpha for the Global Macro Strategy estimated using the seven factor model for the whole sample period. For illustrative purposes, we chose to present here these time-varying alphas only for one of the strategies, Global Macro, which, as was just explained, was found to have the most volatile alpha of the nine strategies. The figure shows the volatility exhibited by the alpha of Global Macro. This volatility of the alphas questions the constant alpha assumption implicit in the estimation through regression models. We also present, in Figure 9, and for illustrative purposes, the estimated iHMM-GGM beta of the Global Macro Strategy with respect to one of the seven factors of Fung and Hsieh [52], the commodity lookback straddle factor, using the whole sample period. It can be clearly observed the high volatility of beta for Global Macro with respect to this specific factor, particularly during the first half of the sample period. Macro iHMM Time alpha 1995 2000 2005 2010 -1.5 -1.0 -0.5 0.0 0.5 1.0 1.5 2.0 050 100 150 0.20 0.25 0.30 0.35 Macro Lin Reg Index alpha_linreg Figure 8. (CSGlobalMacroAlpha file, 4 September 2011): Estimated iHMM-GGM monthly alpha compared to alpha estimated from a linear regression for the Global Macro Strategy with respect to the seven factor model of Fung and Hsieh [52] (January 1994 to June 2009). Source: Datastream, David Hsieh’s website, and own calculations.
Econometrics 2016,4, 13 19 of 23 Macro Commodity lookback straddle beta 1995 2000 2005 2010 -0.5 0.0 0.5 Figure 9. 6b (CSGlobalMacroBeta file, 4 September 2011): Estimated iHMM-GGM beta of the Global Macro Strategy with respect to the commodity lookback straddle factor of Fung and Hsieh ([52], January 1994 to June 2009). Source: Datastream, David Hsieh’s website, and own calculations. Figure 10 shows the graphical models associated to three critical time-points (months) during the sample: September 1998 (collapse of Long-Term Capital Management), March 2000 (beginning of the end of the 1990s bull market), and September 2008 (fall of Lehman Brothers and beginning of the global financial crisis). These graphs were built by adding any edge that had a greater than 80% posterior inclusion probability for the respective timepoint. It is striking to note that the three graphs are exactly the same, this is, they depict the same relation among four hedge fund strategies (global macro, fixed income, long/short and managed futures) and the average of the hedge fund indices for all strategies. The graphs suggest that in times of market crisis the aforementioned strategies belonged to the same cluster as they exhibited a high degree of covariation and dependence. Our findings from Figure 10 could be very useful, for example, for funds of hedge funds (or any other allocators), as they indicate that in times of market distress hedge funds dedicated to these strategies may suffer heavy losses and, thus, allocators investing into these categories of hedge funds will not be as diversified as originally thought. We can compare our results to previous studies. For example, Khandani and Lo [57], who also used CSFB/TREMONT hedge fund indices, documented that the global macro strategy had a correlation of only between 25% and 50% with both the long/short and the fixed income strategies between the following two sub-periods: April of 1994 to December of 2000, and January of 2001 to June of 2007; and that the fixed income strategy had a correlation of less than 25% with the long/short strategy during the same sub-periods. However, Figure 10 strongly suggests that the global macro and fixed income strategies were both linked to the long/short strategy but not between them during those months of market stress, and that the managed futures strategy was linked to the other three strategies. In turn, [57] also find, contrary to our results, a strong correlation (greater than 50%) between 29 pairs of hedge fund strategies and during the same sub-periods. Thus, our results may also uncover the dangers of inferring the benefits of diversifying into specific hedge funds styles in times of market crisis when using traditional correlation analysis. In sum, the understanding of the joint distribution of hedge fund styles under different clusters provides a portfolio view of their marginal impact and behavior. For example, if one identifies a crisis cluster and observes the inferred distribution of returns in that cluster, one may be able to construct a portfolio that weights funds more appropriately, not according to their Sharpe or other performance measures,
Econometrics 2016,4, 13 20 of 23 but around the expected behavior of the distribution in times of crisis. Our findings are also important in the context of stress testing of portfolios during periods of market turmoil. FI HF LS F GM FI HF LS F GM FI HF LS F GM Figure 10. The left graph is for September 2008, the center graph is for March 2000 and the right graph is for September 1998. Source: Datastream, David Hsieh’s website, and own calculations. 7. Conclusions and Extensions We expand a novel statistical method to compute nonparametric dynamic alphas and betas, as well as returns clusters in a Bayesian fashion. This model provides a nonparametric, nonstationary, fully flexible approach to modelling sets of complex relationships between return series, regardless of the data generating process. We show through a simulation exercise that for both alpha and beta the estimation results based on the RMSE and MAD measures improve significantly upon popular statistical methods such as the regular OLS and DLMs, a special case of which being the Kalman filter. In a second exercise, we use hedge fund index returns from CSFB/Tremont for the period January 1994 to June 2009 and find that, using a Gaussian Graphical Model applied to Fung and Hsieh’s [52] seven-factor model, our estimated alphas are, on average, 0.75% per month. The average alphas of global macro, long/short equity and event driven were above the average for all strategies. The other strategies still had positive alphas, except for short-bias, although their average alphas were below the average for all strategies. Our estimated average alphas for all strategies are 0.13% per month (1.6% per year) higher than the alphas for all strategies estimated through OLS for the same period and, thus, our methodology reveals that average hedge fund alpha could be underestimated when measured using OLS. Furthermore, our estimated alphas and betas exhibit high volatility, particularly in periods which can be identified as times of stressful market events. Consistent with previous
Econometrics 2016,4, 13 21 of 23 research, hedge fund returns were found to be non-normal. We also found that estimated alphas for the global macro strategy were highly volatile. The alphas of equity market neutral, long/short equity and emerging markets exhibited the lowest volatilities of the nine strategies. These findings question the use of parametric and stationary (non dynamic) models to estimate alphas and betas. The Bayesian procedure that we use in this paper has as advantages that it describes hedge fund returns in a dynamic, nonstationary and nonparametric form and, at the same time, allows us to extract new nonparametric, nonstationary and dynamic alphas and betas. The procedure for measuring alpha that we propose in this paper was applied at the index level. Future research should investigate whether the same conclusions hold when individual hedge fund data is used. The use of individual hedge fund data also facilitates the measurement of performance persistence by grouping outperforming and underperforming hedged funds in each strategy and measuring their subsequent performance. Another extension would be to apply our procedure using other risk factors besides those of Fung and Hsieh [52] seven-factor model. In this regard, some authors have proposed specific factors for certain hedge fund strategies. For example, [58] propose a risk factor model for hedge funds dedicated to the fixed-income strategy. Finally, [21] document that mutual funds, which are catered mainly to small investors and are tightly regulated, as opposed to hedge funds, have recently begun offering funds that use trading strategies similar to those that are typical of hedge funds, as they include short sales and the use of derivatives, intended to take advantage of investment opportunities. These trading strategies should generate nonlinear payoffs and, thus, future research could also be extended to measuring mutual fund alpha using a more general procedure such as the one we propose here. The focus of this paper has been on a descriptive approach to the returns for decision-makers. Future work will focus on forecasting, in cases where enough persistence exists in the series to provide sufficient structure. In this paper we do not presume that there is a stationary distribution to revert to, but instead a nonstationary process that reflects the less predictable characteristics of the returns. One of the by-products of the methodology that we have presented in this paper is the time-varying probability that any two strategies belong to the same cluster. This information could be used by decision-makers in a number of different ways. For example, one approach would be through the understanding of the evolution of that relationship over time. However, another approach that could be more novel and meaningful, would be to analyze whether a set of strategies belong to the same cluster in key moments in time. Those key periods could be defined as periods with large negative returns, very large positive returns or any other particular feature that could be of relevance to the decision-maker, which in turn could help understand the relationship features between strategies with a higher granularity, as well as the expected marginal impact of additions to their portfolio in key periods. For example, Figure 10 showed that the graphical representation of the relationship between four specific strategies and aggregate hedge fund indices is the same in three key periods of time (September of 1998, March of 2000, and September of 2008). Whether the decision-maker expects that relationship to exist or expects further independence during times of market stress is outside the scope of this paper. However, the availability of this by-product brings in itself a significant number of areas of future research, including the analysis of the dynamics of these graphical structures over time and their impact on portfolio construction. Acknowledgments: We would like to thank Mark Jensen for his helpful comments and feedback; We also thank Tim Bollerslev and other participants at the Department of Statistical Science 25th Anniversary Conference at Duke University. Author Contributions: All authors contributed equally to the paper. Conflicts of Interest: The authors declare no conflict of interest. References 1. Rodriguez, A.; Lenkoski, A.; Dobra, A. Sparse covariance estimation in heterogeneous samples. Electron. J. Stat. 2011,5, 981–1014.
Econometrics 2016,4, 13 22 of 23 2. Titman, S.; Tiu, C. Do the Best Hedge Funds Hedge? Rev. Financ. Stud. 2011,24, 123–168. 3. Mamaysky, H.; Spiegel, M.; Zhang, H. Estimating the Dynamics of Mutual Fund Alphas and Betas. Rev. Financ. Stud. 2008,21, 233–264. 4. Ferson, W.; Schadt, R. Measuring Fund Strategy and Performance in Changing Economic Conditions. J. Financ. 1996,51, 425–461. 5. Patton, A.; Ramadorai, T. On the High-Frequency Dynamics of Hedge Fund Risk Exposures. J. Financ. 2013,68, 597–635. 6. Fung, W.; Hsieh, D.A. Hedge fund replication strategies: implications for investors and regulators. Bank Fr. Financ. Stab. Rev.—Spec. Issue Hedge Funds 2007,10, 55–66. 7. Kazemi, H.; Schneeweis, T. Conditional performance of hedge funds. CISDM Working Paper; Center for International Securities and Derivatives Market: Amherst, MA, USA, 2003. 8. Billio, M.; Getmansky, M.; Pelizzon, L. Dynamic risk exposures in hedge funds. Comput. Stat. Data Anal. 2012,56, 3517–3532. 9. Ahelegbey, D.F.; Billio, M.; Casarin, R. Bayesian Graphical Models for Structural Vector Autoregressive Processes. J. Appl. Econom. 2015, forthcoming. 10. Bianchi, D.; Billio, M.; Casarin, R.; Guidolin, M. Modeling Contagion and Systemic Risk; Technical Report, 2015. Available online: http://ssrn.com/ abstract=2537986 (accessed on 22 February 2016). 11. Kosowski, R.; Naik, N.; Teo, M. Do hedge funds deliver alpha? A Bayesian and bootstrap analysis. J. Financ. Econ. 2007,84, 229–264. 12. Kosowski, R.; Timmermann, A.; Wermers, R.; White, H. Can mutual fund “stars” reallypick stocks? New evidence from a bootstrap analysis. J. Financ. 2006,61, 2551–2595. 13. Pastor, L.; Stambaugh, R. Investing in equity mutual funds. J. Financ. Econ. 2002,63, 351–380. 14. Fung, W.; Hsieh, D.A. The Risk in Hedge Fund Strategies: Theory and Evidence from Trend Followers. Rev. Financ. Stud. 2001,14, 313–341. 15. Mitchell, M.; Pulvino, T. Characteristics of Risk and Return in Risk Arbitrage. J. Financ. 2001,56, 2135–2176. 16. Agarwal, V.; Naik, N. Risk and portfolio decision involving hedge funds. Rev. Financ. Stud. 2004,17, 63–98. 17. Baks, K.P.; Metrick, A.; Wachter, J. Should investors avoid all actively managed mutual funds? A study in Bayesian performance evaluation. J. Financ. 2001,56, 45–86. 18. Jones, C.S.; Shanken, J. Mutual fund performance with learning across funds. J. Financ. Econ. 2005,78, 507–552. 19. Avramov, D.; Werner, R. Investing in Mutual Funds when Returns are Predictable. J. Financ. Econ. 2006, 81, 339–377. 20. Busse, J.; Irvine, P. Bayesian alphas and mutual fund persistence. J. Financ. 2006,61, 2251–2288. 21. Agarwal, V.; Boyson, N.; Naik, N.Y. Hedge Funds for Retail Investors? An Examination of Hedged Mutual Funds. J. Financ. Quant. Anal. 2009,44, 273–305. 22. Dempster, A. Covariance selection. Biometrics 1972,28, 157–175. 23. Aguilar, O.; West, M. Bayesian Dynamic Factor Models and Portfolio Allocation. J. Bus. Econ. Stat. 2000, 18, 338–357. 24. Carvalho, C.M.; Lopes, H.F.; Aguilar, O. Dynamic stock selection strategies: A structured factor model framework (with discussion). In Bayesian Statistics 9; Bernardo, J.M., Bayarri, M., Berger, J.O., Dawid, A., Heckerman, D., Smith, A., West., M., Eds.; Oxford University Press: Oxford, UK, 2011; pp. 69–90. 25. Fama, E.F.; French, K.R. The Cross-Section of Expected Stock Returns. J. Financ. 1992,47, 427–465. 26. Fama, E.F.; French, K.R. Common risk factors in the returns on stocks and bonds. J. Financ. Econ. 1993, 33, 3–56. 27. Carhart, M.M. On persistence in mutual fund performance. J. Financ. 1997,52, 57–82. 28. Carvalho, C.M.; West, M. Dynamic Matrix-Variate Graphical Models. Bayesian Anal. 2007,2, 69–98. 29. Dobra, A.; Eicher, T.; Lenkoski, A. Modeling uncertainty in macroeconomic growth determinants using Gaussian graphical models. Stat. Methodol. 2010,7, 292–306. 30. Wang, H. Sparse Seemingly Unrelated Regression Modelling: Applications in Econometrics and Finance. Comput. Stat. Data Anal. 2010,54, 2866–2877. 31. Wang, H.; Reeson, C.; Carvalho, C.M. Dynamic Financial Index Models: Modeling Conditional Dependencies via Graphs; Technical Report; McCombs School of Business, University of Texas: Austin, TX, USA, 2011. 32. Letac, G.; Massam, H. Wishart distributions for decomposable graphs. Ann. Stat. 2007,35, 1278–1323.
Econometrics 2016,4, 13 23 of 23 33. Jones, B.; Carvalho, C.; Dobra, A.; Hans, C.; Carter, C.; West, M. Experiments in stochastic computation for high-dimensional graphical models. Stat. Sci. 2005,20, 388–400. 34. Lenkoski, A.; Dobra, A. Computational Aspects Related to Inference in Gaussian Graphical Models with the G-Wishart Prior. J. Comput. Gr. Stat. 2011,20, 140–157. 35. Dobra, A.; Lenkoski, A.; Rodríguez, A. Bayesian inference for general Gaussian graphical models with application to multivariate lattice data. J. Am. Stat. Assoc. 2011,106, 1418–1433. 36. Rasmussen, C.; Williams, C. Gaussian Processes for Machine Learning, 1 ed.; MIT Press: Cambridge, MA, USA, 2006. 37. Ferguson, T.S. A Bayesian Analysis of Some Nonparametric Problems. Ann. Stat. 1973,1, 209–230. 38. Ferguson, T.S. Prior Distributions on Spaces of Probability Measures. Ann. Stat. 1974,2, 615–629. 39. Teh, Y.W.; Jordan, M.I.; Beal, M.J.; Blei, D.M. Sharing Clusters among Related Groups: Hierarchical Dirichlet Processes. J. Am. Stat. Assoc. 2006,101, 1566–1581. 40. Van Gael, J.; Saatci, Y.; Teh, Y.W.; Ghahramani, Z. Beam sampling for the Infinite Hidden Markov Model. In Proceedings of the 25th International Conference on Machine Learning (ICML), Helsinki, Finland, 5–9 July 2008. 41. Fruewirth-Schnatter, S. Finite Mixture and Markov Switching Models; Springer: New York, NY, USA, 2006. 42. Antoniak, C. Mixtures of Dirichlet processes with applications to Bayesian nonparametric problems. Ann. Stat. 1974,2, 1152–1174. 43. Song, Y. Modelling Regime Switching and Structural Breaks with an Infinite Hidden Markov Model. J. Appl. Econom. 2014,29, 825–842. 44. Fox, E.; Sudderth, E.; Jordan, M.; Willsky, A. A Sticky HDP-HMM with Application to Speaker Diarization. Ann. Appl. Stat. 2011,5, 1020–1056. 45. Petris, G.; Petrone, S.; Campagnoli, P. Dynamic Linear Models with R, 2nd ed.; Springer: Heidelberg, Germany, 2009. 46. West, M.; Harrison, J. Bayesian Forecasting and Dynamic Models, 2nd ed.; Springer: Heidelberg, Germany, 1997. 47. Schneeweis, T.; Spurgin, R. Multi-Factor Models of Hedge Fund, Managed Futures, and Mutual Fund Return and Risk Characteristics. J. Altern. Invest. 1998,1, 1–24. 48. Liang, B. On the performance of hedge funds. Financ. Anal. J. 1999,55, 72–85. 49. Fung, W.; Hsieh, D.A. Performance Characteristics of Hedge Funds and Commodity Funds: Natural vs. Spurious Biases. J. Financ. Quant. Anal. 2000,35, 291–307. 50. Fung, W.; Hsieh, D.A. The Risk in Hedge Fund Strategies: Theory and Evidence from Trend Followers. Rev. Financ. Stud. 2001,14, 313–341. 51. Agarwal, V.; Naik, N.Y. Multi-Period Performance Persistence Analysis of Hedge Funds. J. Financ. Quant. Anal. 2000,35, 327–342. 52. Fung, W.; Hsieh, D.A. Hedge Fund Benchmarks: A Risk Based Approach. Financ. Anal. J. 2004,60, 65–80. 53. Fung, W.; Hsieh, D.A.; Naik, N.; Ramadorai, T. Hedge Funds:Performance, Risk and Capital Formation. J. Financ. 2008,63, 1777–1803. 54. Boyson, N.; Stulz, R.; Stahel, C. Hedge Fund Contagion and Liquidity Shocks. J. Financ. 2010,65, 1789–1816. 55. Bekaert, G.; Harvey, C.; Ng, A. Market Integration and Contagion. J. Bus. 2005,78, 39–69. 56. Ibbotson, R.G.; Chen, P.; Zhu, K.X. The ABCs of Hedge Funds: Alphas, Betas, and Costs. Financ. Anal. J. 2011,67, 15–25. 57. Khandani, A.; Lo, A. What Happened to the Quants in August 2007? J. Invest. Manag. 2007,5, 29–78. 58. Fung, W.; Hsieh, D.A. The Risk in Fixed-Income Hedge Fund Styles. J. Fixed Income 2002,12, 6–27. c 2016 by the authors; licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons by Attribution (CC-BY) license (http://creativecommons.org/licenses/by/4.0/).