scieee AI-readable full text Open interactive document viewer

“Wrong” skewness and endogenous regressors in stochastic frontier models: an instrument-free copula approach with an application to estimate firm efficiency in Vietnam

Haschka, Rouven E.

Abstract

EconStor is a publication server for scholarly economic literature, provided as a non-commercial public service by the ZBW.

Full text

Haschka, Rouven E. Article — Published Version “Wrong” skewness and endogenous regressors in stochastic frontier models: an instrument-free copula approach with an application to estimate firm efficiency in Vietnam Journal of Productivity Analysis Provided in Cooperation with: Springer Nature Suggested Citation: Haschka, Rouven E. (2024) : “Wrong” skewness and endogenous regressors in stochastic frontier models: an instrument-free copula approach with an application to estimate firm efficiency in Vietnam, Journal of Productivity Analysis, ISSN 1573-0441, Springer US, New York, NY, Vol. 62, Iss. 1, pp. 71-90, https://doi.org/10.1007/s11123-024-00722-6 This Version is available at: https://hdl.handle.net/10419/315409 Standard-Nutzungsbedingungen: Die Dokumente auf EconStor dürfen zu eigenen wissenschaftlichen Zwecken und zum Privatgebrauch gespeichert und kopiert werden. Sie dürfen die Dokumente nicht für öffentliche oder kommerzielle Zwecke vervielfältigen, öffentlich ausstellen, öffentlich zugänglich machen, vertreiben oder anderweitig nutzen. Sofern die Verfasser die Dokumente unter Open-Content-Lizenzen (insbesondere CC-Lizenzen) zur Verfügung gestellt haben sollten, gelten abweichend von diesen Nutzungsbedingungen die in der dort genannten Lizenz gewährten Nutzungsrechte. Terms of use: Documents in EconStor may be saved and copied for your personal and scholarly purposes. You are not to copy documents for public or commercial purposes, to exhibit the documents publicly, to make them publicly available on the internet, or to distribute or otherwise use the documents in public. If the documents have been made available under an Open Content Licence (especially Creative Commons Licences), you may exercise further usage rights as specified in the indicated licence. http://creativecommons.org/licenses/by/4.0/ Journal of Productivity Analysis (2024) 62:71–90 https://doi.org/10.1007/s11123-024-00722-6 “Wrong”skewness and endogenous regressors in stochastic frontier models: an instrument-free copula approach with an application to estimate firm efficiency in Vietnam Rouven E. Haschka 1,2 Accepted: 9 March 2024 / Published online: 3 April 2024 © The Author(s) 2024 Abstract Stochastic frontier models commonly assume positively skewed inefficiency. However, if the data speak against this assumption, sample-failure problems are often cited, but less attention is paid to economic reasons. We consider this phenomenon as a signal of distinctive population characteristics stemming from the inefficiency component, emphasizing its potential impact on evaluating market conditions. Specifically, we argue more generally that “wrong”skewness could indicate a lack of competition in the market. Moreover, endogeneity of model regressors presents another challenge, hindering the identification of causal relationships. To tackle these issues, this paper proposes an instrument-free estimation method based on Gaussian copulas to model the dependence between endogenous regressors and composite errors, while accommodating positively or negatively skewed inefficiency through simultaneous identification. Monte Carlo simulation experiments demonstrate the suitability of our estimator, comparing it with alternative methods. The contributions of this study are twofold. On the one hand, we contribute to the literature on stochastic frontier models by providing a comprehensive method for dealing with “wrong”skewness and endogenous regressors simultaneously. On the other hand, our contribution to an economic understanding of “wrong”skewness expands the comprehension of market behaviors and competition levels. Empirical findings on Vietnamese firm efficiency indicate that endogeneity hinders the detection of “wrong”skewness and suggests a lack of competitive market conditions. The latter underscores the importance of policy interventions to incentivize firms in non-competitive markets. Keywords Stochastic frontier ●Wrong skewness ●Endogenous regressors ●Copula function 1 Introduction The classical assumption in production stochastic frontier (SF) models is that inefficiency exhibits positive skewness, while the noise term is symmetrically distributed, resulting in the composite error, i.e., the regression residuals, having negative skewness. 1 However, in empirical applications, the residuals may exhibit positive skewness. 2 Although the methodological SF literature has proposed some approaches to handle skewness issues (e.g., Hafner et al. 2018), another substantial empirical challenge arises in the case of regressor endogeneity. For example, feedback mechanisms linking output to input can introduce endogeneity if producers adjust their inputs based on the inputs that yield the *Rouven E. Haschka [email protected] 1Chair of Business Analytics and Data Science, Zeppelin University, Am Seemoser Horn 20, D-88045 Friedrichshafen, Germany 2Institute of Strategy and Management, Corvinus University, Fóvám tér 8, H-1093 Budapest, Hungary 1Our focus is on the production SF model. In the cost SF model, the classical assumptions imply positive skewness of the regression residuals. 2Waldman (1982)first demonstrated that if the residuals from the SF model exhibit “wrong”skewness, i.e., positive under the production SF model, inefficiency variance effectively becomes zero. Consequently, efficiency scores tend to one, leading to false conclusions of high efficiency (Hafner et al. 2018; Parmeter and Racine 2013). Green and Mayes (1991) argue that this either indicates “super efficiency” (all firms in the industry operate close to the frontier) or the inappropriateness of the SF analysis technique to measure inefficiencies. 1234567890();,: 1234567890();,: highest marginal outcome (Siebert 2017). Endogeneity can also be introduced if firms know their own inefficiency and adjust their inputs accordingly, but this is unobserved to the analyst (Haschka and Herwartz 2022). While there are many economic reasons for endogeneity, potential economic understandings of “wrong”skewness are far less explored. The occurrence of “wrong”skewness has been attributed to poor samples (Almanidis and Sickles 2011;Hafneretal. 2018). Simar and Wilson (2009)confirm that small samples can lead to incorrect skewness measures despite correct skewness in the underlying population. To address this issue, researchers should prioritize increasing the sample size rather than making changes to the model specification (Almanidis and Sickles 2011, p. 201). However, what if “wrong” skewness occurs in large samples, such that explaining it by poor sampling or bad luck might not be justified? Methodologically, three reasons for “wrong”skewness have been considered in the literature: (i) asymmetry of idiosyncratic noise, (ii) dependence between idiosyncratic noise and inefficiency, or (iii) “wrong”skewness of the inefficiency component. In contrast, existing economic explanatory approaches for “wrong”skewness deal with the issue in general (for a recent summary, see Papadopoulos and Parmeter 2023), but do not delve into investigating where (i.e., from which component) it might originate. By assuming that “wrong”skewness is due to the inefficiency component, our contribution to the economic explanation aims to question the characteristics of the market. If “wrong”skewness is detected in a market that is expected to be competitive, we argue that it suggests that market competition is not generating adequate incentives for producers to improve their efficiency, thus indicating that the market may not be as competitive as expected. In the absence of competitive pressure that forces producers to increase efficiency to avoid falling behind competitors, we might observe many inefficient producers and only a few efficient ones. In such situations, assuming “wrong”skewness is due to the inefficiency term offers explanations for competition levels and market dynamics. Since any empirical detection of “wrong”skewness requires estimation, this depends on the performance and consistency of the estimation method used. Even if the underlying (production) model is correctly specified, endogeneity can lead to biased results, making it difficult to detect “wrong”skewness. Endogeneity can be loosely defined as regressor-error dependence, which is particularly important for SF models, as this dependence can stem from correlation with inefficiency or idiosyncratic noise (Griffiths and Hajargasht 2016; Mutter et al. 2013). While the linkage of endogenous covariate information and composite errors can not only lead to biased estimates for causal effects if applied methods build upon assumptions of regressor exogeneity, but the residual distribution may also be distorted, making it difficult to capture skewness correctly. The standard approach to handle the endogeneity problem in SF models is to use likelihood-based instrumental variable (ML-IV) estimation methods (Amsler et al. 2016; Haschka and Herwartz 2022; Kutlu 2010; Prokhorov et al. 2021; Tran and Tsionas 2013). However, a general drawback of ML-IV methods is that they rely on the availability of consensual outside information to construct the instruments. Against this background, we propose an estimation approach to handle endogenous regressors while simultaneously identifying “correct”or “wrong”skewness to assess market competition levels as an “empirical test of concept”, aiming to contribute to the economic understanding of “wrong”skewness. Given the often limited availability or weakness of suitable instrumental variables (IVs) in SF models, we conceptualize our approach as a methodological extension of the IV-free joint regression model using copulas introduced by Park and Gupta (2012). The core idea is to construct the joint distribution of the endogenous regressor and composite error, enabling simultaneous identification of “correct”or “wrong”skewness without the need for IVs. While copula-based endogeneity corrections have been extensively studied and successfully applied in classical SF models (Karakaplan and Kutlu 2015; Tran and Tsionas 2015; Tsionas 2017), their general applicability in SF models with “wrong”skewness remains unexplored. Our proposed approach builds upon and generalizes several existing methods, integrating models presented by Hafner et al. (2018), Park and Gupta (2012), and Tran and Tsionas (2015) into a unified framework. We conduct a series of Monte Carlo simulation experiments to demonstrate the suitability of the proposed estimator and compare it with alternative methods. Since there is currently no method capable of simultaneously addressing both endogeneity and skewness issues, our comparison includes methods that assume exogeneity (Hafner et al. 2018)or“correct”skewness (Tran and Tsionas 2013,2015). The empirical application aims to provide an unbiased understanding of the determinants of firm performance using data from 16,474 Vietnamese firms in 2015. Our findings lead to three major conclusions. First, we identify significant regressor endogeneity, challenging the estimation of firm productivity in Vietnam. Under the exogeneity assumption, marginal effects are overestimated, suggesting increasing returns to scale (RTS). However, accounting for endogeneity reveals constant RTS, aligning with the Vietnamese government’s priority of steady growth rates over rapid expansion. Second, the detection of “wrong”skewness is hindered by endogeneity, as explored within the Monte Carlo simulations. Despite lower-than-implied efficiency levels, accounting for endogeneity without addressing “wrong” skewness results in even lower efficiency levels, as the skewness might be falsely attributed to endogeneity. Third, 72 Journal of Productivity Analysis (2024) 62:71–90 empirical evidence points to moderate efficiency levels and “wrong”skewness, indicating a growing number of inefficient firms in the market, which contradicts the assumption of competitiveness, given our considerations that “wrong” skewness is due to market forces are valid. This lack of incentives to improve efficiency may be attributed to factors such as corruption and the constraints of the communist regime, hindering the establishment of liberal and competitive market conditions. Policy interventions are therefore necessary to create incentives for firms to optimize their processes and enhance efficiency. The paper begins by discussing the presence of “wrong” skewness in competitive markets, followed by a brief review of the (IV-free) SF literature addressing endogeneity. Section 3 introduces the model and discusses the copula approach to handle regressor endogeneity in SF models with “wrong”skewness when instrumental information is unavailable. In Section 4, we assess the finite sample performance of the proposed approach through Monte Carlo simulations. The empirical application is detailed in Section 5, followed by the concluding remarks in Section 6. 2 Background In this section, we first summarize potential explanations for the occurance of “wrong”skewness which have been discussed in the literature. Subsequently, we elaborate on economic perspectives that are based on the informative nature of detecting “wrong”skewness in (competitive) markets. Finally, we provide a literature overview focusing on endogeneity in SF models, with particular emphasis on instrument-free approaches. 2.1 Reasons for “wrong”skewness Since “wrong”skewness has primarily been considered an empirical phenomenon (Almanidis and Sickles 2011; Hafner et al. 2018; Waldman 1982), the prevailing reason for its detection is often attributed to small sample sizes (Simar and Wilson 2009, pp. 8–9). In cases where the true skewness is correct but “wrong”skewness is observed in a small sample, an inadequate sample size is typically identified as the cause. However, other reasons behind detecting “wrong”skewness have received less attention (for an excellent recent review, see Papadopoulos and Parmeter 2023). While there are some studies that detect “wrong”skewness in empirical applications (e.g., Almanidis and Sickles 2011; Hafner et al. 2018; Parmeter and Racine 2013), they do not delve into discussing potential characteristics in the population for this finding (one exception is Haschka and Wied 2022). On the one hand, certain characteristics of the data structure may contribute to “wrong”skewness. The asymmetry of the idiosyncratic error term can lead to multimodality in the distribution of efficiency scores, which in turn can cause “wrong”skewness (e.g., Badunenko and Henderson 2024; Bonanno et al. 2017; Horrace et al. 2024; Son et al. 1993). Additionally, unmodeled dependence between idiosyncratic noise and inefficiency can be another contributing factor (e.g., Bonanno et al. 2017; Bonanno and Domma 2022; Smith 2008). On the other hand, from an economic perspective, specific characteristics of the underlying population, such as unique features of the market in which firms operate, could also explain “wrong”skewness. As suggested by Papadopoulos and Parmeter (2023), when encountering skewness issues, researchers are advised to first consider the market’s specific attributes and potential peculiarities. Subsequently, they should reassess their arguments and determine whether the skewed result reflects an inherent characteristic of the population or is merely a consequence of a flawed sample. In reviewing these contributions, two points stand out. First, the literature that discusses methodological reasons for the occurrence of “wrong”skewness fails to relate them to characteristics in the population. Specifically, explaining market mechanisms that introduce dependence between idiosyncratic noise and inefficiency or cause asymmetry in the distribution of idiosyncratic noise requires a sound economic understanding. For instance, to what extent should unobserved production shocks simultaneously increase (or reduce) efficiency, and what accounts for the prevalence of positive shocks over negative ones (or vice versa)? Second, if we start addressing skewness issues more generally by discussing economic reasons, the question arises as to which of the model components requires an adjustment to reflect this peculiarity. Beyond the dependence within the composite error or asymmetry of idiosyncratic noise, skewness issues canalsobeattributedtoinefficiency. What insights can we derive regarding market characteristics by presuming that “wrong”skewness stems from the inefficiency component? 2.2 Economic explanations and market competition Stochastic frontier models often assume positive skewness in the inefficiency distribution to align with competitive market dynamics in which producers operate (Aigner et al. 1977). This assumption is grounded in economic reasoning, particularly in microeconomic production models that assume producers strive to optimize output given the inputs they use. In competitive markets, producers should minimize costs and maximize outputs, thus operating near the efficiency frontier due to competitive pressure that incentivizes efficiency improvements. Inefficiency is seen as deviations from the frontier, with highly inefficient producers likely exiting the market (Haschka and Herwartz 2020). Thus, specifying stochastic frontier models with positively skewed inefficiency distributions, such as the common half-normal distribution Journal of Productivity Analysis (2024) 62:71–90 73 (Kumbhakar et al. 2020), is justified for evaluating producer efficiency in competitive markets. What if only a small fraction of the firms attain a level of productivity close to the frontier while a large fraction attains considerable inefficiencies? According to Carree (2002), such a situation that is at odds with the assumption of positively skewed inefficiency might be found in industries characterized by alternating cycles of innovation and imitation, with periods in which a few firms innovate and improve their efficiency, while many firms remain inefficient, yielding “wrong”skewness. In subsequent periods, these firms imitate the innovations and efficiency levels converge, yielding correct skewness. However, these examples impose specific requirements on the market, such as the necessity for innovations to drastically and suddenly enhance efficiencies, the occurrence of these “leapfrog innovations”in a cyclical manner, and that innovation markets are distinguished by a clear distinction between innovation leaders and followers. Furthermore, Torii (1992) mentions that “wrong”skewness in inefficiency results from technological progress and the non-immediate replacement of assets within each producer, leading to misalignment when a few firms quickly renew their capital stock while the majority do so slowly. However, this implies that “wrong”skewness disappears in the long run once all firms have renewed their capital stock (Torii 1992). This suggests that increased competitive pressure leads to a more rapid renewal of capital stock by firms, resulting in a shorter duration for the phenomenon of “wrong”skewness to be observable. Both of these explanations implicitly assume that the inefficiency term is responsible for the skewness issues, albeit without explicitly labeling it as such. While they describe very specific situations, our explanatory approach more generally aims to outline that the existence of negative skewness in the inefficiency distribution contradicts the expectations of a competitive market environmen. Given the absence of other reasons for “wrong”skewness (e.g., poor samples), an observation that the majority of producers operate at lower efficiency levels, with only a few operating close to the frontier, could be explained by limited competition in the market, where producers may not face sufficient pressure to operate at their maximum efficiency levels (Haschka and Herwartz 2022). Papadopoulos and Parmeter (2023)mention markets with heavy regulation or entry barriers as potential reasons why the majority of established firms sit comfortably near higher inefficiency values without seeing a need to reduce inefficiency. 3 Moreover, factors such as limited market transparency (Møllgaard and Overgaard 2001), technological constraints (Ortega 2010), market imperfections (Cohen and Winn 2007), seller’s markets (Redmond 2013), or structural reasons (Haschka and Wied 2022) could contribute to this lack of competition. Unlike oligopolistic markets, the absence of competition does not necessarily result from market concentration or a small number of producers. In the absence of competitive market mechanisms, producers lack the necessary incentives to improve their efficiency levels. To illustrate these considerations and our contribution to the economic reasoning that explains “wrong”skewness by attributing it to the inefficiency component, Fig. 1depicts (potential) reasons and implications that emerge from skewness issues discussed in the literature. While no concerns are expressed if the skewness is correct, 4 methodological, economic, and small sample sizes have been identified as reasons for “wrong”skewness in the literature. While economic explanatory approaches deal with the problem in general and do not discuss which model component could be responsible, methodological explanatory approaches do not inquire about the economic causes. Likewise, attributing it solely to poor samples or data issues is an oversimplification, especially if it appears in larger samples (in small samples, however, it can never be ruled out with acceptable degree of certainty that the sample is poor). In contrast, we suggest a lack of competitive pressure and associated incentives for producers to enhance their efficiency levels as another potential explanation. Since our explanation is aimed at “wrong”skewness originating from the inefficiency term, existing economic explanations in the literature can be linked (Carree 2002; Papadopoulos and Parmeter 2023;Torii1992). Assuming the inefficiency term is the source, empirical identification of negative skewness likely provides valuable insights into the general competitive market dynamics (indicated by the green blocks in Fig. 1). A more detailed examination should then follow to determine why there are no incentives to increase efficiency in this market. 2.3 Endogeneity and the use of copulas in SF models While skewness issues require careful model specification when estimating SF models, endogeneity can greatly hinder the detection of causal effects. 5 Regressor endogeneity can have various causes. If producers have some a priori information on potentially inefficient output generation, it seems likely that the choice of production inputs is adjusted 3While this explanation could also apply to the scenario described by Torii (1992), in the example by Papadopoulos and Parmeter (2023), “wrong”skewness may also persist in the long run. 4As shown in the Monte Carlo simulations in Section 4, endogeneity can lead to correct skewness being detected in a sample even though the true skewness is “wrong”. This is relevant because “wrong” skewness can also be justified in the population (Haschka and Wied 2022). 5For recent approaches to cope with endogeneity in nonlinear models (including SFA), the reader may consult the special volume of the J Econom. entitled Endogeneity Problems in Econometrics, edited by Kumbhakar and Schmidt (2016). 74 Journal of Productivity Analysis (2024) 62:71–90 accordingly (Haschka and Herwartz 2022). In effect, the described unobserved correlation between production input factors and stochastic inefficiency is among the most common forms of endogeneity in the context of efficiency modeling (Cincera 1997). More generally, since output generation is typically seen as a reflection of input activities, successful output generation might also lead to further input activities, inducing endogeneity as a result of patterns of reverse causality. Furthermore, technology shocks that affect investment decisions provide a third origin of endogeneity. Since such shocks are unobserved to analysts, they manifest in model terms assessing productive efficiency (Haschka and Herwartz 2020). Endogeneity bias might also occur when firms respond to demand or supply shocks (that are unobserved to the analyst) by adjusting their inputs, such as the number of employees (Ehrenfried and Holzner 2019). For instance, global health shocks, energy crises, or political tensions might trigger unexpected hiring or investment decisions (Reeb et al. 2012). Lastly, the presence of omitted variables, such as subsidies large enough to have a significant impact on output generation, can also give rise to endogeneity bias. Traditional approaches for dealing with endogenous regressors in stochastic frontier settings often involve instrumental variable estimation (Amsler et al. 2016; Griffiths and Hajargasht 2016). These methods utilize exogenous instruments to exploit their informational content and typically employ two-stage-least squares (Amsler et al. 2016; Griffiths and Hajargasht 2016), ML-IV (Amsler et al. 2016; Haschka and Herwartz 2022), control functions (Centorrino and Pérez-Urdiales 2023), or GMM (Shee and Stefanou 2015; Tran and Tsionas 2013) for estimation. However, the validity of instruments remains debatable. Instruments may be scarce, weak, or even unavailable, prompting researchers to explore IV-free alternatives. The use of copulas has gained increasing attention in stochastic frontier settings, although many applications are not aimed at endogeneity corrections. Amsler and Schmidt (2021) identify three different motivations for the use of copulas in the SF literature: (i) allowing idiosyncratic noise and inefficiency to be correlated in an otherwise standard SF model (e.g., Amsler et al. 2016,2017; El Mehdi and Hafner 2014; Smith 2008; Wiboonpongse et al. 2015); 6 (ii) allowing dependence between different composite errors Fig. 1 Implications of detecting correct and “wrong”skewness that are discussed in the literature. The literature references include only studies that refer to “wrong”skewness. What we derive from an empirical detection of “wrong”skewness is shown in green 6This allows addressing one of the potential reasons for “wrong” skewness (see Fig. 1). Journal of Productivity Analysis (2024) 62:71–90 75 and/or other types of errors; for example, to model autocorrelation in panel data (e.g., Amsler et al. 2014; Das 2015; Lai and Kumbhakar 2020), or across different equations in a multi-equation model (e.g., Carta and Steel 2012; Haschka and Herwartz 2022; Huang et al. 2018); and (iii) allowing non-standard types of dependence between the errors in a multi-equation system (e.g., Amsler et al. 2021). The potentials arising from (i) and (ii) have led to the possibility of taking the endogeneity of regressors into account. That is, it allows for correlation between the regressors and idiosyncratic noise and/or inefficiency (Amsler et al. 2016,2017). Following instrumental variable theory, Amsler et al. (2016) assume that the endogenous regressor can be decomposed into a part that is correlated with the error and a part that is truly exogenous (i.e., the instrument). Because this decomposition introduces a new equation, this class of models may be seen as multiequation-type. They use the Gaussian copula to obtain the joint distribution of this correlated part, idiosyncratic noise, and the inefficiency term. Amsler et al. (2017) generalize this approach and allow for environmental variables to affect inefficiency. To avoid the assumption of decomposability of the endogenous regressors and therefore not belong to the class of multi-equation models, copula approaches directly model regressor-error dependence, and are increasingly explored in SF settings. This class of models uses copula functions to approximate the joint distribution of endogenous regressors and composite errors without requiring instruments. In the first step, data-driven cumulative distribution functions (cdfs) of endogenous regressors are obtained. These, along with an assumed distribution for composite errors, are used as plug-in estimates for the copula function in the second step, and estimates are derived based on the joint distribution. Tran and Tsionas (2015) directly construct this joint distribution using empirical cdfs and Gaussian copula (see also Tsionas 2017). Karakaplan and Kutlu (2015) rearrange the model proposed by Tran and Tsionas (2015) and show that targeting the joint distribution is not necessary because, with two-stage generated regressors, focusing on the marginal distribution of composite errors for ML estimation is sufficient. Papadopoulos (2021) develops a two-tier SF model to handle latent variables, building on the copula approach of Tran and Tsionas (2015). Note that in the first strand of literature, copulas are employed instead of a closed-form expression for the likelihood function to model the dependence between the error terms, yet instruments are still utilized. This distinction is crucial because this literature avoids many of the identification difficulties encountered by the second stream of literature. The identification problem the second strand faces arises if endogenous regressors have the same distribution as the errors, or if the distributions are very close. In that case, model identification without IVs breaks down because copulas fail to distinguish noise from variation due to endogenous regressors (Tran and Tsionas 2015). Assuming normality of the idiosyncratic noise component and the half-normal distribution for inefficiency is a natural choice, since these assumptions lead to the closed skew normal distribution (CSN) for composite errors (Domınguez-Molina et al. 2003; González-Farıas et al. 2004); a distribution that is well-defined parametrically. Since joint estimation using copulas requires both the cumulative distribution function (cdf) and the probability density function (pdf) of the composite error distribution, the use of the CSN distribution allows the approach to be implemented in a straightforward manner. The fact that little attention is paid to skewness issues in SF models becomes even clearer when reviewing these contributions, since extant studies entirely base endogeneity-robust SF modeling on the assumption of “correctly”skewed inefficiency. To our knowledge, regressor endogeneity and “wrong”skewness have not been simultaneously addressed so far. Therefore, we aim to offer a simple solution to this issue by building on previous work by Hafner et al. (2018), Park and Gupta (2012), and Tran and Tsionas (2015). 3 Copula-based handling of endogenous regressors under “wrong”skewness Consider the typical stochastic frontier model: yi¼x0 iβþz0 iδþviui |fflfflffl{zfflfflffl} ei ;i¼1;¼;n;ð1Þ where yiis the output of producer i,xiis L× 1 vector of exogenous inputs, ziis K× 1 vector of endogenous inputs, β and δare L× 1 and K× 1 vector of unknown parameters, respectively, viis a symmetric random error, uiis the onesided random disturbance representing technical efficiency, and the composite error is therefore ei=vi−ui. We assume that xiis uncorrelated with viand ui, but ziis allowed to be correlated with viand possibly with ui, and this generates the endogeneity problem. We also assume that uiand viare independent and leave the skewness of uiunrestricted. The discussion that follows can be easily extended for the case where the (exogenous) environmental variables are included in the distribution of ui(Battese and Coelli 1995; Haschka and Herwartz 2022). 3.1 “Wrong”skewness of inefficiency distribution According to standard SF practices, we assume that vi Nð0;σ2 vÞcaptures two-sided idiosyncratic noise. Following 76 Journal of Productivity Analysis (2024) 62:71–90 Hafner et al. (2018), we distinguish two cases for uiwhich characterize the shape of distribution of composite errors ei=vi−ui: ‘Correct’skewness : uiN0;1½Þ ð0;γ2Þ;γ>0ð2Þ ‘Wrong’skewness : uiN0;a0jγj½Þ ða0jγj;γ2Þ;γ<0ð3Þ The assumption in (2) is well-disseminated and describes “correct”skewness of ui(and thus also ei) because the density of uiis strictly decreasing in 0;1½Þ(Kumbhakar and Lovell 2003). 7 By contrast, “wrong”skewness is induced by (3), where a0≈1.389 is the non-trivial solution of ϕð0Þ Φð0Þ¼ a0þϕa0 ðÞϕð0Þ Φa0 ðÞΦð0Þand the density of uiis strictly increasing and bounded in [0, a0∣γ∣]. It is worth highlighting that expectations of both uiand eiremain unaffected by the sign of skewness (Hafner et al. 2018). Thus, inefficiency variance and sign of skewness are directly related because γ>0 (γ< 0) induces correct (“wrong”) skewness but E½uiand E½eiare not subject to the sign of γ. The density of ei=vi−uiis given by: ‘Correct’skewness : gþ eðeÞ¼2 σϕe σ Φeγ σσv  ;ð4Þ ‘Wrong’skewness : g eðeÞ¼ 1 σΦa0 ðÞΦð0ÞðÞ ϕea0γ σ  ΦAwþa0σ σv  ΦAw ðÞ hi ; Aw¼ea0γ σ γ σv; ð5Þ with σ2¼γ2þσ2 vand Regþ eðeÞde ¼Reg eðeÞde. Note that under “correct”skewness, eCSNð0;σ2;γ σvσ;0;1Þ, while as shown by Haschka and Wied (2022), under “wrong”skewness, it is eCSN12a0γ;σ2;γ=σ γ=σ  ;  a0σ 0  ;σ2 v0 0σ2 v  Þ. The shape of inefficiency distribution under “correct”and “wrong”skewness is shown in Panel (a) of Fig. 2, and the corresponding distributions of composite errors in Panel (b). For γ> 0, the distribution of u (e) has positive (negative) skewness, whereas for γ< 0 its skewness is negative (positive). Note that composite error distributions in both cases are only determined by σvand γ. Accordingly, we can distinguish “correct”and “wrong” skewness without the necessity to identify further parameters (Hafner et al. 2018). 8 3.2 Joint estimation using copulas Let F(z1,…,zK,e) and f(z1,…,zK,e) be the joint distribution and the joint density of (z1,…,zK) and e, respectively. In practice, F(⋅) and f(⋅) are typically unknown and hence need to be estimated. Following Park and Gupta (2012), we adopt a copula approach to construct this joint density. The copula essentially captures dependence in the joint distribution of endogenous regressors and composed errors. Let ωz;i¼ðFz1ðz1iÞ;¼;FzK ðzKiÞÞ0and ωe,i=G(ei;σv,γ) denote the margins ðωz;i;ωe;iÞ02½0;1Kþ1based on a Fig. 2 Densities of uand efor γ=2.5, i.e., correct skewness (dotted lines) and γ=−2.5, i.e., wrong skewness (solid lines); with σv=0.5. aDensity of u;bDensity of e 7In general, “correct”skewness in the production SF model means positive skewness of uiand in consequence negative skewness of ei=vi−uidue to symmetry of vi. 8The adopted one-sided distribution is parsimonious. However, other approaches that allow for a data-driven choice of correct or “wrong” skewness either involve the identification of multiple parameters that determine inefficiency distribution (see, e.g. Tsionas 2007, for Weibull inefficiency), or an a priori determination of the sign of skewness (COLS or MOLS). While Li (1996) argues that a one-sided error component with unbounded range always has a positive skewness, Johnson et al. (1995) shows that the two-parameter Weibull distribution can have positive and (small) negative skewness for specific parameter combinations. Journal of Productivity Analysis (2024) 62:71–90 77 probability integral transform. The F’s denote the respective marginal cumulative distributions functions of observed endogenous regressors and G(ei;σv,γ) is the cumulative distribution function of the CSN distribution for errors, which is subject to the sign of γ. Building on Tran and Tsionas (2015), we replace F1z1i ðÞ;¼;Fpzpi by their respective empirical counterparts in a first stage. Given observed samples of zji,j=1, …,p;i=1, …,n, we use the empirical cumulative distribution function of zj, i.e., ^ Fj¼1 nþ1Pn i¼11zji z0j  . 9 Using a Gaussian copula, ^ ξz;i¼ðΦ1ð^ Fz1ðz1iÞÞ;¼; Φ1ð^ FzK ðzKiÞÞÞ0, and ^ ξe;i¼Φ1ð^ Gð^ ei;^σv;^γÞÞ follow a standard multivariate normal distribution of dimension (K+1) with correlation matrix Ξ. 10 Then, the joint density can be derived as fðzi;eiÞ¼ 1 ffiffiffiffiffiffiffiffiffi detðΞÞ pexp 1 2 ^ ξz;i ^ ξe;i ! 0 Ξ1I  ^ ξz;i ^ ξe;i ! ! gðei;σv;γÞQ K k¼1 fzk zki ðÞ; ð6Þ where ξe,iand g(ei;σv,γ) is again subject to “correct”or “wrong”skewness. The copula density in the first row links the error and all explanatory variables to encode information about the entire dependence in the model whereas densities in the second row describe marginal behavior. The marginal densities fzk zki ðÞin (6) do not contain any parameter of interest and can be dropped when deriving the likelihood, since they enter as normalizing constants. Before deriving the likelihood, we brieflydiscussmodel identification. Under our setting, model identification requires the distribution of endogenous regressors to be different from that of the composite error (for a more detailed discussion on identification issues, see Haschka 2022b;ParkandGupta 2012). Accordingly, the model is identified as long as γis not zero (or very close to zero) and endogenous regressors are not normally distributed. However, model identification breaks down if both (i) γ=0 (such that the composite error is normal) and (ii) endogenous regressors are normal. In this case, the joint distribution of endogenous regressors and composite error is multivariate normal, which implies that E½ejzis a linear function, making it impossible to identify the linear effect δwithout instrumental variable information (Haschka 2022b). Thus, external instrumental information is needed to provide model identification (Tran and Tsionas 2013). Consequently, the identification problem has important implications when ∣γ∣→0. In this scenario, identification requires the endogenous regressors to be non-normally distributed. Therefore, in empirical applications, assessing the marginal distribution of endogenous regressors before estimation is a common approach in the empirical literature using copulabased identification (e.g., Datta et al. 2017; Haschka and Herwartz 2022;Papiesetal.2017). To obtain a simultaneous choice of “correct”or “wrong” skewness that is determined by the sign of γ, we follow Haschka (2024) and use an indicator function for the likelihood. As an alternative to using indicator function in the likelihood, Hafner et al. (2018) argue that choice of “correct” or “wrong”skewness can be made a priori by inspecting skewness of the OLS residual. However, in our approach, the sign of γis not predetermined but is instead estimated simultaneously with all other parameters. This approach is adopted because any prior determination of residual skewness could be influenced by (potential) endogeneity. Accordingly, we have To explicitly consider the case of only fully efficient firms, the likelihood also allows for γ=0. Here, the marginal distribution of the errors is a normal distribution with mean zero and variance σ2 v,itisξ0 e;i¼ei=σv. Note that our approach nests those by Hafner et al. (2018), Tran and Tsionas (2015), and Park and Gupta (2012). In case of exogeneity of all regressors, i.e. ρk=0∀k=1, …,K, the likelihood in (7) collapses to that in Hafner et al. (2018); in case of “correct”skeweness, i.e., Lðθjy;z;xÞ/1ðγ>0ÞQ n i¼1 1 ffiffiffiffiffiffiffiffiffi detðΞÞ pexp 1 2 ^ ξz;i ^ ξþ e;i ! 0 Ξ1I  ^ ξz;i ^ ξþ e;i ! 0 @1 Agþðei;σv;γÞ þ1ðγ<0ÞQ n i¼1 1 ffiffiffiffiffiffiffiffiffi detðΞÞ pexp 1 2 ^ ξz;i ^ ξ e;i ! 0 Ξ1I  ^ ξz;i ^ ξ e;i ! ! gðei;σv;γÞ þ1ðγ¼0ÞQ n i¼1 1 ffiffiffiffiffiffiffiffiffi detðΞÞ pexp 1 2 ^ ξz;i ^ ξ0 e;i ! 0 Ξ1I  ^ ξz;i ^ ξ0 e;i ! 0 @1 Aϕðei;σvÞ: ð7Þ 9The rescaling factor 1/(n+1) instead of 1/nensures that the empirical cumulative distribution is well bounded in (0, 1). 10 In general, any other copula that is capable of modeling multivariate dependency structures can also be used. According to Papadopoulos (2022), the Gaussian copula is most flexible and has many desirable properties. Furthermore, if the true dependence is different from what the Gaussian copula assumes, literature has demonstrated its robustness to capture various non-Gaussian dependencies (Becker et al. 2022;Haschka 2022b;ParkandGupta2012); although the true dependence should not be nonparametric (Haschka 2022a) or asymmetric (Papadopoulos 2022). 78 Journal of Productivity Analysis (2024) 62:71–90 decreasing returns to scale (RTSGMM =0.8949, CI = (0.818, 0.972); and RTScopula =0.891, CI =(0.8110, 0.971)). This suggests that, on average, as firms increase their inputs, the growth rate of output diminishes. In contrast, the proposed estimator paints a different picture, depicting a scenario of constant returns to scale. This is visible as the sum of coefficients in the production function is roughly 1 (RTSProposed =0.9906, CI =(0.906, 1.08)). 14 While increasing RTS suggest that it should be easy for firms to scale up, decreasing RTS urges firms to assess their expansion strategies critically, as indiscriminate scaling might not yield proportional increases in output, and considerations for optimizing resource allocation and operational efficiency are paramount. Constant returns to scale are indicative of a more consistent and predictable production process, allowing firms to plan and allocate resources with greater confidence. The question that now arises is which results seem most economically feasible for the case of Vietnam. Vietnam stood out as the sole emerging economy in Southeast Asia to avoid recession in 2009 amidst the global crisis. Moreover, Vietnam has demonstrated sustained growth rates over the past few decades (Cling et al. 2010). However, explaining increasing returns to scale would be difficult given the predominance of small firms in the dataset (O’Toole and Newman 2017). Although there is significant growth potential in the Vietnamese economy (Bai et al. 2019), the government prioritizes achieving stable and consistent economic growth over pursuing rapid growth at the expense of stability (Nguyen et al. 2018). This perspective favors constant returns to scale, as it means that the government’s economic goals can be attained while maintaining stability (Nghiem Tan et al. 2021). This suggests that MLE, GMM, and copula estimators may be flawed due to (remaining) endogeneity. 5.3 Endogeneity bias and efficiency levels Additional evidence in favor of endogeneity is provided by significant estimates of correlations between production inputs and errors when using copula and the proposed estimators. Specifically, the correlation coefficients are estimated as ^ ρe;log wages ¼0:2510 and ^ ρe;log assets ¼ 0:2442 for the proposed estimator. These metrics directly reflect the interdependence and provide valuable economic information by allowing an evaluation of how firms adapt to fluctuations in inefficiency or random disturbances, on average (Haschka and Herwartz 2022). Substantial positive correlation estimates suggest notable adjustments in inputs in response to implicit shifts in production technology or idiosyncratic shocks. It seems intuitive for both correlation estimates to be positive. For example, adverse external technological shocks are likely to diminish efficiency, indirectly leading to reduced output. At the same time, an increase in production inputs becomes necessary to maintain the output level. The distribution of efficiency scores is shown in Fig. 4. Considering firm efficiency, we find rather high mean firm efficiency when using MLE, with an average score of 0.8785. These results initially seem plausible, because it is in line with other efficiency levels documented in the literature (Le and Harvie 2010; Le et al. 2018; Nguyen et al. 2018; Tran et al. 2008;Vu2003). However, it should be mentioned that none of these studies consider potential regressor-endogeneity. Accounting for endogeneity while Table 4 Estimation results using MLE (Hafner et al. 2018), GMM (Tran and Tsionas 2013), copula (Tran and Tsionas 2015), and the proposed estimator MLE GMM Copula Proposed Est. SE Est. SE Est. SE Est. SE Const 1.140 0.0625 0.9814 0.0752 1.055 0.0800 1.241 0.0812 log wages 0.9102 0.0114 0.6211 0.0198 0.5994 0.0201 0.6493 0.0210 log assets 0.2428 0.0086 0.2738 0.0188 0.2916 0.0199 0.3413 0.0213 σv0.5918 0.0142 0.6561 0.0295 0.6290 0.0308 0.8993 0.0301 γ0.6345 0.0310 0.8181 0.0393 0.8544 0.0409 −0.8803 0.0404 ρe;log wages 0.3310 0.0391 0.2510 0.0404 ρe;log assets 0.1822 0.0365 0.2442 0.0361 Mean Efficiency 0.7985 0.4252 0.4516 0.6552 Sector Dummies Yes Yes Yes Yes Regional Dummies Yes Yes Yes Yes Standard errors of the copula-based estimators (copula and proposed) are obtained by means of bootstrap procedures with 1999 replications. Efficiency scores are calculated using the estimator by Jondrow et al. (1982). The skewness of the OLS residuals is −0.0981 14 Regarding the estimated RTS obtained by the proposed estimator, one can see that if a firm simultaneously increases its wages and assets by 1%, it can expect .9906% more revenues. Since this value is not significantly different from 1, we can conclude that returns to scale are constant. Journal of Productivity Analysis (2024) 62:71–90 85 assuming “correct”skewness through GMM and copula estimators substantially decreases the mean efficiency scores to 0.4252 (GMM) and 0.4516 (copula), respectively. While these values seem very low, the proposed approach reveals a mean efficiency of 0.6552, still indicating a considerable shortfall of about 0.35% from maximum feasible output of Vietnamese firms. This is in line with Haschka et al. (2023), who also consider potential endogeneity and find similarly low efficiency levels. Important factors identified by the literature on low efficiency scores in Vietnam are corruption and the level of local financial development (Haschka et al. 2023,2021). Previous studies have highlighted a direct correlation between corruption and inefficiency (Nguyen and Van Dijk 2012; Rand and Tarp 2012), as well as a negative impact of higher local financial development on firm efficiency in Vietnam (O’Toole and Newman 2017). These findings align with broader research indicating that while financial development tends to bolster technical efficiency in highly efficient economies, its effects are diminished or even adverse in less-efficient ones (Arestis et al. 2006; Rioja and Valev 2004). Another possible explanation for low-efficiency levels is the connection with the adjustment of inputs, i.e., as a reflection of endogeneity (Haschka and Herwartz 2022). If firms are aware of their own low inefficiency levels, and increase their inputs according to that, a positive (unobserved) input-inefficiency dependence might be present. Since the simulations show that a positive correlation leads to an overestimation of the efficiency levels (see also Tran and Tsionas 2015), and the proposed estimator indicates a positive dependence between inputs and errors, this could provide another explanation for the actually lower efficiency levels. A positive sign of both correlation estimates is intuitively reasonable. For instance, (adverse) external technological shocks are likely to reduce efficiency and result indirectly in less output. At the same time, the production inputs have to be simultaneously increased to retain the output level. Exemplifying such shocks, one might notice environmental regulations of the Vietnamese government (Ho 2015). Consequently, stronger production restrictions or sharper regulations might be considered as potential manifestations of endogenous technological shocks. While the OLS residuals point to correct skewness and MLE also favors the traditional SF specification, the proposed estimator indicates the presence of “wrong”skewness after accounting for endogeneity. Although MLE can also detect “wrong”skewness, the presence of endogenous regressors likely hinders that. The pronounced difference in the efficiency scores obtained by GMM, copula, and the proposed estimator is interesting insofar as the latter indicates “wrong”skewness, but at the same time delivers higher efficiencies. We assume that since GMM and copula cannot deal with skewness issues, this was falsely identified as endogeneity, and the efficiency values are therefore underestimated (as could also be shown in the simulations). 5.4 Competition and market characteristics The insights offered by the proposed estimator provide evidence supporting the presence of endogenous regressors and “wrong”skewness. While the former has already been emphasized in the existing empirical development literature, the latter has not yet been acknowledged. While correct skewness coupled with low efficiency scores would probably be accepted, “wrong”skewness gives reason to question the market conditions. As mentioned in Section 2, we argue that “wrong”skewness contradicts the assumption of a competitive market situation. Yet, unless alternative explanations for the “wrong”skewness can be ruled out, such as a poor sample or asymmetry in idiosyncratic noise, relying solely on an economic rationale may lack persuasiveness. The notion that “wrong”skewness is attributed to bad luck with the sample becomes less plausible given the substantial size of our dataset, which is representative for Vietnam (O’Toole and Newman 2017). While it would be conceivable that asymmetry exists in the two-sided noise term or this term may be correlated with inefficiency, a shift in the model specification towards asymmetric two-sided distributions or a dependency between inefficiency and noise would necessitate a sound economic rationale. Fig. 4 Distribution of efficiency scores obtained by the four estimators via TEi¼E½expf^ uig(see Eqns. (8) and (9)). aMLE; bGMM; cCopula; dProposed 86 Journal of Productivity Analysis (2024) 62:71–90 High efficiencies combined with the “wrong”skewness revealed by the proposed estimator, collectively cast doubt on the prevailing market conditions and raise concerns about competition levels. These findings might be attributed to a lack of incentives for firms to optimize their efficiency, resulting in many firms lacking the pressure to improve their operations. This situation can be attributed to various factors, despite the already mentioned corruption in Vietnam, burdens imposed by the communist regime, which hinder the emergence of liberal and competitive market conditions (Rand and Tarp 2012; Sahut and Teulon 2022), or a lack of market orientation of many companies (Evangelista et al. 2013). Despite extensive reforms in this regard, their effectiveness appears to be limited (Gupta et al. 2014; Tran et al. 2008). On the one hand, Vietnam established a central committee dedicated to combating corruption and enacted an anti-corruption law in 2005, followed by ratification of the UN Convention Against Corruption in 2009. Despite these proactive measures, the country’s standing in the Corruption Perception Index of 2018 remained relatively low, with a ranking of 117th among 180 countries. This position represented a decline of ten places compared to the previous year, 15 indicating ongoing challenges in addressing corruption effectively. On the other hand, a lack of market orientation of firms in Vietnam is often attributed to the influence of cultural, economic, and institutional characteristics. As noted by Evangelista et al. (2013), market orientation can have a significant impact on business efficiency, when its adoption is influenced by a combination of internal organizational dynamics and external market forces. Therefore, further policy interventions are necessary to provide firms with incentives to optimize their processes and enhance efficiency since the market alone fails to generate adequate incentives in this regard. Specifically, steps to combat corruption are already being taken in the right direction, while incentives to strengthen the market orientation of Vietnamese producers could prove advantageous and contribute further to the overall improvement of business processes. 6 Conclusion Under the traditional production SF specification, composite errors are assumed to have negative skewness. Violations of this assumption, commonly termed “wrong” skewness, have been highlighted in the literature on SF analysis (Choi et al. 2021; Curtiss et al. 2021; Daniel et al. 2019). While earlier discussions attributed such skewness issues to dataset peculiarities like small or poor samples (Almanidis and Sickles 2011; Hafner et al. 2018; Simar and Wilson 2009), recent studies increasingly delve into economic rationales (see Papadopoulos and Parmeter 2023, for a review). We contribute to this discourse by outlining that when assuming that “wrong”skewness stems from the inefficiency term, it might indicate a lack of market incentives, leading producers to perceive no need to address their inefficiencies. While various methodological approaches exist to address skewness issues empirically (e.g., Hafner et al. 2018), our work contributes to the economic explanation of “wrong”skewness. However, existing studies have not considered the potential endogeneity of regressors. Endogeneity of production inputs is likely if firms adjust their resource allocation based on their own inefficiency levels, which may go unnoticed by the analyst. In the presence of endogeneity, examining firm processes becomes challenging for analysts, making it even more difficult to uncover underlying market characteristics. Against this background, the methodological scope of this paper is to propose an approach for estimating SF models while simultaneously identifying inefficiency skewness, considering potential endogeneity of regressors. Adapting the approach by Park and Gupta (2012), we utilize a Gaussian copula function to construct the joint distribution of endogenous regressors and composite errors, capturing mutual dependency without relying on instrumental variables. Our model distinguishes between “correct”and “wrong”skewness in one model without imposing a priori sign restrictions on inefficiency skewness or needing to identify additional parameters determining skewness. We assess the finite sample behavior of the proposed approach through Monte Carlo simulations. We analyze the determinants of firm performance in Vietnam using data from 16,474 unaffiliated firms in 2015. Our findings reveal three key insights. Firstly, endogeneity significantly affects firm productivity estimation, indicating constant returns to scale rather than increasing returns. This finding aligns with the government’s goal for steady growth. Secondly, endogeneity complicates the detection of “wrong”skewness, leading to lower efficiency estimates. Lastly, evidence of low efficiency and “wrong”skewness suggests a lack of competitiveness due to factors like corruption and government constraints. Policy interventions are crucial to incentivize firms to optimize processes and improve efficiency. Data availability The dataset generated during the current study is not publicly available as it contains proprietary information that the authors acquired through a license. Information on how to obtain it and reproduce the analysis is available from the corresponding author on request. 15 https://www.transparency.org/cpi2018. Journal of Productivity Analysis (2024) 62:71–90 87 Funding Open Access funding enabled and organized by Projekt DEAL. Compliance with ethical standards Conflict of interest The author declares no competing interests. Publisher’s note Springer Nature remains neutral with regard to jurisdictional claims in published maps and institutional affiliations. Open Access This article is licensed under a Creative Commons Attribution 4.0 International License, which permits use, sharing, adaptation, distribution and reproduction in any medium or format, as long as you give appropriate credit to the original author(s) and the source, provide a link to the Creative Commons licence, and indicate if changes were made. The images or other third party material in this article are included in the article’s Creative Commons licence, unless indicated otherwise in a credit line to the material. If material is not included in the article’s Creative Commons licence and your intended use is not permitted by statutory regulation or exceeds the permitted use, you will need to obtain permission directly from the copyright holder. To view a copy of this licence, visit http://creativecommons. org/licenses/by/4.0/. References Aigner D, Lovell CK, Schmidt P (1977) Formulation and estimation of stochastic frontier production function models. J Econom 6:21–37 Almanidis P, Sickles R (2011) The skewness issue in stochastic frontiers models: fact or fiction? In van Keilegom I, Wilson PW (eds) Exploring research frontiers in contemporary statistics and econometrics, Springer, Berlin/Heidelberg, DE, p 201–227 Amsler C, Prokhorov A, Schmidt P (2014) Using copulas to model time dependence in stochastic frontier models. Econom Rev 33:497–522 Amsler C, Prokhorov A, Schmidt P (2016) Endogeneity in stochastic frontier models. J Econom 190:280–288 Amsler C, Prokhorov A, Schmidt P (2017) Endogenous environmental variables in stochastic frontier models. J Econom. 199:131–140 Amsler C, Prokhorov A, Schmidt P (2021) A new family of copulas, with application to estimation of a production frontier system. J Prod Anal 55:1–14 Amsler C, Schmidt P (2021) A survey of the use of copulas in stochastic frontier models. In Parmeter C, Sickles RC (eds) Advances in efficiency and productivity analysis, Springer, p 125–138 Arestis P, Chortareas G, Desli E (2006) Financial development and productive efficiency in OECD countries: an exploratory analysis. Manchester Sch 74:417–440 Badunenko O, Henderson DJ (2024) Production analysis with asymmetric noise. J Prod Anal. 61:1–18 Bai J, Jayachandran S, Malesky EJ, Olken BA (2019) Firm growth and corruption: empirical evidence from Vietnam. Econ J 129:651–677 Battese GE, Coelli TJ (1995) A model for technical inefficiency effects in a stochastic frontier production function for panel data. Empir Econ 20:325–332 Becker J-M, Proksch D, Ringle CM (2022) Revisiting Gaussian copulas to handle endogenous regressors. J Acad Mark Sci 50:46–66 Bonanno G, De Giovanni D, Domma F (2017) The “wrong skewness” problem: a re-specification of stochastic frontiers. J Prod Anal 47:49–64 Bonanno G, Domma F (2022) Analytical derivations of new specifications for stochastic frontiers with applications. Mathematics 10:3876 Breitung J, Mayer A, Wied D (2023) Asymptotic properties of endogeneity corrections using nonlinear transformations (March 3, 2023). Available at arXiv. https://arxiv.org/abs/2207.09246 Carree MA (2002) Technological inefficiency and the skewness of the error component in stochastic frontier analysis. Econ Lett 77:101–107 Carta A, Steel MF (2012) Modelling multi-output stochastic frontiers using copulas. Comput Stat Data Anal 56:3757–3773 Centorrino S, Pérez-Urdiales M (2023) Maximum likelihood estimation of stochastic frontier models with endogeneity. J Econom 1:82–105 Choi K, Kang HJ, Kim C (2021) Evaluating the efficiency of Korean festival tourism and its determinants on efficiency change: parametric and non-parametric approaches. Tour Manag 86:104348 Cincera M (1997) Patents, R&D, and technological spillovers at the firm level: Some evidence from econometric count models for panel data. J Appl Econom 12:265–280 Cling J-P, Chi NH, Razafindrakoto M, Roubaud F (2010) How deep was the impact of the economic crisis in Vietnam? A focus on the informal sector in Hanoi and Ho Chi Minh City. Washington, DC, World Bank Cohen B, Winn MI (2007) Market imperfections, opportunity and sustainable entrepreneurship. J Bus Ventur 22:29–49 Curtiss J, Jelínek L, Medonos T, Hruška M, Hüttel S (2021) Investors’ impact on Czech farmland prices: a microstructural analysis. Eur Rev Agric Econ 48:97–157 Daniel BC, Hafner CM, Simar L, Manner H (2019) Asymmetries in business cycles and the role of oil prices. Macroecon Dyn 23:1622–1648 Das A (2015) Copula-based stochastic frontier model with autocorrelated inefficiency. Centr Eur J Econom Model Econom 7:111–126 Datta H, Ailawadi KL, Van Heerde HJ (2017) How well does consumer-based brand equity align with sales-based brand equity and marketing-mix response? J Mark 81:1–20 Domınguez-Molina JA, González-Farıas G, Ramos-Quiroga R (2003) Skew-normality in stochastic frontier analysis. Comun Téc. No. I 3–18 Ehrenfried F, Holzner C (2019) Dynamics and endogeneity of firms’ recruitment behaviour. Labour Econ 57:63–84 El Mehdi R, Hafner CM (2014) Inference in stochastic frontier analysis with dependent error terms. Math Comput Simul 102:104–116 Evangelista F, Thuy PN et al. (2013) Does it pay for firms in Asia’s emerging markets to be market-oriented? Evidence from Vietnam. J Bus Res 66:2412–2417 González-Farıas G, Domınguez-Molina JA, Gupta A (2004) The closed skew normal distribution. In Genton, M (ed) Skew elliplical distributions and their applications: a journey beyond normality, Chapman and Hall/CRC, p 25–42 Green A, Mayes D (1991) Technical inefficiency in manufacturing industries. Econ J 101:523–538 Griffiths WE, Hajargasht G (2016) Some models for stochastic frontiers with endogeneity. J Econom 190:341–348 Gupta R, Yang J, Basu PK (2014) Market efficiency in emerging economies–case of Vietnam. Int J Bus Glob 13:25–40 Haschka RE (2024) Endogeneity in stochastic frontier models with “wrong”skewness: copula approach without external instruments. Statistical Methods & Applications (forthcoming). https:// doi.org/10.1007/s10260-024-00750-4 Hafner CM, Manner H, Simar L (2018) The “wrong skewness”problem in stochastic frontier models: a new approach. Econom Rev 37:380–400 88 Journal of Productivity Analysis (2024) 62:71–90 Haschka RE (2022) Bayesian inference for joint estimation models using copulas to handle endogenous regressors. Available at SSRN. https://papers.ssrn.com/sol3/papers.cfm?abstract_id=4235194 Haschka RE (2022) Handling endogenous regressors using copulas: a generalisation to linear panel models with fixed effects and correlated regressors. J Mark Res 59:860–881 Haschka RE, Herwartz H (2020) Innovation efficiency in European high-tech industries: evidence from a Bayesian stochastic frontier approach. Res Policy 49:104054 Haschka RE, Herwartz H (2022) Endogeneity in pharmaceutical knowledge generation: an instrument-free copula approach for Poisson frontier models. J Econ. Manag Strat 31:942–960 Haschka RE, Herwartz H, Silva Coelho C, Walle YM (2023) The impact of local financial development and corruption control on firm efficiency in Vietnam: evidence from a geoadditive stochastic frontier analysis. J Prod Anal 60:203–226 Haschka RE, Herwartz H, Struthmann P, Tran VT, Walle YM (2021) The joint effects of financial development and the business environment on firm growth: evidence from Vietnam. J Comp Econ 50:486–506 Haschka RE, Wied D (2022) Estimating fixed-effects stochastic frontier panel models under “wrong’skewness with an application to health care efficiency in Germany. Available at SSRN https://papers.ssrn.com/sol3/papers.cfm?abstract_id=4079660 Ho H-A (2015) Business compliance with environmental regulations: evidence from Vietnam. Vietnam Econ Annu Meet 26:1–21 Horrace WC, Parmeter CF, Wright IA (2024) On asymmetry and quantile estimation of the stochastic frontier model. J Prod Anal 61:19–36 Huang T-H, Liu N-H, Kumbhakar SC (2018) Joint estimation of the Lerner index and cost efficiency using copula methods. Empir Econ 54:799–822 Johnson NL, Kotz S, Balakrishnan N (1995) Continuous univariate distributions, vol 2, John Wiley & Sons Jondrow J, Lovell CK, Materov IS, Schmidt P (1982) On the estimation of technical inefficiency in the stochastic frontier production function model. J Econom 19:233–238 Karakaplan MU, Kutlu L (2015) Handling endogeneity in stochastic frontier analysis. Available at SSRN 2607276 Kumbhakar SC, Lovell CK (2003) Stochastic frontier analysis, Cambridge University Press Kumbhakar SC, Parmeter CF, Zelenyuk V (2020) Stochastic frontier analysis: foundations and advances. In Ray RG, Subhash C Chambers, Kumbhakar SC (eds) Handbook of production economics, Springer Kumbhakar SC, Schmidt P (2016) Editors’introduction to the special volume “endogeneity problems in econometrics”. J Econom 190:209–211 Kutlu L (2010) Battese-Coelli estimator with endogenous regressors. Econ Lett 109:79–81 Lai H-P, Kumbhakar SC (2020) Estimation of a dynamic stochastic frontier model using likelihood-based approaches. J Appl Econometr 35:217–247 Le V, Harvie C (2010) Firm performance in Vietnam: evidence from manufacturing small and medium enterprises. Department of Economics, University of Wollongong, Working Paper 04-10. https://ro.uow.edu.au/commwkpapers/221 Le V, Vu X-BB, Nghiem S (2018) Technical efficiency of small and medium manufacturing firms in Vietnam: a stochastic metafrontier analysis. Econ Anal Policy 59:84–91 Li Q (1996) Estimating a stochastic production frontier when the adjusted error is symmetric. Econ Lett 52:221–228 Møllgaard HP, Overgaard PB (2001) Market transparency and competition policy. Rivista Polit Econ 91:11–64 Mutter RL, Greene WH, Spector W, Rosko MD, Mukamel DB (2013) Investigating the impact of endogeneity on inefficiency estimates in the application of stochastic frontier analysis to nursing homes. J Prod Anal 39:101–110 Nelder JA, Mead R (1965) A simplex method for function minimization. Comput J 7:308–313 Nghiem Tan L, Hau Long L, TRAN TVT (2021) Determinants of technical efficiency of microenterprises in Vietnam. J Asian Financ Econ Bus 8:829–838 Nguyen DP, Ho VT, Vo XV (2018) Challenges for Vietnam in the globalization era. Asian J Law Econ 9:20180002 Nguyen TT, Van Dijk MA (2012) Corruption, growth, and governance: private vs. state-owned firms in Vietnam. J Bank Financ 36:2935–2948 Ortega MJR (2010) Competitive strategies and firm performance: technological capabilities’moderating roles. J Bus Res 63:1273–1281 O’Toole C, Newman C (2017) Investment financing and financial development: evidence from Vietnam. Rev Financ 21:1639–1674 Papadopoulos A (2021) Measuring the effect of management on production: a two-tier stochastic frontier approach. Empir Econ 60:3011–3041 Papadopoulos A (2022) Accounting for endogeneity in regression models using Copulas: a step-by-step guide for empirical studies. J Econom Methods 11:127–154 Papadopoulos A, Parmeter CF (2023) The wrong skewness problem in stochastic frontier analysis: a review. J Prod Anal 61:1–14 Papies D, Ebbes P, Van Heerde HJ (2017) Addressing endogeneity in marketing models. In Leeflang P, Wieringa J, Bijmolt T, Pauwels K (eds) Advanced methods for modeling markets, vol 1 of International series in quantitative marketing, 18, Basel: Springer, p 581–627 Park S, Gupta S (2012) Handling endogenous regressors by joint estimation using Copulas. Mark Sci 31:567–586 Parmeter CF, Racine JS (2013) Smooth constrained frontier analysis. In Chen X, Swanson N (eds) Recent advances and future directions in causality, prediction, and specification analysis: essays in honor of Halbert L. White Jr, New York: Springer, p 463–489 Prokhorov A, Schmidt P (2009) Likelihood-based estimation in a panel setting: Robustness, redundancy and validity of copulas. J Econom 153:93–104 Prokhorov A, Tran KC, Tsionas MG (2021) Estimation of semiand nonparametric stochastic frontier models with endogenous regressors. Empir Econ 60:3043–3068 Rand J, Tarp F (2012) Firm-level corruption in Vietnam. Econ Dev Cult Change 60:571–595 Redmond W (2013) Three modes of competition in the marketplace. Am J Econ Sociol 72:423–446 Reeb D, Sakakibara M, Mahmood IP (2012) From the editors: endogeneity in international business research. J Int Bus Stud 43:211–218 Rioja F, Valev N (2004) Finance and the sources of growth at various stages of economic development. Econ Inquiry 42:127–140 Sahut J-M, Teulon F (2022) The challenges of the transition to market economies for post-communist East Asian countries. PostCommunist Econ 34:283–292 Shee A, Stefanou SE (2015) Endogeneity corrected stochastic production frontier and technical efficiency. Am J Agric Econ 97:939–952 Siebert RB (2017) A structural model on the impact of prediscovery licensing and research joint ventures on innovation and product market efficiency. Int J Ind Organ 54:89–124 Simar L, Wilson PW (2009) Estimation and inference in cross-sectional, stochastic frontier models. Econom Rev 29:62–98 Smith MD (2008) Stochastic frontier models with dependent error components. Econom J 11:172–192 Son TVH, Coelli T, Fleming E (1993) Analysis of the technical efficiency of state rubber farms in Vietnam. Agric Econ 9:183–201 Journal of Productivity Analysis (2024) 62:71–90 89 Torii A (1992) Technical efficiency in Japanese industries. In Caves RE (ed) Industrial efficiency in six nations, Cambridge: MIT Press, p 31–119 Tran KC, Tsionas EG (2013) GMM estimation of stochastic frontier models with endogenous regressors. Econ Lett 118:233–236 Tran KC, Tsionas EG (2015) Endogeneity in stochastic frontier models: copula approach without external instruments. Econ Lett 133:85–88 Tran TB, Grafton RQ, Kompas T (2008) Firm efficiency in a transitional economy: evidence from Vietnam. Asian Econ J 22:47–66 Tsionas EG (2007) Efficiency measurement with the Weibull stochastic frontier. Oxford Bull Econ Stat 69:693–706 Tsionas MG (2017) “When, where, and how”of efficiency estimation: Improved procedures for stochastic frontier modeling. J Am Stat Assoc 112:948–965 Vu QN (2003) Technical efficiency of industrial state-owned enterprises in Vietnam. Asian Econ J 17:87–101 Waldman DM (1982) A stationary point for the stochastic frontier likelihood. J Econom 18:275–279 Wiboonpongse A, Liu J, Sriboonchitta S, Denoeux T (2015) Modeling dependence between error components of the stochastic frontier model using copula: application to intercrop coffee production in Northern Thailand. Int J Approx Reason 65:34–44 90 Journal of Productivity Analysis (2024) 62:71–90