A semiparametric approach to estimating reference price effects in sales response models
Abstract
EconStor is a publication server for scholarly economic literature, provided as a non-commercial public service by the ZBW.
Full text
Aschersleben, Philipp; Steiner, Winfried J. Article — Published Version A semiparametric approach to estimating reference price effects in sales response models Journal of Business Economics Provided in Cooperation with: Springer Nature Suggested Citation: Aschersleben, Philipp; Steiner, Winfried J. (2022) : A semiparametric approach to estimating reference price effects in sales response models, Journal of Business Economics, ISSN 1861-8928, Springer, Berlin, Heidelberg, Vol. 92, Iss. 4, pp. 591-643, https://doi.org/10.1007/s11573-022-01083-y This Version is available at: https://hdl.handle.net/10419/309850 Standard-Nutzungsbedingungen: Die Dokumente auf EconStor dürfen zu eigenen wissenschaftlichen Zwecken und zum Privatgebrauch gespeichert und kopiert werden. Sie dürfen die Dokumente nicht für öffentliche oder kommerzielle Zwecke vervielfältigen, öffentlich ausstellen, öffentlich zugänglich machen, vertreiben oder anderweitig nutzen. Sofern die Verfasser die Dokumente unter Open-Content-Lizenzen (insbesondere CC-Lizenzen) zur Verfügung gestellt haben sollten, gelten abweichend von diesen Nutzungsbedingungen die in der dort genannten Lizenz gewährten Nutzungsrechte. Terms of use: Documents in EconStor may be saved and copied for your personal and scholarly purposes. You are not to copy documents for public or commercial purposes, to exhibit the documents publicly, to make them publicly available on the internet, or to distribute or otherwise use the documents in public. If the documents have been made available under an Open Content Licence (especially Creative Commons Licences), you may exercise further usage rights as specified in the indicated licence. https://creativecommons.org/licenses/by/4.0/
Vol.:(0123456789) Journal of Business Economics (2022) 92:591–643 https://doi.org/10.1007/s11573-022-01083-y 1 3 ORIGINAL PAPER A semiparametric approach toestimating reference price effects insales response models PhilippAschersleben1 · WinfriedJ.Steiner1 Accepted: 18 February 2022 / Published online: 9 April 2022 © The Author(s) 2022 Abstract It is well known that store-level brand sales may not only depend on contemporaneous influencing factors like current own and competitive prices or other marketing activities, but also on past prices representing customer response to price dynamics. On the other hand, non- or semiparametric regression models have been proposed in order to accommodate potential nonlinearities in price response, and related empirical findings for frequently purchased consumer goods indicate that price effects may show complex nonlinearities, which are difficult to capture with parametric models. In this contribution, we combine nonparametric price response modeling and behavioral pricing theory. In particular, we propose a semiparametric approach to flexibly estimating price-change or reference price effects based on store-level sales data. We compare different representations for capturing symmetric vs.asymmetric and proportional vs. disproportionate price-change effects following adaptation-level and prospect theory, and further compare our flexible autoregressive model specifications to parametric benchmark models. Functional flexibility is accommodated via P-splines, and all models are estimated within a fully Bayesian framework. In an empirical study, we demonstrate that our semiparametric dynamic models provide more accurate sales forecasts for most brands considered compared to competing benchmark models that either ignore price dynamics or just include them in a parametric way. Keywords Sales and price response modeling· Functional flexibility· Bayesian P-splines· Behavioral pricing· Prospect theory· Reference price JEL Classification M31 * Philipp Aschersleben [email protected] 1 Department ofMarketing, Clausthal University ofTechnology, Julius-Albert-Str. 2, 38678Clausthal-Zellerfeld, Germany
592 P.Aschersleben, W.J.Steiner 1 3 1 Introduction Sales response functions describe the relationship between the sales of a product (or an entire product category) as dependent variable and predictors that are believed to influence sales as independent variables. In brand sales models based on store-level data, these predictors typically represent prices of (substitute) brands and other marketing variables related to promotional activities (like displays and feature advertising) as well as trend or seasonal indicators. In this context, retailers and academic researchers face several challenges, for example how to process the usually large amount of information (predictors) to arrive at a parsimonious model, how to specify the functional relationships between metric predictors (like prices) and sales, and/or how to accommodate dynamic effects in the model. More specifically, several streams of research for modeling (store) sales response to price variations have developed over the last 40 years. Among them, researchers have focused on choosing the right functional form to adequately capture the relationship between (own or competitive) prices and sales. Non- or semiparametric regression models have been proposed here in recent years in order to capture strong and/or complex nonlinearities in price response that could actually be proven for frequently purchased consumer goods in many empirical studies and are difficult to handle with parametric models. A second stream has addressed price dynamics in (store) sales response models by adding variables for lagged (or even lead) price effects, by including price terms for reference price or price-change effects, or by considering time-varying price parameters. Reference price effects are more commonly studied with disaggregate consumer data (i.e., household-level data), and have been less frequently incorporated into response models based on aggregate sales data. Obviously, reference price effects are much more difficult to model with aggregate data compared to disaggregate data. In the latter case, purchase incidence, brand choice, and purchase quantity decisions of consumers can be more easily separated at the individual consumer or household level, and reference price effects can in principle influence all three decisions (although they are most popular in models that have its focus on brand choice only). Beyond reference price effects, other forms of dynamics like stockpiling, state dependence (brand loyalty) or customer holdover, or consumer learning have been shown to be also very relevant at the individual consumer or household level, see for example Neslin and vanHeerde (2009) and vanHeerde and Neslin (2017) for an overview. Sales response models lack this micro-foundation: both the three different consumer decisions and the possibly individually different dynamics are confounded in aggregate data, making it challenging to disentangle them (see, e.g., Neslin and Shoemaker 1989; vanHeerde etal. 2000, 2004). Using aggregate data, reference price effects can therefore only be interpreted for an aggregate of households in the sense that they refer to prices paid or observed in previous periods (e.g., weeks) rather than to prices paid or observed at previous individual purchase occasions. This may be less of a problem if goods like in the food sector are purchased on a regular (weekly) basis and are frequently advertised. On the other hand, if consumer goods have longer interpurchase times
593 1 3 A semiparametric approach toestimating reference price effects… household data can also suffer from some problems like for example a different composition of consumers across periods when modeling dynamic effects over time (e.g., weeks). And, modeling reference price effects with aggregate data has also its pros due to the greater managerial relevance of aggregate data compared to household-level data. Although household-level data are richer for explaining customers’ purchasing behavior (as indicated above), they have been often criticized by managers for their potential lack of representativeness, which may cause share estimates to differ from those based on store-level data (cf. van Heerde 1999, p. 21). While household-level data cover only a subset of all customers purchasing at a retail store, store-level data cover all these customers, which is a weighty argument from the perspective of a store manager in favor of using aggregate data. For a comprehensive discussion on the pros and cons of the different data types, see van Heerde (1999,pp. 20–21). As a result of the discussion above, it is however important to separate reference price effects from other dynamic effects like stockpiling or customer holdover when relying on aggregate data.1 In this paper, we combine nonparametric sales response modeling with the estimation of price dynamics where the latter are captured by reference price effects. The fact that no other study has tackled this frontier up to now can probably be explained by the much higher popularity of studying reference price effects with household-level data and the greater difficulties to disentangle different sources of dynamics with aggregate data, as discussed above. We try to fill this research gap and propose a semiparametric model to flexibly estimating price-change or reference price effects based on store-level sales data. We compare different representations for accommodating symmetric versus asymmetric and proportional versus disproportionate price-change effects following adaptation-level and prospect theory, and further compare our flexible autoregressive model specifications to parametric benchmark models. Since management decisions should be based on the model with the highest predictive performance (vanHeerde etal. 2002), our primary focus is on the predictive model performance rather than on solving the problem how to tease out different dynamic effects with aggregate data. Actually, focusing on prediction as our main goal relaxes the problem that it is more difficult to disentangle different dynamic effects with aggregate sales data compared to disaggregate consumer data. Nevertheless, we address this problem and separate reference price effects from other dynamic effects like stockpiling and customer holdover by including lagged sales as autoregressive model component. In an empirical study, we demonstrate that our semiparametric dynamic models can provide more accurate sales forecasts compared to competing benchmark models that either ignore price dynamics or just include them in a parametric way. The main benefit of the proposed model is therefore to help managers to predict sales better. It has been shown before for other models that semiparametric modeling can provide (much) better predictions than parametric modeling (see the literature review in Sect.2 below); as such, one main academic contribution of the paper is to 1 We thank two anonymous reviewers for stimulating the discussion on the pros and cons of the different data types to model reference price effects.
594 P.Aschersleben, W.J.Steiner 1 3 show that this also holds for reference price models. In addition, we discuss likely implications of our model for related optimal pricing decisions in our outlook onto future research perspectives at the end of the article. The rest of the paper is organized as follows. Section2 provides a compact review of the relevant literature on price response modeling based on aggregate data, reflecting the road from parametric to more flexible model specifications as well as the different options of addressing price dynamics, including reference price effects in particular. In Sect. 3, we introduce our Bayesian model estimation framework. Using scanner data for refrigerated orange juice brands sold by a large supermarket chain we compare different model specifications (nonparametric vs.parametric, dynamic vs.static, alternative options of specifying reference price effects) for predictive performance and discuss implications regarding estimated price elasticities in Sect.4. We conclude in Sect.5 with a summary of the most important findings, managerial implications, and an outlook on future research opportunities. 2 Literature review This section provides an overview of relevant literature for our proposed approach, referring to the functional form of price response models, the incorporation of price dynamics in such models, and the few approaches that have so far combined functional flexibility and the estimation of dynamic price effects in sales response models. Early approaches used strictly parametric modeling to estimate sales/price response functions, as a rule using the sales variable in logarithmic form (e.g., Hruschka 1997; Montgomery 1997; Foekens et al. 1999; Kopalle et al. 1999; vanHeerde etal. 2000, 2002; Hruschka 2006a, b; Andrews etal. 2008). Using log sales instead of sales enables to capture nonlinearities in sales response, however the observed data are still projected “into a Procrustean bed of a fixed parameterization” (Härdle 1990; as cited in vanHeerde 1999,p.28). In other words, parametric models only provide consistent estimates if the apriori assumed functional form is correct (e.g., Leeflang etal. 2000). The use of more flexible semi- or nonparametric models can help to overcome this problem, as these allow to ‘extract’ the shape of functional relationships directly from data without prior knowledge about the functional form (e.g., vanHeerde 2017). vanHeerde etal. (2001) have shown for several food categories that the use of a kernel regression approach can improve the predictive performance of brand sales models based on store-level data compared to parametric modeling. Hruschka (2006a) and Hruschka (2007) used neural nets (multilayer perceptrons) to capture nonlinearities in sales response and reported much better log marginal densities as well as high posterior model probabilities or superior cross-validated predictive densities compared to strictly parametric modeling for all brands considered. Other researchers proposed spline approaches to model sales response and could reveal strong nonlinearities in price effects which in addition were shaped very differently at the individual brand level. Kalyanam and Shively (1998) used stochastic cubic splines, Haupt and
595 1 3 A semiparametric approach toestimating reference price effects… Kagerer (2012) and Haupt etal. (2014) applied B-splines, Hruschka (2000) considered both B-splines and cubic smoothing splines, and Steiner etal. (2007), Weber and Steiner (2012), Lang etal. (2015), and Weber etal. (2017) employed Bayesian P-splines. Except for Kalyanam and Shively (1998) and Hruschka (2000), who focused on model fit (the first also used marginal posterior model probabilities, the latter applied AIC), the mentioned spline applications provided further evidence of (much) more accurate sales predictions when using nonparametric instead of parametric response models. These findings are of great importance since (store) managers should prefer the model specification with the best possible predictive performance (vanHeerde etal. 2002). More flexible specifications have the potential to work better and to provide superior forecasts than parametric ones if the sales data at hand include complex nonlinear relationships in price response which are difficult to ‘read out’ with parametric models (e.g., Lang etal. 2015). Beyond the choice of the right functional form, one can think about the incorporation of time-dependent (price) effects leading to dynamic instead of static sales or price response models. In the context of price promotions there is empirical evidence that lags or leads of prices can have an impact on current brand or current category sales volumes. vanHeerde etal. (2000) and vanHeerde etal. (2004) used leads and lags of price indices reflecting promotional price cuts with different types of promotional support, and reported significant and in parts also very substantial dynamic effects. Nijs etal. (2001) and Horváth and Fok (2013) accounted for price dynamics by fitting VARX models. Nijs et al. (2001) examined category-demand effects and found that the strong positive short-term effects of price promotions almost completely dissipate over time. Horváth and Fok (2013) analyzed cross-price effects and found evidence of preemptive switching in a way that a brand’s price promotion in one period can decrease a substitute brand’s sales in subsequent periods. Foekens et al. (1999) and Kopalle et al. (1999) proposed varying parameter models to account for dynamic (pricing) effects in store-level sales response models, both using the widespread multiplicative functional form for modeling price response. Foekens etal. (1999) reparameterized a brand’s own-price elasticity as to depend on cumulated previous price discounts (amount and time) for both the brand considered and competing brands, and reported that the magnitude and timing of preceding price cuts can have a significant impact on own-price elasticities at the current period. Kopalle etal. (1999) reparameterized own- and cross-price parameters as functions of geometrically-weighted averages of past discounts and in addition developed a normative model for related pricing decisions. Talking about lagged prices and dynamic price response modeling is further closely connected to the topic of reference prices. Adaptation-level theory, as proposed by Helson (1964), states that the perception of a new stimulus is performed relative to an ‘adaptation level’. Applied to a pricing context, the adaptation level for judging a newly observed price information is called reference price. In other words, the reference price constitutes an internal price standard of the consumer and works as the adaptation level the consumer compares the current price of a brand observed or paid to. Prospect theory also suggests that consumers evaluate alternatives based on a comparison to a standard or reference point, but further distinguishes how consumers value potential gains versus losses from making a decision (Kahnemann and
596 P.Aschersleben, W.J.Steiner 1 3 Tversky 1979). The value of a new price information therefore not only depends on the (absolute) difference to a reference point but also onwhether the price difference represents a gain or a loss for the consumer. Consequently, prospect theory offers an option to evaluate price-change effects much more differentiated compared to adaptation-level theory (see Sect.3.1 for a more detailed description of the value function underlying prospect theory). In the pricing literature, adaptation-level and prospect theory are most frequently applied in the context of (brand) choice modeling, i.e., in models that are based on disaggregate consumer data (for an introduction to this topic see Neslin and vanHeerde 2009, for a detailed literature review see Mazumdar etal. 2005 and Neumann and Böckenholt 2014, and for a recent application see Boztuğ etal. 2014 and Baumgartner etal. 2018). Exceptions are for example Kucher (1987) and Natter and Hruschka (1997), who incorporated reference price effects into market share models (i.e., using aggregate data). At this point, it is important to note that the reference price as a construct to model dynamic effects of past prices can be either operationalized as the price of the previous purchase occasion (referred to as pricechange effect) or determined via more complex reference price formation mechanisms based on several past prices (referred to as price-deviation effect), compare Kucher (1987). Gedenk (2002, p. 249) has provided an overview of studies for either stream, and Briesch etal. (1997) discussed different reference price formation mechanisms. Accordingly, because reference prices of consumers cannot be directly measured or be determined in aggregate sales data, some authors used either the price of the last period or the average of several previous prices as proxy for reference prices in aggregate response models (also seeGedenk 2002, pp.247–249). Referring to adaptation-level theory, Simon (1982,pp. 208–213) still a little earlier argued that it seems realistic to assume not only an absolute price effect but also a price-change effect on brand sales, i.e., that the price of the last period should have an impact on the response to the current period’s price. In a first step, he proposed two versions of linear price-change response models, one where a brand’s sales depend on the difference between its current price and its previous price, and another where the relative price change instead of the absolute price change was used as independent variable. Both linear price-change response models assume a symmetric and proportional sales response to price increases and decreases. In addition, he also proposed a hyperbolic sine sales response function to accommodate non-proportional (but still symmetric) price-change effects, with the relative price change as its argument. He motivated the use of this functional form using the same behavioral rationale as is inherent to the well-known Gutenberg price response model: small price changes may have only marginal (below average) effects on sales, while large price changes should yield disproportionately large sales effects (e.g., Hruschka 2000). Note that the three models did not include a contemporaneous (static) price effect, however Simon (1982,p.211) mentioned that the model could be extended accordingly in case of a sufficiently large data base (as it is given today with store-level scanner data). Later, Simon (1992,pp.253–254) also considered the possibility of an asymmetric price-change response by expanding his linear pricechange response model (with the absolute price difference as argument) to capture
597 1 3 A semiparametric approach toestimating reference price effects… sales effects from price increases (losses) and price decreases (gains) separately. He motivated the model extension with an own empirical study where he observed a significant price-change effect for price decreases (but not for price increases) on the one hand, as well as by referring to prospect theory, which vice versa suggests a higher price elasticity for price increases compared to price decreases due to loss aversion of consumers on the other hand. In the context of price assessment, Diller (2008, pp. 140–143) linked prospect theory to reference prices and proposes (among others) a reverse s-shaped decreasing function to capture asymmetric price-change or price-deviation effects. The shape of the function looks similar to the logistic price response function but shows a steeper progression for losses than for gains, following prospect theory. As an alternative, he suggested an s-shaped decreasing response function in order to accommodate the existence of possible lower and upper price thresholds. Consequently, the latter function is not in line with prospect theory but resembles the shape of the Gutenberg function with a flatter middle part on the one hand and much more elastic parts for larger deviations from the reference point on the other hand. Both functions allow to capture a disproportionate price response pattern. Importantly, Diller (2008,pp.360–361) pointed out that the choice of the right functional form depends on the empirical data at hand which makes it necessary to compare different parametric approaches in empirical applications. As mentioned above, using nonparametric estimation techniques can remedy this dilemma by letting the data determine the shape of price-deviation or price-change effects without apriori assumptions about functional forms. There are certainly pros as well as cons to decide upon whether to use only the price of the last period or a more complex reference price formation mechanism based on prices of several previous periods in an aggregate sales response model. Rinne (1981,pp.29–30) already used the previous price as reference price (in terms of the last seen price) in his sales response model, as well as Kucher (1987,p. 179) did due to the “exceptional position” of the previous price among all past prices. In addition, there is empirical evidence that individual consumers would not access price information that lies much beyond the immediate past purchase occasion, simply due to difficulties in accurately remembering prices further back (see Krishnamurthi etal. 1992, and the literature cited therein). If consumers buy products on their shopping trips on a weekly basis, this restricted memory capacity argument with its focus on the immediately last price as reference price is also applicable to aggregate data. Also, established reference price formation models for disaggregate data use periodical (weekly) updates for a brand’s reference price at the individual consumer level even if a consumer did not buy that brand in the last period (cf. Erdem etal. 2010). Accordingly, “updating reference prices only when households make purchases would underestimate the reference price” (Erdem et al. 2010,p.310). This assumes that consumers are monitoring brand prices over periods and therefore are aware of a brand’s price in the previous period, which is realistic for frequently purchased consumer goods (at least for segments of consumers). Based on these arguments, we focus on price-change response models using aggregate store-level sales data and the price of the last period as proxy for the reference
598 P.Aschersleben, W.J.Steiner 1 3 price of an aggregate of consumers shopping at a retailer, hence we use the more parsimonious option to operationalize reference prices. Beyond the approaches discussed above, only very few authors have explicitly considered prospect theory for modeling price effects in store-level sales response models. Based on the reference price model of Greenleaf (1995), Kopalle et al. (1996) assumed that demand for a brand is a linear function of price(s) and a pricedeviation effect, the latter which is operationalized with two additively separable terms representing gains and losses. The authors developed optimal dynamic pricing strategies and showed that when (a sufficiently large number of) consumers weigh losses stronger than gains, as suggested by prospect theory, every day low pricing (i.e., setting constant prices) is optimal for a retailer. Conversely, if (enough) consumers weigh gains stronger than losses, a hi-lo strategy (cyclical pricing) would be the optimal retailer strategy (for a similar result see Fibich etal. 2007). Assuming asymmetric reference price effects with loss-averse consumers, Fibich etal. (2003) demonstrated that for an infinite planning horizon the optimal pricing strategy ‘converges’ at a steady-state price, which turns out slightly lower than without considering reference price effects. Pauwels etal. (2007) proposed smooth transition regression models to explore threshold-based price elasticities and found evidence for larger threshold sizes for gains than for losses. VanHeerde etal. (2004) addressed both functional flexibility and price dynamics in brand sales models (with store-level data). Based on vanHeerde etal. (2000), who used leads and lags of price discount variables to capture price dynamics, and based on vanHeerde etal. (2001), who applied kernel regression to flexibly estimate price discount effects, the authors combined both features (leads and lags, local polynomial regression) in order to decompose the sales effect of promotions into the three different sources cross-brand effects, cross-period effects, and category expansion effects. Finally, Natter and Hruschka (1997) have been previously the only ones who estimated reference price effects within a flexible approach (via a neural network) and based on aggregate data, however their approach was directed on market share modeling rather than sales response modeling. To the best of our knowledge, no study so far has employed nonparametric regression to flexibly estimate asymmetric reference price (price-change) effects in store-level sales response models, and we attempt to fill this research gap in the literature with our study. Table 1 summarizes the literature on (store-level) sales response models with focus on estimating price effects discussed above, distinguishing between approaches that have addressed either functional flexibility, or price dynamics (in the form of using lead or lagged prices, reference prices, or time-varying parameters), or both features. In the Appendix A, we provide an overview of advantages of using nonparametric regression techniques in general and especially for capturing gain and loss effects and summarize further convenient properties of applying Bayesian P-splines as we do in this article. In the following, we present a semiparametric brand sales model which accounts for price dynamics via (asymmetric) price-change effects. Our model therefore combines reference price effects with functional flexibility. Our model specification is based on Weber et al. (2017), who assumed a brand’s sales to depend on the brand’s own price and prices of substitute brands, further
605 1 3 A semiparametric approach toestimating reference price effects… would perform very well, since it is nowadays well-established that price response is usually nonlinear for frequently purchased consumer goods, as considered here (see the literature review in Sect. 2). Nevertheless, the linear model represents a natural benchmark model, especially as it constituted the starting approach in the German-language pricing literature for modeling price-change effects (also compare Sect.2). We eventually moved the empirical results related to the linear model to the Appendix for the interested reader, because it actually performed much worse than the other models in our application, as expected. In the strictly parametric models, the unknown smooth price functions fj in (5) (and in (6)–(11), respectively) are replaced by parametric linear effects. The linear and exponential models differ only in the specification of the dependent variable (and consequently in the related autoregressive part), which is unit sales ( q ) in the former case and log unit sales ( log(q) ) in the latter case. The multiplicative model uses log unit sales as dependent variable and furthermore log-transformations for all price covariates. Accordingly, the three benchmark models can be expressed as follows (omitting the terms for store intercepts, promotional activities, and seasonality, which are identical across the models for simplification, to highlight the differences between the three parametric models): • linear: • exponential: • multiplicative: For the linear and exponential models, pref represents one of the four different specifications for capturing the price-change or reference price effect (simple price-dif- ference term vs.separate gain and loss variables, measured in absolute vs.relative terms) as in (8)–(11), and pci denotes the cross-price terms, as introduced in (5). For the multiplicative model, we use the ratio between the previous price and the current price, Δpt=pt−1∕pt , instead of the price difference as equivalent specification for the price-change effect (corresponding to the difference in log prices). Like for the price-difference term, as in (1), negative values of the log price-ratio term correspond to losses and positive values to gains. Note that an operationalization of the reference price term as in (2) is not reasonable for the multiplicative model, resulting in only two specifications for the price-change effect (simple log price ratio, separate log price ratios for gains and losses). Further note that like for the semiparametric model we again estimated each two nested variants of the three parametric benchmark models, once as static model without the reference price term and without the autoregressive part (static), and once as simpler dynamic model without (12) q st =𝛼1pst +𝛼2pref st +𝛿qs;t−1+ ∑i 𝛼c i p c i st + ⋯ (13) log (qst)=𝛼1pst +𝛼2pref st +𝛿log(qs;t−1)+ ∑i 𝛼c i p c i st + ⋯ (14) log (qst)=𝛼1log(pst)+𝛼2log(pref st )+𝛿log(qs;t−1)+ ∑i 𝛼c i log(p c i st )+ ⋯
606 P.Aschersleben, W.J.Steiner 1 3 the reference price term but including the autoregressive part (dyn-ar), compare Sect.3.3. Overall, we estimate and validate 22 different models (each 6 variants of the linear, exponential, and semiparametric models, as well as 4 variants of the multiplicative model), and all models are estimated within a fully Bayesian framework using the public domain software package BayesX (Brezger etal. 2005). Table2 summarizes the capabilities of the four types of models to estimate asymmetric or disproportionate price-change effects, depending on the specification of the reference price term(s). In case a simple price-difference (or price-ratio) term is used, the linear model only allows a symmetric and proportional price-change effect, the exponential and multiplicative models a symmetric and disproportionate effect, and the semiparametric model an asymmetric and disproportional effect. If separate gain and loss terms are used, the linear model enables an asymmetric but only proportional price-change effect, the exponential and multiplicative models an asymmetric and disproportionate effect, and the semiparametric model once again an asymmetric and disproportionate effect. The linear model is not in line with prospect theory, that suggests asymmetric and disproportionate price-change effects. The exponential model is able to reproduce disproportionately increasing values of gains and losses (if modeled separately) as suggested, e.g., by the Gutenberg function. The multiplicative model is still a bit more flexible than the exponential model and is not only able to capture disproportionately increasing gains and losses but also to mimic disproportionately decreasing returns to scale (again if gains and losses are modeled as separate terms), as suggested, e.g., by the logistic function. The semiparametric model can approximate any curvature from data including concave and convex shapes as well as (asymmetrically) s-shaped and reverse s-shaped patterns as provided, e.g., by logistic or Gutenberg functions. Table 2 Capabilities of the linear, exponential, multiplicative, and semiparametric models to capture asymmetries and nonlinearities (disproportionalities) of the price-change effect (operationalized via a single price-difference or price-ratio term, or via two separate terms for gains and losses) Model type Capability to estimate... ...asymmetric effects via... ...disproportionate effects via... ...price difference (price ratio) ...gains and losses ...price difference (price ratio) ...gains and losses Linear No Yes No No Exponential No Yes Yes Yes Multiplicative No Yes Yes Yes Semiparametric Yes Yes Yes Yes
607 1 3 A semiparametric approach toestimating reference price effects… 4 Empirical study This section describes the data we use in our empirical study, provides technical details on model estimation and model validation, and presents and discusses the corresponding results. 4.1 Data For our empirical analysis, we use scanner data for refrigerated orange juice sold by a large supermarket in the Chicago metropolitan area (Dominick’s Finer Foods). The data were provided by the James M. Kilts Center of the University of Chicago and contain weekly unit sales of 64 oz. packages for m∈{1, …,M=8} brands that can be divided into three price-quality tiers: the premium brand tier with two brands, the national brand tier with five brands, and the private label brand tier represented by the supermarket’s own store brand. The data were collected in s∈{1, …,S=81} stores of the supermarket chain and cover a time span of t∈{1, …,Ts} weeks each, where Ts∈[75, 88] . Descriptive statistics for weekly brand prices, market shares, unit sales, and the share of weeks with a different store-specific price compared to the previous week (referred to as “price changes”) are displayed in Table3. For a more parsimonious model specification, we capture cross-price effects at the tier level by using the sales-weighted mean price across the competing brands belonging to a considered price-quality tier per store and week (see, e.g., Kopalle etal. 1999). The three cross-price variables are denoted as p(prem) , p(nat) , and p(priv) in the following, referring to the premium, national, and store Table 3 Descriptive statistics for weekly brand prices, market shares, and unit sales *The share of price changes is calculated as the share of weeks with a different store-specific price compared to the previous week ( pt≠pt−1 ) Brand Retail price ($) / Share of price changes (%)* Market share (%) Unit sales Range Mean SD pt≠ p t−1 Mean SD Mean SD Premium brands Florida Natural [1.54, 3.35] 2.85 0.33 39.3 4.7 6.5 27.1 46.0 Tropicana Pure [1.29, 3.87] 2.96 0.57 47.3 12.3 13.5 74.8 97.8 National brands Citrus Hill [0.99, 3.07] 2.31 0.35 42.8 8.0 12.7 53.5 157.3 Florida Gold [0.99, 3.08] 2.19 0.40 43.9 5.2 7.9 33.3 63.4 Minute Maid [1.27, 3.17] 2.23 0.43 55.7 10.1 13.7 51.6 76.4 Tree Fresh [0.99, 2.69] 2.16 0.31 43.0 7.6 8.4 48.7 92.0 Tropicana [1.41, 2.99] 2.21 0.38 56.7 18.2 21.0 112.0 157.7 Private brand Dominick’s [0.99, 2.69] 1.76 0.42 47.9 34.5 25.7 314.3 540.5
608 P.Aschersleben, W.J.Steiner 1 3 brand tiers, and they are captured in (5) by the cross-price terms with index ci∈ {(prem), (nat), (priv)} . Since there is only one store brand (the retailer’s own brand), sales response models for this brand include only the two competitive price variables p(prem) and p(nat) . Note that using this more parsimonious specification of cross-price effects still allows the estimation of cross-effects across tiers, even if the information about prices of competing brands within a tier per store and week is concentrated as a sales-weighted average of the corresponding individual prices. This is important as previous research has shown that cross-price effects across tiers can be substantial; especially if higher-tier brands are temporarily reduced in price during a promotion, it can be expected that they can steal sales from lower-tier brands. For the dynamic models, we additionally need the lagged price (i.e., the price of the previous period, pt−1 ) to model price-change effects. For this reason, we eventually dropped the first week in the data for the estimation of all models (including the static model versions) to preserve comparability across the estimation results. Note that the share of weeks with a different store-specific price compared to the previous week ranges between 39.3% for Florida Natural and 56.7% for Tropicana. The data further provide information on the use of displays and odd prices by the retailer that we include in our models as control variables together with an indicator variable for a holiday in the current week to capture seasonality.3 Table4 provides a summary of all variables included in our models as well as an overview of the related effects estimated in the (most complex) semiparametric models as example. 4.2 Model estimation andvalidation Eilers and Marx (1996), who originally introduced the P-spline approach into the statistical literature, recommended to use between 20 and 40 equidistant knots within the range of observed levels of an independent variable of interest. That way, sufficient flexibility for the spline should be guaranteed (i.e., not less than 20 knots) and at the same time overfitting can be avoided (i.e., not more than 40 knots). For our empirical study, we use 20 knots which is also in line with Lang and Brezger (2004) and represents the default setting in the BayesX software. Strictly speaking, we generally use 20 knots for estimating all contemporaneous own- and cross-price effects. However, in order to provide a fair model comparison, we use only 10 knots each for estimating gain and loss effects in the models with two separate price-change terms (cf.(10) and (11)) whereas 20 knots for estimating the price-change effect in the models that use only one single price-difference (price-ratio) term (cf.(8) and (9)). This is reasonable since the gain and loss terms capture only one branch of the value function and therefore cover only about half of the data range of the price-difference term each. 3 Without loss of generality, seasonal effects could be addressed differently or in addition to the holiday covariate, for example by including quarter dummies. If available, one should also consider other promotional activities like feature advertising as well as prices at competing retailers as further control variables.
609 1 3 A semiparametric approach toestimating reference price effects… Table 4 Summary of variables used in our models and overview of estimated effects in the semiparametric reference price model as specified in (5) Variable Effect Description – 𝛽s Store-specific random intercept for store s pst f1 Unknown smooth nonlinear decreasing function of the brand’s own price pref st f2 Unknown smooth nonlinear increasing function capturing the reference price effect (see (6)–(11) for the different specifications of the reference price term pref ) pci st f c i Unknown smooth nonlinear increasing functions for cross-price effects, where prices of competing brands ( pci st ) are captured at the tier level and ci∈ {(prem), (nat), (priv)} denotes the three tiers of premium brands, national brands, and the private label (store) brand of the retailer v′ st 𝜸 v′ st contains variables for promotional activities and seasonality, and 𝜸 captures the corresponding parametric effects: - Ddisp,st - Indicator variable for own-display activities per week and store (indicates whether a particular brand was promoted inside the store) - S(tier) disp,st - Indicator variable for cross-display activities at the tier level per week and store, computed as sales-weighted share of display activities of the competing brands in the same price-quality tier (i.e., excluding own-display activities captured by Ddisp ) - D9,st , D99,st - Mutually exclusive indicator variables denoting whether a brand was offered at a 9-ending or a 99-ending price (odd pricing) - Dhldy,st - Indicator variable capturing a holiday in the current week log(qs;t−1) 𝛿 Autoregressive effect of the one-period lagged (log) unit sales – 𝜀st Gaussian error term with mean zero and variance 𝜎2
610 P.Aschersleben, W.J.Steiner 1 3 All models, the static and dynamic ones as well as the semiparametric and parametric ones, are estimated with BayesX using a Gibbs sampler to draw from the posterior distribution. We use a total of 12,000 iterations, with a burn-in period of 2000 iterations and a thinning value of 10 to minimize the autocorrelation of the samples, i.e., we finally saved D=1000 draws from the Markov chain. To account for parameter uncertainty, model performance (see (16)–(19) below and Sect.4.3.3) is assessed using the individual parameter (Gibbs) draws instead of using the posterior means of estimated parameters (e.g., Montgomery 1997; Hruschka 2006b; Lang etal. 2015). That is, predictions 𝜂 st for (log) unit sales ( log(qst) and qst , respectively) are calculated as the mean across 1000 draw-based predictions 𝜂 st,d : Since we are interested in predictions for a brand’s unit sales (instead of log unit sales) and especially to be able to compare the performance of models estimated in the log sales space (exponential, multiplicative, and semiparametric model) versus models estimated in the sales space (linear model), conditional mean predictions for unit sales are computed for the exponential, multiplicative, and semiparametric sales response models via qst =exp( 𝜂 st + 𝜎 2∕2) (see, e.g., Greene 2008,p.100). For the linear models, qst =𝜂 st . We compare the different models with regard to their prediction accuracy by using two error measures: the Root Mean Squared Sales Prediction Error ( RMSE , see, e.g., van Heerde etal. 2001) and the Root Median Squared Sales Prediction Error (RMedSE, see, e.g., Franses and Ghijsels 1999): In particular, we compute the Average Root Mean or Median Squared Sales Prediction Error ( ARMSE / ARMedSE ) in holdout samples based on a C -fold cross-valida- tion with C=10 folds. That is, we randomly split the total sample of observations for a brand into 10folds, use each time C−1=9 parts of the sample for model estimation, calculate the RMSE or RMedSE for the remaining part (holdout), and finally average over the 10individual RMSE / RMedSE values: (15) 𝜂 st = 1 D D ∑ d=1 𝜂 st,d . (16) RMSE = √ √ √ √ 1 S S ∑ s=1 1 Ts Ts ∑ t=1 1 D D ∑ d=1( qst,d−qst ) 2 , (17) RMedSE = √ √ √ √ √ med s=1,…,S t=1,…,Ts 1 D D ∑ d=1(qst,d−qst)2 . (18) ARMSE = 1 C C ∑ c=1 √ √ √ √ √ 1 S S ∑ s=1 1 T(c) s T(c) s ∑ t=1 1 D D ∑ d=1(q(c) st,d−q(c) st )2 ,
611 1 3 A semiparametric approach toestimating reference price effects… Note that all (lagged) prices are assumed to be always known both during model estimation and model validation, i.e., even though observations of a certain previous period may not be explicitly part of a respective estimation or holdout sample. We use the (A)RMedSE measure in addition to the more widespread (A)RMSE measure in order to correct for the possibility of huge misses due to outliers in holdout samples (Franses and Ghijsels 1999): suppose that, due to the random split of the sample, the range of values for one of the price variables in one of the holdout exercises would be larger than the corresponding range of levels across the other folds used for model estimation. In this case, sales forecasting for holdout observations at price levels in domains not covered by the estimation sample becomes an extrapolation. Unlike parametric functions, whose shape is globally affected or determined by one or only few parameters, (P-)splines fit the data locally which gives them their high flexibility to capture more complex shapes. This local fitting, however, makes splines or any other nonparametric regression technique at the same time more sensitive at the boundaries of the data range, since predictions outside the data range would be guided only by the nearest domain of the spline. As a consequence, it can happen that extrapolated sales predictions for ‘new’ price levels turn out exorbitantly high or low if the spline is very steep at the boundaries of the data range. Using the RMedSE as measure for the cross-validation procedure guides against this ‘extrapolation problem’ (as opposed to RMSE ). Note that extrapolation is generally not recommendable per se, but could theoretically appear here due to the random sample split. Since the ‘extrapolation problem’ did occur in very few instances when computing the RMSE measure, we removed the corresponding observations (21 observations representing as a rule isolated exceptionally low prices or high gains) from the data to preserve the comparability both between the different model specifications and the two performance measures, leaving a total of 54,841 observations for model estimation.4 4.3 Estimation results 4.3.1 Estimated effects In the following, we at first illustrate our estimation results for the semiparametric model using the brand “Citrus Hill” as an example. Figure2 shows plots of the abs-diff model given in (8), i.e., using absolute differences for estimating the pricechange effect. Depicted are the estimated mean effects for the price variables and the (19) ARMedSE = 1 C C ∑ c=1 √ √ √ √ √ med s=1,…,S t=1,…,T(c) s 1 D D ∑ d=1(q(c) st,d−q(c) st )2 . 4 The following observations were removed: pst < 1.50 for “Florida Natural” (4 observations) as well as the corresponding observations for “Tropicana Pure” where p(prem) st < 1.50 ; (Δrel p st ) + > 0.40 for “Florida Gold” (2 obs.); pst < 1.40 (3 obs.), (Δpst ) + > 1.40 (1 obs.), and p(nat) st > 2.80 (1 obs.) for “Tropicana”; pst < 1.20 (6 obs.) for “Minute Maid” (for a similar procedure, see also Lang etal. 2015).
612 P.Aschersleben, W.J.Steiner 1 3 lagged sales variable including 95% pointwise credible intervals as well as partial residuals (e.g., Fahrmeir etal. 2013,p.77), and the estimated effects for own-display use, tier-specific cross-displays, 9- and 99-ending prices, and the holiday covariate. Note that estimated median effects (blue lines, almost always hidden) coincide with the mean effects (red lines) for all price variables and the lagged sales variable. First, the estimated effects and effect sizes show face validity. The own-price effect turns out much stronger than any of the cross-price effects. Since “Citrus Hill” is a national brand it could further be expected that its unit sales are more strongly affected by brands of both the premium and national brand tier than by the store brand (as becomes evident from the nearly flat cross-tier price effect with respect to the store brand, see middle-right panel). Moreover, the own-price effect shows a threshold effect near the price of 2.00$ , and all price effects have very tight Fig. 2 Estimation results for the semiparametric abs-diff model using the brand “Citrus Hill” as example: estimated effects and partial residuals for price and lagged sales variables including 95% pointwise credible intervals (red lines, gray-shaded credible intervals), as well as estimated effects for display, price ending, and holiday covariates including error bars
613 1 3 A semiparametric approach toestimating reference price effects… confidence bands. 99-ending prices have a much larger effect size than other prices ending in 9, and a holiday in a week leads to a significant decrease in the brand’s unit sales in this week. Both the own-display effect and the cross-display effect of the competing national brands are not significant, as the credible intervals include the zero point, respectively. The more interesting part is the estimated dynamic price effect, displayed in the top-middle panel of Fig.2. Note that positive values of Δpt correspond to gains and negative values to losses (compare (1)). We observe that the loss part of the spline is rather flat (except for very large losses), while the gain part is steeply increasing over the whole price difference range to theright of the zero-point. That is, a higher current price compared to the price of the previous week shows only a small effect, while price cuts strongly stimulate sales. Consequently, customers seem to value gains more than losses for “Citrus Hill”, which contradicts the assumption of loss aversion as suggested by prospect theory (see Sect.3.1). Finally, the autoregressive model part displayed in the top-right panel shows a significant positive effect of one-period lagged sales on the sales of the current period ( 𝛿=0.11, p<0.05 ) and suggests a (moderate) customer holdover effect rather than stockpiling across the aggregate of consumers for “Citrus Hill”. Note that this lagged sales effect turns out small in comparison to the own-price and reference price effects. Remember that one would expect a significant negative parameter estimate for lagged sales in case of distinct stockpiling across (parts of) consumers.5 Fig. 3 Estimated price-change effects from the exponential (left panels), multiplicative (top-middle panel), and semiparametric models (right panels) for the brand “Citrus Hill”, capturing the price difference in absolute terms (abs-diff, top panels) or relative terms (rel-diff, bottom panels). See Fig.6 in the Appendix for a variant relating to the ’sales space’ 5 The bivariate correlation between log sales and lagged log sales is 0.22 for “Citrus Hill”.
614 P.Aschersleben, W.J.Steiner 1 3 In Fig.3, we focus on the dynamic model part and compare the estimated pricechange effects for (1)the exponential (left panels), multiplicative (top-middle panel), and semiparametric response models (right panels) and (2)for the two options to capture the price difference Δpst either in absolute monetary units (top panels, absdiff) or by a percentage change (bottom panels, rel-diff). Remember that we did not estimate a rel-diff version of the multiplicative model (see Sect.3.4). Independent of the specification of the price difference in absolute or relative terms, the advantage of using a flexible regression approach becomes obvious: the spline clearly fits the data much better than the two parametric models. Without loss of generality, the effect plots refer to the space where the models were estimated, i.e., the log sales space for all these models. Note that this implies that estimated price effects for the exponential (or log-linear) and multiplicative (or log-log) models turn out linear in the log-space but exponential in the sales space. To illustrate this, we additionally plotted the estimated price-change effects for all three types of models in the sales space, see Fig.6 in the Appendix. Here, we observe that the exponential and multiplicative models tend to a convex shape for absolute differences (with a slightly better fit of the multiplicative model for high gain values), but that both parametric models are far too inflexible to capture the strong kink for the gain effect inherent to the data near the upper bound of the pricedifference range. The price-change effect is determined by only one parameter estimate in both the exponential and the multiplicative model, which makes them rather inflexible compared to the spline model, at least for this kind of price variable (the Fig. 4 Estimated price-change effects from the exponential (left panels), multiplicative (middle panels), and semiparametric models (right panels) for the brand “Citrus Hill”, capturing gain and loss effects in absolute monetary units (abs-gl) with two separate terms. See Fig.7 for a variant relating to the ’sales space’ and Fig.8 for a variant with gains and losses defined in percentage terms (please find both figures in the Appendix)
621 1 3 A semiparametric approach toestimating reference price effects… Natural”, the specification of gains and losses in relative terms (rel-gl) turns out to be at least as good or even superior to measuring the price difference in monetary units (abs-gl) for improving the predictive model performance. Taking a look at the results for the linear sales response model in Table 9, we find that the static version of the exponential model (i.e., not addressing price or other dynamics at all) already outperforms the best dynamic linear specification. This suggests, that the linear sales response model is highly misspecified since it is not able to accommodate the expected nonlinearities in price response for frequently purchased consumer goods, like orange juice. For the multiplicative models, we see a parallel development in predictive model performance as for the exponential response models: modeling the reference price effect with a single price-difference term (abs-diff) is much less helpful or again even decreases predictive validity for some brands over the static multiplicative model compared to the use of separate gain and loss variables (abs-gl). Improvements in predictive validity from the latter dynamic model (abs-gl) range between −7 % for Florida Gold and −21 % for Tropicana Pure. Note that although the static multiplicative models clearly outperform their exponential counterparts, differences in ARMedSE values between the best dynamic exponential and multiplicative models are very small or even marginal, which is reflected by the fact that the multiplicative (exponential) model predicts better for five (three) brands. None of the two parametric model types is therefore superior when price dynamics are accommodated, both perform similarly well. The following findings are obtained for the semiparametric sales response model with flexibly estimated price effects. First, the static semiparametric model (i.e., ignoring price and other dynamics) always provides more accurate sales predictions than the best dynamic exponential or multiplicative model versions capturing the price-change effect via a single reference price term (abs-diff and rel-diff). Improvements in ARMedSE from accommodating functional flexibility range up to −11 % for the brand “Tree Fresh” here (percentages not displayed in the table). For “Tree Fresh”, the static flexible model even outperforms each of the dynamic exponential and multiplicative model versions (i.e., including the abs-gl and rel-gl models). Second, adding price dynamics further improves the predictive performance of the flexible approach, but contrary to the class of exponential or multiplicative models all four dynamic specifications (abs-diff, rel-diff, abs-gl, rel-gl) perform pretty close. Again, once price effects are accommodated flexibly it does not seem to make a great difference whether the price-change effect is captured by one single price-dif- ference term or two separate variables for gains and losses (as was already evident from our discussions of the estimated effects in Sect.4.3.1 and price elasticities in Sect.4.3.2). In particular, we find that accommodating reference price effects in the semiparametric sales response model leads to noticeable improvements in predictive accuracy over the static flexible model of a minimum of −5 % for all brands and dynamic specifications, and more than −10% up to −21 % for four out of the eight brands (“Dominick’s”: −10 %, abs-gl, −11 %, rel-gl; “Citrus Hill”: −11 %, rel-diff, rel-gl; “Tropicana Pure”: −13 %, abs-diff, abs-gl, −14 %, rel-diff, rel-gl; “Tropicana”: −18 %, abs-diff, −19 %, rel-diff, −20 %, abs-gl, −21 %, rel-gl).
622 P.Aschersleben, W.J.Steiner 1 3 Table 6 Out-of-sample predictive performance of the competing models evaluated by the average root median squared sales prediction error (ARMedSE) in holdout samples for each brand Quality tiers: Premium National Private Flor. Natrl. Trop. Pure Citrus Hill Flor. Gold Min. Maid Tree Fresh Tropicana Dominick’s Exponential static 6.06 14.64 10.37 10.14 12.38 8.51 26.84 96.49 – – – – – – – – dyn-ar 6.01 14.60 10.37 10.05 12.20 8.20 26.91 95.34 (−0.8%) (−0.3%) (±0.0%) (−0.9%) (−1.5%) (−3.6%) (+0.3%) (−1.2%) abs-diff 5.54 13.71 9.26 9.65 12.05 7.83 25.72 95.27 (−8.6%) (−6.4%) (−11%) (−4.8%) (−2.7%) (−8.0%) (−4.2%) (−1.3%) rel-diff 5.85 14.46 9.70 9.73 12.18 7.88 26.23 97.04 (−3.5%) (−1.2%) (−6.5%) (−4.0%) (−1.6%) (−7.4%) (−2.3%) (+0.6%) abs-gl 𝟓.𝟏𝟖 11.26 8.19 9.20 11.23 7.50 22.61 84.48 (−15%) (−23%) (−21%) (−9.3%) (−9.3%) (−12%) (−16%) (−12%) rel-gl 5.19 𝟏𝟏.𝟏𝟒 𝟕.𝟗𝟑 𝟗.𝟏𝟕 𝟏𝟏.𝟏𝟒 𝟕.𝟒𝟐 𝟐𝟏.𝟕𝟖 𝟖𝟐.𝟗𝟔 (−14%) (−24%) (−24%) (−9.6%) (−10%) (−13%) (−19%) (−14%) Multiplicative static 5.93 13.84 9.89 9.95 11.89 8.08 26.68 90.24 – – – – – – – – dyn-ar 5.94 13.77 9.82 9.87 11.81 7.86 26.75 89.52 (+0.2%) (−0.5%) (−0.7%) (−0.8%) (−0.7%) (−2.7%) (+0.3%) (−0.8%) abs-diff 5.55 13.30 8.90 9.62 11.60 7.55 25.11 90.46 (−6.4%) (−3.9%) (−10%) (−3.3%) (−2.4%) (−6.6%) (−5.9%) (+0.2%) abs-gl 𝟓.𝟐𝟔 𝟏𝟎.𝟗𝟖 𝟕.𝟖𝟗 𝟗.𝟐𝟔 𝟏𝟎.𝟖𝟏 𝟕.𝟐𝟔 𝟐𝟐.𝟐𝟔 𝟖𝟏.𝟕𝟑 (−11%) (−21%) (−20%) (−6.9%) (−9.1%) (−10%) (−17%) (−9.4%)
623 1 3 A semiparametric approach toestimating reference price effects… Table 6 (continued) Quality tiers: Premium National Private Flor. Natrl. Trop. Pure Citrus Hill Flor. Gold Min. Maid Tree Fresh Tropicana Dominick’s Semiparametric static 5.39 12.64 8.66 9.53 11.27 6.98 24.85 88.05 – – – – – – – – dyn-ar 5.35 12.49 8.75 9.50 11.19 6.85 24.88 88.04 (−0.7%) (−1.2%) (+1.0%) (−0.3%) (−0.7%) (−1.9%) (+0.1%) (±0.0%) abs-diff 5.04 10.99 7.85 8.78 10.27 6.61 20.37 79.88 (−6.5%) (−13%) (−9.4%) (−7.9%) (−8.9%) (−5.3%) (−18%) (−9.3%) rel-diff 5.09 𝟏𝟎.𝟗𝟑 𝟕.𝟔𝟕 𝟖.𝟕𝟏 𝟏𝟎.𝟏𝟕 6.60 20.12 79.50 (−5.6%) (−14%) (−11%) (−8.6%) (−9.8%) (−5.4%) (−19%) (−9.7%) abs-gl 4.98 11.00 7.82 8.80 10.23 𝟔.𝟓𝟗 19.94 79.13 (−7.6%) (−13%) (−9.7%) (−7.7%) (−9.2%) (−5.6%) (−20%) (−10%) rel-gl 𝟒.𝟗𝟕 𝟏𝟎.𝟗𝟑 7.68 8.74 10.19 6.60 𝟏𝟗.𝟔𝟐 𝟕𝟖.𝟑𝟔 (−7.8%) (−14%) (−11%) (−8.3%) (−9.6%) (−5.4%) (−21%) (−11%) Best semipar. vs. best expon. −4.1 % −1.9 % −3.3 % −5.0 % −8.7 % −11 % −9.9 % −5.5 % Best semipar. vs. best multipl. −5.5 % −0.5 % −2.8 % −5.9 % −5.9 % −9.2 % −12 % −4.1 % Best models per brand and type of model (i.e., exponential, multiplicative, and semiparametric) are marked in bold
624 P.Aschersleben, W.J.Steiner 1 3 Finally, the last rows in Table6 contrast the best flexible model with the best exponential and multiplicative models. Accordingly, improvements in predictive accuracy from semiparametric instead of nonlinear parametric modeling of price effects (own-price, cross-price, and price-change effects) lie between −3 % and −11 % (exponential model) or −3 % and −12 % (multiplicative model) for seven out of eight brands, with “Tree Fresh” and “Tropicana” benefiting most from accommodating functional flexibility, respectively. For “Tropicana Pure”, the exponential and multiplicative models with separate variables for gains and losses already do a good job and semiparametric modeling does not pay off here. The latter is important to note, since nonparametric techniques are only more powerful if nonlinearities (here nonlinear effects in price response) are too complex to be captured by parametric nonlinear models. The results for the second predictive performance measure, the Average Root Mean Squared Sales Prediction Error ( ARMSE ), closely resemble those for the ARMedSE measure in many aspects, which is why we have put the corresponding results in the Appendix (see Table8). First, including the lagged sales variable but no price dynamics (dyn-ar) only marginally improves or even worsens the predictive performance compared to the static model for all brands except “Dominick’s”, independent of the type of model. For “Dominick’s”, improvements are moderate ranging between −3 % and −4 % across model types. Second, the best (dynamic) exponential and multiplicative models again perform similarly well, and no recommendation can be made in favor of one or the other. The multiplicative (exponential) model performs somewhat better for five (three) out of the eight brands. Third, relative improvements in ARMSE for the best semiparametric models over the best exponential (multiplicative) models are similar than for ARMedSE and lie between −3 % and −15 % ( −3 % and −14 %) for seven out of eight brands, with “Tree Fresh” and “Tropicana” as before benefiting most from addressing functional flexibility. And fourth, once price effects are modeled flexibly the four different dynamic model specifications (abs-diff, rel-diff, abs-gl, rel-gl) come pretty close in their prediction accuracy. On the other hand, there are some differences in the patterns of the ARMSE versus the ARMedSE results which are noteworthy. For the two parametric models, using a single price-dif- ference term to capture the reference price effect (abs-diff, rel-diff) is no more consistently inferior to separating gain and loss effects with two individual price terms (abs-gl, rel-gl), as can be seen for the two brands “Tropicana” and “Dominick’s”. For “Minute Maid”, accommodating price dynamics does not pay off at all when measured by ARMSE (as opposed to ARMedSE ), independent from the type of model (exponential, multiplicative, semiparametric) and specification of the reference price term (abs-diff, rel-diff, abs-gl, rel-gl). For “Tree Fresh”, the improvements from the semiparametric model over the two parametric models are large ( −14 % compared to the exponential model, −10 % compared to the multiplicative model), but once price effects are modeled flexibly adding price dynamics does not provide further benefits. And finally, semiparametric modeling does not pay off at all for the premium brand “Florida Natural” here. In total, for parametric models (including the linear model) the picture is more clear when using ARMedSE as measure of predictive accuracy. Here, the predictive
625 1 3 A semiparametric approach toestimating reference price effects… performance always strongly benefited from capturing gains and losses with two separate covariates compared to a simple price-difference term. This clear implication in favor of separating gains and losses did not hold for all brands if ARMSE was employed to assess the predictive performance. Moreover, once price effects were modeled flexibly, adding a reference price term always further improves ARMedSE , while this did not apply to all brands when using ARMSE instead. Nevertheless, our findings clearly suggest the use of the semiparametric approach as method of choice to assess the predictive performance. First, the semiparametric model enabled better predictions for seven out of eight brands regardless of which measure was used, with sometimes very large improvements in predictive accuracy compared to all parametric models (see the brands “Tropicana” and “Tree Fresh” in Tables6 and 8). For the remaining brand, semiparametric modeling was not inferior to parametric modeling, respectively. Second, because of the latter aspect (if one cannot do worse with the flexible approach), one need not care to find the best parametric model at the individual brand level. And third, once price effects are modeled flexibly, it does not longer seem to make a great difference which dynamic specification is employed to adequately capture reference price effects. This is due to the high flexibility of P-splines (local fitting property) to uncover large differences (different curvatures) between gain and loss effects, even if these were not modeled separately but only via a single price-difference term. 4.3.4 A note onloss aversion The previous discussions have shown that the semiparametric approach is characterized by an at least as good or (considerably) better predictive performance than the parametric models considered, and that once (reference) price effects are modeled flexibly the decision whether price-change effects should be accounted for by a single price-difference term or two separate terms for gains and losses is obviously of minor importance. Since prospect theory is a prominent behavioral concept stating that consumers should weigh losses of a certain amount stronger than gains of the same amount (also compare Sects.2 and 3), and because we did not find lossaversion but instead gain-seeking behavior of consumers for all considered brands in the refrigerated orange juice brand category without exception, a closer look on loss-gain ratios seems worthwhile. Loss-gain ratio statistics are more widespread in a brand choice modeling context (e.g., Neumann and Böckenholt 2014), and have not been applied yet to sales response models to the best of our knowledge. For our parametric models with separate gain and loss terms (gl models), the gain-loss ratio can easily be determined by dividing the estimated parameter for losses ( 𝛼2L ) by the corresponding one for gains ( 𝛼2G ): (20) 𝜆 = 𝛼 2L 𝛼 2G ,
626 P.Aschersleben, W.J.Steiner 1 3 see (13) and (14) in connection with the abs-gl model in (10) for the derivation of the parameters.8 For the semiparametric models, the loss-gain ratio extends to a flexible nonlinear ratio based on the derivative of the estimated effect curve. The calculation is nonetheless rather similar to the simple loss-gain ratio for parametric models: we divide the derivative of the loss part by the one of the gain part and aggregate them to a weighted mean, with the number of observations supporting the particular points as weights: where X contains all price-difference observations in the data set (full range) or those contained in a particular predefined subrange, respectively. Table7 displays as examples the loss-gain ratios for the exponential, multiplicative, and semiparametric abs-diff and abs-gl models, referring to either the full range of price differences observed for the brands or to small, mid-sized, and large price differences at a more disaggregate subrange level (measured each time in monetary units). Note that for the parametric models with a single price-difference term (absdiff) the loss-gain ratio implicitly amounts to 𝜆=1 (therefore not included in the table), while this does not apply for the semiparametric model. Considered aggregated over the entire price difference ranges (see columns full rg.), the loss-gain ratios for the parametric and the semiparametric models turn out consistently (much) smaller than 1, which suggests that consumers buying refrigerated orange juice brands at Dominick’s Finer Foods stores are not loss-averse as a rule (as was already visible from the estimated price effect curves for losses and gains), contrary to what prospect theory postulates. Taking a look at the more disaggregated results within the separate price-difference intervals |Δ p t| ≤ 0.50 , |Δ p t|∈(0.50$, 1.00$] , and |Δ p t| > 1.00 confirms this finding in principle. Here, we find loss-aversion only very sporadically and only for some brands, and in no single case for (absolute) price differences below 0.50 . Additionally, the loss-gain ratios are lowest across the three subranges here for most brands implying that consumers weigh low gains much stronger than low losses compared to price differences larger than 0.50 (an exception is “Tree Fresh”, where consumers are least loss-averse for large price differences greater than 1.00 ). Clear loss aversion of consumers (i.e., with a loss-gain ratio greater than 2) is evident in only two cases: for large price changes of “Florida Gold”, and for mid-sized price changes of “Tree Fresh”. Note that the loss-gain ratios hardly differ between the two semiparametric model versions (absdiff vs.abs-gl), which could be expected based on our previous findings. Finally, the loss-gain ratios obtained from the exponential and the multiplicative abs-gl models are well below1 for all brands (never larger than 0.5 and mostly not exceeding0.25), which speaks again in favor of estimating gain and loss effects with two separate terms when using the parametric models. (21) 𝜆 (x)= f � 2L ( x ) f� 2G (x) ⇒ 𝜆=1 | X |∑ x∈X 𝜆(x) , 8 We use the abs-gl model version in the following, because the rel-gl version is not reasonable for the multiplicative model, compare Sect.3.4.
627 1 3 A semiparametric approach toestimating reference price effects… Neumann and Böckenholt (2014) reported an average loss-gain ratio of 1.49 (indicating moderate loss aversion) based on a meta analysis of 33 studies conducted in a parametric brand choice modeling context (i.e., using random utility models of brand choice), which contradicts the findings of our study at first glance. However, the authors also showed that loss aversion can vary substantially depending on, e.g., product characteristics and the used model specification. For example, loss-aversion turned out significantly stronger for durable product categories which commonly bare a higher financial risk than nondurables (like, e.g., orange juice). Further, lossgain ratios were found to be significantly lower for so-called sticker shock models that (like in our approach) use a price main effect in addition to gain and loss terms (as opposed to the use of gain and loss terms only). The meta study of Neumann and Böckenholt (2014) was motivated by the fact that previous findings on loss aversion regarding price effects were very inconsistent, i.e., some studies indicated strong evidence for loss aversion while others not at all (for details, see the literature cited therein).9 Mazumdar etal. (2005) emphasized much earlier that the empirical evidence on asymmetric reference price effects is mixed, referring to a number of empirical studies on brand choice not supporting loss aversion. Interestingly, Natter and Hruschka (1997) also could not find loss aversion of consumers in an empirical study for laundry detergent brands based on their estimated market share models (i.e., based on aggregate data), and reported larger coefficients for gains than for losses (in terms of absolute magnitudes). They provided a number of possible explanations for their findings that could favor gain-seeking behavior over loss aversion: price cuts are frequently supported by POS advertising like displays and have therefore a stronger effect; costs of brand switching motivate consumers to utilize price cuts on their more preferred brands; and/or there exists a high share of brand switchers who are attracted by lower prices. Still, it is important Table 7 Loss-gain ratios for abs-diff models and abs-gl models Note that the loss-gain ratio implicitly equals1 for the parametric abs-diff models, since in this case only one parameter is estimated for the price-difference term. Roman values indicate gain-seeking behavior, values in italics indicate loss aversion Brand abs-diff model abs-gl model Semiparametric Exp. Mult. Semiparametric Full rg. ≤0.50 ⋯ >1.00 Full rg. Full rg. Full rg. ≤0.50 ⋯ >1.00 Flor. Natrl. 0.53 0.35 1.20 0.46 0.16 0.18 0.20 0.17 1.08 0.31 Trop. Pure 0.07 0.13 0.03 0.08 0.12 0.25 0.08 0.08 0.04 0.10 Citrus Hill 0.15 0.11 0.26 0.74 0.25 0.31 0.08 0.07 0.30 0.34 Flor. Gold 0.55 0.12 1.37 2.80 0.42 0.46 0.15 0.07 1.28 2.26 Min. Maid 0.22 0.09 0.28 0.49 0.09 0.00 0.09 0.04 0.39 0.32 Tree Fresh 0.77 0.54 2.59 0.18 0.23 0.25 0.63 0.60 2.14 0.17 Tropicana 0.30 0.10 0.57 1.03 0.06 0.10 0.13 0.06 0.76 0.80 Dominick’s 0.28 0.12 0.32 1.29 0.01 0.03 0.14 0.10 0.32 1.09 9 We thank an anonymous reviewer for pointing us to the study of Neumann and Böckenholt (2014).
628 P.Aschersleben, W.J.Steiner 1 3 to mention that our results of course depend on the specification of the dynamic parts in our models and our decision to separate reference price effects from other dynamic effects by including a one-period lagged sales term. As discussed before in the introduction and in Sect.3.2, it is generally difficult to disentangle different dynamic (price) effects with aggregate sales data, and the small or negligible loss effects for price increases observed in our data might have been underestimated due to an unidentified stockpiling effect if one existed. In other words, since post-promo- tion dips caused by stockpiling are hard to detect in aggregate sales response models, for example as a result of very different (re)purchasing patterns of individual households, the negative effect of losses might be undervalued and actually larger. But even then, this should not be a problem for sales prediction, which is the main objective of our proposed model. 5 Conclusions In this article, we proposed a semiparametric approach to flexibly estimating reference price effects in brand sales models. In particular, we focused on the so-called price-change response of consumers (prominently introduced by Simon 1982), using aggregate store-level sales data and the price of the last period as proxy for the reference price of (an aggregate of) consumers. We compared different options to capture this dynamic price-change effect, following adaptation-level and prospect theory. While adaptation-level theory states that consumers evaluate a new price information for a brand relative to an adaptation level (which is the brand’s price of the last period in our context), prospect theory goes one step further and claims that consumers should value losses of a certain amount stronger than gains of the same amount (corresponding to price increases and price decreases of the same amount in our context), and that the value function is convex for losses and concave for gains. Accommodating functional flexibility for price effects via nonparametric regression helps to simultaneously analyze a potential asymmetry and/or disproportionality of the price-change response without the need to assume a certain functional form for it apriori. In other words, by letting the data determine the shape of the price-change effect we can easily verify if the implications of these behavioral theories hold for the data at hand. We further compared the semiparametric approach to parametric benchmark models in order to assess the added value of using nonparametric regression for estimating price (change) effects. To compare the predictive performance of our models, we conducted an empirical study using store-level scanner data of the Dominick’s Finer Foods (DFF) data base for refrigerated orange juice. For model specification, we assumed a brand’s sales to depend on the brand’s own price, the brand’s sales of the previous period, prices of substitute brands, promotional activities, store-specific and holiday effects, as well as on the brand’s previous price to capture the price-change or reference price effect in the following ways: via a single price-difference term versus two separate price terms for perceived gains and losses, where the price change with respect to the previous price is measured in absolute monetary units or as a percentage change, respectively. We further estimated two nested variants of our semiparametric model
629 1 3 A semiparametric approach toestimating reference price effects… in order to evaluate the impact of accounting for (price) dynamics on the predictive model performance: a static variant without the reference price term and without the lagged sales effect, and a simpler dynamic variant without the reference price term but including one-period lagged sales as autoregressive part. To assess the added value of employing nonparametric regression for estimating the price-change effect flexibly (as well as own- and cross-price effects), we also compared our semiparametric approach to the exponential (log-linear) and the multiplicative (log-log) sales response function (as well as to the simple linear model) as parametric benchmark models. 5.1 Results andmanagerial implications The main results of our empirical study can be summarized as follows: first, accounting for price-change or reference price effects can substantially improve the predictive performance of brand sales models (as measured by the cross-validated average root median or mean squared errors, ARMedSE or ARMSE ). For the parametric models (linear, exponential, and multiplicative), accommodating gain and loss effects with two separate price-terms (abs-gl, rel-gl) largely improves ARMedSE values (i.e., reduces prediction errors) in holdout samples, whereas using a single price-difference term (abs-diff, rel-diff) provides only small (marginal) improvements or even decreases the predictive accuracy measured by ARMedSE compared to the static model. A look at the estimated effects and the corresponding partial residuals of the competing dynamic model specifications reveals the reason for the clearly worse performance of the abs-diff and rel-diff models: obviously, the pricechange effect is asymmetric for nearly all orange juice brands analyzed, however using a single price-difference term for both gains and losses does not allow the detection of asymmetrical effects of price changes. The use of separate variables for perceived gains versus perceived losses helps to overcome this limitation. In contrast, the semiparametric models do not have this shortcoming: due to their much greater flexibility they are able to capture such asymmetries and therefore possibly different shapes for gain and loss effects, even if gain and loss effects are not separated from each other with two different price terms. This explains why the four different dynamic specifications for the reference price effect perform similarly well here. In other words, once price (change) effects are accommodated flexibly the form of the specification of the reference price term gets secondary. The predictive validity results based on the ARMSE measure closely resemble those for the ARMedSE measure with one noticeable exception. For the parametric models, modeling gain and loss effects via two separate reference price terms did no consistently improve the predictive performance for all brands compared to using a single price-difference term only. In these cases, however, taking price dynamics into account did improve the predictive accuracy either only marginally or not at all, independent of the form of including the reference price term. Second, as the spline functions are furthermore able to account for disproportionate effects of any shape, each of the semiparametric model variants provided more accurate sales predictions than its linear, exponential, or multiplicative
630 P.Aschersleben, W.J.Steiner 1 3 counterparts for all brands considered but one (for each of the two predictive validity measures). For the one brand, the semiparametric model predicted similarly well or marginally better nevertheless. Interestingly, even the static semiparametric model leads to lower ARMedSE values than each of the two dynamic exponential or multiplicative models when capturing the price-change effect with only a single price-difference term. This also held for all but one brand (two brands) for the exponential (multiplicative) model when the predictive accuracy was evaluated by the ARMSE measure. This underlines the power of nonparametric estimation techniques in the present context. Overall, improvements in predictive accuracy from accommodating price effects flexibly over the best dynamic exponential or multiplicative models ranged up to −11 % ( −12 %) in terms of ARMedSE and up to −15 % ( −14 %) in terms of ARMSE at the individual brand level. Third, referring to the shapes of the flexibly estimated price-change effects, we observed rather steep, disproportionate gain effects while rather flat loss effects for nearly all brands. Accordingly, consumers seem to weigh gains much stronger than losses in the refrigerated orange juice category, which contradicts prospect theory. Loss-gain ratios smaller than1 as a rule underlined this finding. For the gain effect, decreasing returns to scale (as suggested by prospect theory) were found only for two of the national brands. For most of the other brands, the estimated gain effect curves turn out neither strictly concave nor strictly convex, showing more or less complex nonlinearities, which in addition differed across brands. This is exactly the strength of nonparametric modeling: there is no need to search for the right functional specification(s) in advance, the shape of effects is estimated directly from the data. Also note that we controlled in our models for other dynamic effects like stockpiling or customer holdover by including oneperiod lagged sales. Still, it could be that larger stockpiling effects may have gone undetected because post-promotion dips are generally hard to detect in aggregate response models. In this case, the negative effect of losses might have been undervalued. On the other hand, refrigerated orange is a less stockable product which suggests that the estimation bias in the estimated loss effects should not be that large, if one existed. From a managerial point of view, using our more complex semiparametric approach seems worth the effort as it provides several advantages. First, as already discussed above, predictions turned out never worse, often better and sometimes considerably superior to those from any of the parametric models compared to. Second, semiparametric modeling especially pays off if price effects involve complex nonlinearities which are difficult or not at all to capture via parametric models. Even if complex nonlinearities (e.g., strong kinks or several thresholds) are not at work and improvements from using the more complex model would be not that big or only small, one need not care about the problem which parametric model to use for which brand to arrive at the best possible brand sales predictions. Our study has shown that nonlinearities for gain effects may be complex and may further turn out very differently at the individual brand level which favors the use of a flexible estimation techniques. Third, free software (BayesX) is available to easily estimate the semiparametric model (as well as the parametric models).
637 1 3 A semiparametric approach toestimating reference price effects… Table 8 Out-of-sample predictive performance of the competing models evaluated by the Average Root Mean Squared Sales Prediction Error (ARMSE) in holdout samples for each brand Quality tiers: Premium National Private Flor. Natrl. Trop. Pure Citrus Hill Flor. Gold Min. Maid Tree Fresh Tropicana Dominick’s Exponential static 31.26 61.57 129.47 55.15 𝟓𝟐.𝟖𝟎 74.68 107.86 365.53 – – – – – – – – dyn-ar 31.63 62.14 128.33 55.37 53.05 75.33 107.38 354.40 (+1.2%) (+0.9%) (−0.9%) (+0.4%) (+0.5%) (+0.9%) (−0.4%) (−3.0%) abs-diff 28.06 59.29 119.17 54.11 53.61 73.55 105.90 337.35 (−10%) (−3.7%) (−8.0%) (−1.9%) (+1.5%) (−1.5%) (−1.8%) (−7.7%) rel-diff 30.96 62.10 123.31 54.68 53.05 74.41 𝟏𝟎𝟓.𝟖𝟕 354.75 (−1.0%) (+0.9%) (−4.8%) (−0.9%) (+0.5%) (−0.4%) (−1.8%) (−2.9%) abs-gl 𝟐𝟓.𝟑𝟖 53.17 107.35 𝟓𝟑.𝟑𝟗 55.14 71.31 110.79 367.24 (−19%) (−14%) (−17%) (−3.2%) (+4.4%) (−4.5%) (+2.7%) (+0.5%) rel-gl 𝟐𝟓.𝟑𝟖 𝟓𝟎.𝟑𝟎 𝟗𝟓.𝟑𝟖 53.67 54.84 𝟕𝟎.𝟓𝟕 108.45 𝟑𝟐𝟗.𝟐𝟕 (−19%) (−18%) (−26%) (−2.7%) (+3.9%) (−5.5%) (+0.5%) (−9.9%) Multiplicative static 28.18 57.79 115.52 56.11 𝟓𝟏.𝟗𝟒 71.44 108.40 357.67 – – – – – – – – dyn-ar 28.68 58.31 113.81 56.28 52.11 72.35 106.86 346.89 (+1.8%) (+0.9%) (−1.5%) (+0.3%) (+0.3%) (+1.3%) (−1.4%) (−3.0%) abs-diff 25.93 55.50 100.36 55.04 52.86 70.19 𝟏𝟎𝟑.𝟗𝟕 𝟑𝟑𝟐.𝟓𝟓 (−8.0%) (−4.0%) (−13%) (−1.9%) (+1.8%) (−1.7%) (−4.1%) (−7.0%) abs-gl 𝟐𝟓.𝟏𝟕 𝟓𝟎.𝟏𝟕 𝟗𝟕.𝟑𝟑 𝟓𝟒.𝟓𝟗 56.76 𝟔𝟕.𝟓𝟑 115.37 359.10 (−11%) (−13%) (−16%) (−2.7%) (+9.3%) (−5.5%) (+6.4%) (+0.4%)
638 P.Aschersleben, W.J.Steiner 1 3 Best models per brand and type of model (i.e., exponential, multiplicative, and semiparametric) are marked in bold Table 8 (continued) Quality tiers: Premium National Private Flor. Natrl. Trop. Pure Citrus Hill Flor. Gold Min. Maid Tree Fresh Tropicana Dominick’s Semiparametric static 25.51 53.85 103.92 53.34 𝟒𝟗.𝟖𝟕 61.31 103.30 353.49 – – – – – – – – dyn-ar 25.61 54.37 102.37 53.51 49.94 𝟔𝟎.𝟖𝟖 102.07 340.05 (+0.4%) (+1.0%) (−1.5%) (+0.3%) (+0.1%) (−0.7%) (−1.2%) (−3.8%) abs-diff 26.05 49.05 95.32 51.63 51.38 61.17 91.75 𝟑𝟎𝟑.𝟕𝟓 (+2.1%) (−8.9%) (−8.3%) (−3.2%) (+3.0%) (−0.2%) (−11%) (−14%) rel-diff 𝟐𝟓.𝟑𝟏 48.90 𝟗𝟏.𝟔𝟖 𝟓𝟏.𝟒𝟔 50.92 61.04 90.65 311.86 (−0.8%) (−9.2%) (−12%) (−3.5%) (+2.1%) (−0.4%) (−12%) (−12%) abs-gl 26.04 49.05 96.18 51.62 50.42 61.28 90.27 306.21 (+2.1%) (−8.9%) (−7.4%) (−3.2%) (+1.1%) (±0.0%) (−13%) (−13%) rel-gl 25.67 𝟒𝟖.𝟔𝟑 94.19 51.47 50.03 61.11 𝟖𝟗.𝟓𝟐 304.90 (+0.6%) (−9.7%) (−9.4%) (−3.5%) (+0.3%) (−0.3%) (−13%) (−14%) Best semipar. vs. best expon. −0.3 % −3.3 % −3.9 % −3.6 % −5.5 % −14 % −15 % −7.8 % Best semipar. vs. best multipl. +0.6 % −3.1 % −5.8 % −5.7 % −4.0 % −9.8 % −14 % −8.7 %
639 1 3 A semiparametric approach toestimating reference price effects… Table 9 Out-of-sample predictive performance of the linear models evaluated by the Average Root Median Squared Sales Prediction Error (ARMedSE) in holdout samples for each brand Best models per brand are marked in bold Quality tiers: Premium National Private Flor. Natrl. Trop. Pure Citrus Hill Flor. Gold Min. Maid Tree Fresh Tropicana Dominick’s Linear static 14.82 30.17 48.46 18.37 24.06 23.92 55.11 158.13 – – – – – – – – dyn-ar 14.80 30.34 48.29 18.51 23.64 23.87 58.05 159.35 (−0.1%) (+0.6%) (−0.4%) (+0.8%) (−1.7%) (−0.2%) (+5.3%) (+0.8%) abs-diff 13.03 28.85 43.56 18.38 23.76 23.06 56.88 155.68 (−12%) (−4.4%) (−10%) (+0.1%) (−1.2%) (−3.6%) (+3.2%) (−1.5%) rel-diff 13.63 30.38 45.89 18.76 23.64 23.66 57.58 158.50 (−8.0%) (+0.7%) (−5.3%) (+2.1%) (−1.7%) (−1.1%) (+4.5%) (+0.2%) abs-gl 𝟖.𝟗𝟏 20.68 29.46 15.79 22.78 20.24 48.75 132.18 (−40%) (−31%) (−39%) (−14%) (−5.3%) (−15%) (−12%) (−16%) rel-gl 8.94 𝟐𝟎.𝟓𝟗 𝟐𝟕.𝟒𝟗 𝟏𝟓.𝟓𝟕 𝟐𝟏.𝟖𝟖 𝟏𝟗.𝟗𝟑 𝟒𝟔.𝟔𝟓 𝟏𝟐𝟗.𝟑𝟎 (−40%) (−32%) (−43%) (−15%) (−9.1%) (−17%) (−15%) (−18%) Best semipar. vs. best linear −44 % −47 % −72 % −44 % −54 % −67 % −58 % −39 %
640 P.Aschersleben, W.J.Steiner 1 3 Table 10 Out-of-sample predictive performance of the linear models evaluated by the Average Root Mean Squared Sales Prediction Error (ARMSE) in holdout samples for each brand Best models per brand are marked in bold Quality tiers: Premium National Private Flor. Natrl. Trop. Pure CitrusHill Flor. Gold Min. Maid Tree Fresh Tropicana Dominick’s Linear static 35.74 69.92 130.75 56.55 58.71 77.57 115.98 423.31 – – – – – – – – dyn-ar 35.75 69.89 130.28 56.52 58.65 77.60 114.77 421.38 (±0.0%) (±0.0%) (−0.4%) (−0.1%) (−0.1%) (±0.0%) (−1.0%) (−0.5%) abs-diff 33.33 67.85 124.66 56.00 58.68 76.89 114.08 413.23 (−6.7%) (−3.0%) (−4.7%) (−1.0%) (−0.1%) (−0.9%) (−1.6%) (−2.4%) rel-diff 34.81 69.60 127.05 56.24 58.67 77.28 114.39 420.31 (−2.6%) (−0.5%) (−2.8%) (−0.5%) (−0.1%) (−0.4%) (−1.4%) (−0.7%) abs-gl 𝟐𝟗.𝟕𝟎 59.33 113.91 𝟓𝟒.𝟖𝟐 57.68 75.05 109.44 380.90 (−17%) (−15%) (−13%) (−3.1%) (−1.8%) (−3.2%) (−5.6%) (−10%) rel-gl 29.87 𝟓𝟖.𝟖𝟏 𝟏𝟏𝟏.𝟕𝟑 54.85 𝟓𝟕.𝟐𝟓 𝟕𝟒.𝟖𝟑 𝟏𝟎𝟖.𝟏𝟐 𝟑𝟕𝟔.𝟑𝟓 (−16%) (−16%) (−15%) (−3.0%) (−2.5%) (−3.5%) (−6.8%) (−11%) Best semipar. vs. best linear −15 % −17 % −18 % −6.1 % −13 % −19 % −17 % −19 %
641 1 3 A semiparametric approach toestimating reference price effects… Acknowledgements The data for our empirical study were provided by the James M. Kilts Center, GSB, University of Chicago. All models were estimated with the public domain software package BayesX (Brezger etal. 2005). Further calculations and plots were conducted with the free and open source software R (R Core Team 2021) and the tidyverse collection of R packages (Wickham etal. 2019), amongst others. We thank two anonymous reviewers for their valuable comments and suggestions on our way toward a publishable manuscript. Funding Open Access funding enabled and organized by Projekt DEAL. Declarations Conflict of interests The authors have no relevant financial or non-financial interests to disclose. Open Access This article is licensed under a Creative Commons Attribution 4.0 International License, which permits use, sharing, adaptation, distribution and reproduction in any medium or format, as long as you give appropriate credit to the original author(s) and the source, provide a link to the Creative Commons licence, and indicate if changes were made. The images or other third party material in this article are included in the article’s Creative Commons licence, unless indicated otherwise in a credit line to the material. If material is not included in the article’s Creative Commons licence and your intended use is not permitted by statutory regulation or exceeds the permitted use, you will need to obtain permission directly from the copyright holder. To view a copy of this licence, visit http:// creat iveco mmons. org/ licen ses/ by/4. 0/. References Andrews RL, Currim IS, Leeflang PSH, Lim J (2008) Estimating the SCAN*PRO model of store sales: HB, FM or just OLS? Int J Res Mark 25(1):22–33. https:// doi. org/ 10. 1016/j. ijres mar. 2007. 10. 001 Baumgartner B, Guhl D, Kneib T, Steiner WJ (2018) Flexible estimation of time-varying effects for frequently purchased retail goods: a modeling approach based on household panel data. OR Spectr 40(4):837–873. https:// doi. org/ 10. 1007/ s00291- 018- 0530-6 Boztuğ Y, Hildebrandt L, Raman K (2014) Detecting price thresholds in choice models using a semiparametric approach. OR Spectr 36(1):187–207. https:// doi. org/ 10. 1007/ s00291- 012- 0313-4 Brezger A, Steiner WJ (2008) Monotonic regression based on Bayesian P-splines: an application to estimating price response functions from store-level scanner data. J Bus Econ Stat 26(1):90–104. https:// doi. org/ 10. 1198/ 07350 01070 00000 223 Brezger A, Kneib T, Lang S (2005) BayesX: analyzing bayesian structured additive regression models. J Stat Softw 14(11). https:// doi. org/ 10. 18637/ jss. v014. i11, http:// bayesx. org Briesch RA, Krishnamurthi L, Mazumdar T, Raj SP (1997) A comparative analysis of reference price models. J Consum Res 24(2):202–214. https:// doi. org/ 10. 1086/ 209505 Chan T, Narasimhan C, Zhang Q (2008) Decomposing promotional effects with a dynamic structural model of flexible consumption. J Mark Res 45(4):487–498. https:// doi. org/ 10. 1509/ jmkr. 45.4. 487 Diller H (2008) Preispolitik, 4th edn. Kohlhammer, Stuttgart Eilers PHC, Marx BD (1996) Flexible smoothing with B -splines and penalties. Stat Sci 11(2):89–102. https:// doi. org/ 10. 1214/ ss/ 10384 25655 Erdem T, Katz ML, Sun B (2010) A simple test for distinguishing between internal reference price theories. Quant Mark Econ 8(3):303–332. https:// doi. org/ 10. 1007/ s11129- 010- 9087-7 Fahrmeir L, Kneib T, Lang S, Marx BD (2013) Regression: models, methods and applications. Springer, Heidelberg Fibich G, Gavious A, Lowengart O (2003) Explicit solutions of optimization models and differential games with nonsmooth (asymmetric) reference-price effects. Oper Res 51(5):721–734. https:// doi. org/ 10. 1287/ opre. 51.5. 721. 16758 Fibich G, Gavious A, Lowengart O (2007) Optimal price promotion in the presence of asymmetric reference-price effects. Manag Decis Econ 28(6):569–577. https:// doi. org/ 10. 1002/ mde. 1333
642 P.Aschersleben, W.J.Steiner 1 3 Foekens EW, Leeflang PSH, Wittink DR (1999) Varying parameter models to accommodate dynamic promotion effects. J Econ 89(1–2):249–268. https:// doi. org/ 10. 1016/ S0304- 4076(98) 00063-3 Franses PH, Ghijsels H (1999) Additive outliers, GARCH and forecasting volatility. Int J Forecast 15(1):1–9. https:// doi. org/ 10. 1016/ S0169- 2070(98) 00053-3 Gedenk K (2002) Verkaufsförderung. Vahlen, München Greene WH (2008) Econometric analysis, 6th edn. Pearson Prentice Hall, Upper Saddle River, NJ Greenleaf EA (1995) The impact of reference price effects on the profitability of price promotions. Mark Sci 14(1):82–104. https:// doi. org/ 10. 1287/ mksc. 14.1. 82 Härdle W (1990) Applied nonparametric regression, vol 19. Cambridge Univ. Press, Cambridge Haupt H, Kagerer K (2012) Beyond mean estimates of price and promotional effects in scanner-panel sales-response regression. J Retail Consum Serv 19(5):470–483. https:// doi. org/ 10. 1016/j. jretc onser. 2012. 06. 002 Haupt H, Kagerer K, Steiner WJ (2014) Smooth quantile-based modeling of brand sales, price and promotional effects from retail scanner panels. J Appl Econ 29(6):1007–1028. https:// doi. org/ 10. 1002/ jae. 2347 Helson H (1964) Adaptation-level theory: an experimental and systematic approach to behavior. Harper and Row, New York Horváth C, Fok D (2013) Moderating factors of immediate, gross, and net cross-brand effects of price promotions. Mark Sci 32(1):127–152. https:// doi. org/ 10. 1287/ mksc. 1120. 0748 Hruschka H (1997) Schätzung und normative Analyse ausgewählter Preis-Absatz-Funktionen. Z Betriebs wirtsch 67(8):845–864 Hruschka H (2000) Specification, estimation, and empirical corroboration of Gutenberg’s kinked demand Curve. In: Albach H, Brockhoff KKL, Eymann E, Jungen P, Steven M, Luhmer A (eds) Theory of the firm. Springer Berlin Heidelberg, pp 153–168. https:// doi. org/ 10. 1007/ 978-3- 642- 59661-2_8 Hruschka H (2006a) Relevance of functional flexibility for heterogeneous sales response models: a comparison of parametric and semi-nonparametric models. Eur J Oper Res 174(2):1009–1020. https:// doi. org/ 10. 1016/j. ejor. 2005. 05. 003 Hruschka H (2006b) Statistical and managerial relevance of aggregation level and heterogeneity in sales response models. Mark ZFP J Res Manag 28(JRM 2):94–102. https:// doi. org/ 10. 15358/ 0344- 1369- 2006- JRM-2- 94 Hruschka H (2007) Clusterwise pricing in stores of a retail chain. OR Spectr 29(4):579–595. https:// doi. org/ 10. 1007/ s00291- 006- 0075-y Hruschka H (2017) Functional flexibility, latent heterogeneity and endogeneity in aggregate market response models. Mark ZFP J Res Manag 39(3):17–31. https:// doi. org/ 10. 15358/ 0344- 1369- 2017-3- 17 Kahnemann D, Tversky A (1979) Prospect theory: an analysis of decision under risk. Econometrica 47(2):263–291. https:// doi. org/ 10. 2307/ 19141 85 Kalyanam K, Shively TS (1998) Estimating irregular pricing effects: a stochastic spline regression approach. J Mark Res 35(1):16–29. https:// doi. org/ 10. 1177/ 00222 43798 03500 104 Kopalle PK, Rao AG, Assunção JL (1996) Asymmetric reference price effects and dynamic pricing policies. Mark Sci 15(1):60–85. https:// doi. org/ 10. 1287/ mksc. 15.1. 60 Kopalle PK, Mela CF, Marsh L (1999) The dynamic effect of discounting on sales: empirical analysis and normative pricing implications. Mark Sci 18(3):317–332. https:// doi. org/ 10. 1287/ mksc. 18.3. 317 Krishnamurthi L, Mazumdar T, Raj SP (1992) Asymmetric response to price in consumer brand choice and purchase quantity decisions. J Consum Res 19(3):387. https:// doi. org/ 10. 1086/ 209309 Kucher E (1987) Absatzdynamik nach Preisänderung. Mark ZFP J Res Manag 9(3):177–182 Lang S, Brezger A (2004) Bayesian P-splines. J Comput Graph Stat 13(1):183–212. https:// doi. org/ 10. 1198/ 10618 60043 010 Lang S, Steiner WJ, Weber A, Wechselberger P (2015) Accommodating heterogeneity and nonlinearity in price effects for predicting brand sales and profits. Eur J Oper Res 246(1):232–241. https:// doi. org/ 10. 1016/j. ejor. 2015. 02. 047 Leeflang PSH, Wittink DR, Wedel M, Naert PA (2000) Building models for marketing decisions, vol 9. Springer US, Boston, MA Macé S, Neslin SA (2004) The determinants of pre- and postpromotion dips in sales of frequently purchased goods. J Mark Res 41(3):339–350. https:// doi. org/ 10. 1509/ jmkr. 41.3. 339. 35992 Mazumdar T, Raj SP, Sinha I (2005) Reference price research: review and propositions. J Mark 69(4):84– 102. https:// doi. org/ 10. 1509/ jmkg. 2005. 69.4. 84 Montgomery AL (1997) Creating micro-marketing pricing strategies using supermarket scanner data. Mark Sci 16(4):315–337. https:// doi. org/ 10. 1287/ mksc. 16.4. 315
643 1 3 A semiparametric approach toestimating reference price effects… Natter M, Hruschka H (1997) Ankerpreise als Erwartungen oder dynamische latente Variablen in Marktreaktionsmodellen. Schmalenbachs Z betriebswirtsch Forsch (zfbf) 49(9):747–764 Neslin SA, Schneider Stone LG (1996) Consumer inventory sensitivity and the postpromotion dip. Mark Lett 7(1):77–94. https:// doi. org/ 10. 1007/ BF005 57313 Neslin SA, Shoemaker RW (1989) An alternative explanation for lower repeat rates after promotion purchases. J Mark Res 26(2):205–213. https:// doi. org/ 10. 1177/ 00222 43789 02600 206 Neslin SA, vanHeerde HJ (2009) Promotion dynamics. Found Trends® Mark 3(4):177–268. https:// doi. org/ 10. 1561/ 17000 00010 Neumann N, Böckenholt U (2014) A meta-analysis of loss aversion in product choice. J Retail 90(2):182– 197. https:// doi. org/ 10. 1016/j. jretai. 2014. 02. 002 Nijs VR, Dekimpe MG, Steenkamps JBEM, Hanssens DM (2001) The category-demand effects of price promotions. Mark Sci 20(1):1–22. https:// doi. org/ 10. 1287/ mksc. 20.1. 1. 10197 Pauwels KH, Srinivasan S, Franses PH (2007) When do price thresholds matter in retail categories? Mark Sci 26(1):83–100. https:// doi. org/ 10. 1287/ mksc. 1060. 0207 R Core Team (2021) R: A Language and Environment for Statistical Computing. R Foundationfor Statistical Computing, Vienna, Austria. https:// www.R- proje ct. org/ Rinne HJ (1981) An empirical investigation of the effects of reference prices on sales. PhD dissertation, Purdue University, West Lafayette Rossi PE (2014) Bayesian non- and semi-parametric methods and applications. Princeton Univ. Press, Princeton, NJ Simon H (1982) Preismanagement. Gabler Verlag, Wiesbaden Simon H (1992) Preismanagement: Analyse–Strategie–Umsetzung, 2nd edn. Gabler, Wiesbaden Slonim R, Garbarino E (2009) Similarities and differences between stockpiling and reference effects. Manag Decis Econ 30(6):351–371. https:// doi. org/ 10. 1002/ mde. 1453 Steiner WJ, Brezger A, Belitz C (2007) Flexible estimation of price response functions using retail scanner data. J Retail Consum Serv 14(6):383–393. https:// doi. org/ 10. 1016/j. jretc onser. 2007. 02. 008 van Heerde HJ (1999) Models for sales promotion effects based on store-level scanner data. Labyrint Publication, Capelle a/d IJssel van Heerde HJ (2017) Non- and semiparametric regression models. In: Leeflang PSH, Wieringa JE, Bijmolt THA, Pauwels KH (eds) Advanced methods for modeling markets, international series in quantitative marketing. Springer International Publishing, pp 555–579. https:// doi. org/ 10. 1007/ 978- 3- 319- 53469-5_ 17 vanHeerde HJ, Neslin SA (2017) Sales promotion models. In: Wierenga B, van der Lans R (eds) Handbook of marketing decision models. 254, Springer International Publishing, pp 13–77. https:// doi. org/ 10. 1007/ 978-3- 319- 56941-3_2 van Heerde HJ, Leeflang PSH, Wittink DR (2000) The estimation of pre- and postpromotion dips with store-level scanner data. J Mark Res 37(3):383–395. https:// doi. org/ 10. 1509/ jmkr. 37.3. 383. 18782 van Heerde HJ, Leeflang PSH, Wittink DR (2001) Semiparametric analysis to estimate the deal effect curve. J Mark Res 38(2):197–215. https:// doi. org/ 10. 1509/ jmkr. 38.2. 197. 18842 van Heerde HJ, Leeflang PSH, Wittink DR (2002) How promotions work: SCAN*PRO-based evolutionary model building. Schmalenbach Bus Rev 54(3):198–220. https:// doi. org/ 10. 1007/ BF033 96653 van Heerde HJ, Leeflang PSH, Wittink DR (2004) Decomposing the sales promotion bump with store data. Mark Sci 23(3):317–334. https:// doi. org/ 10. 1287/ mksc. 1040. 0061 Weber A, Steiner WJ (2012) Zur Berücksichtigung von Heterogenität versus funktionaler Flexibilität in Absatzreaktionsmodellen: Eine empirische Studie auf Basis von Handelsdaten. Z Betriebswirtsch 82(12):1337–1365. https:// doi. org/ 10. 1007/ s11573- 012- 0638-0 Weber A, Steiner WJ, Lang S (2017) A comparison of semiparametric and heterogeneous store sales models for optimal category pricing. OR Spectr 39(2):403–445. https:// doi. org/ 10. 1007/ s00291- 016- 0459-6 Wickham H, Averick M, Bryan J, Chang W, McGowan L, Fran, cois R, Grolemund G, Hayes A, Henry L, Hester J, Kuhn M, Pedersen T, Miller E, Bache S, M¨uller K, Ooms J, Robinson D, Seidel D, Spinu V, Takahashi K, Vaughan D, Wilke C, Woo K, Yutani H, (2019) Welcometo the Tidyverse. J Open Sourc Softw 4(43):1686. https:// doi. org/ 10. 21105/ joss. 01686 Publisher’s Note Springer Nature remains neutral with regard to jurisdictional claims in published maps and institutional affiliations.