Abstract
Measure-valued Pólya urn processes (MVPP) are Markov chains with an additive structure that serve as an extension of the generalized k-color Pólya urn model towards a continuum of possible colors. We prove that, for any MVPP on a Polish space , the normalized sequence agrees with the marginal predictive distributions of some random process . Moreover, , , where is a random transition kernel on ; thus, if represents the contents of an urn, then denotes the color of the ball drawn with distribution and —the subsequent reinforcement. In the case , for some non-negative random weights , the process is better understood as a randomly reinforced extension of Blackwell and MacQueen’s Pólya sequence. We study the asymptotic properties of the predictive distributions and the empirical frequencies of under different assumptions on the weights. We also investigate a generalization of the above models via a randomization of the law of the reinforcement.
Keywords:
predictive distributions; random probability measures; reinforced processes; Pólya sequences; urn schemes; Bayesian inference; conditional identity in distribution; total variation distance MSC:
60G57; 60B10; 60G25; 60F05; 60G09
1. Introduction
Let be a sequence of homogeneous random observations, taking values in a Polish space . The central assumption in the Bayesian approach to inductive reasoning is that is exchangeable, that is, its law is invariant under finite permutations. Then, by de Finetti’s theorem, there exists a random probability measure on such that, given , the random variables are conditionally independent and identically distributed with marginal distribution (see [1], Section 3), denoted
Furthermore, is the almost sure (a.s.) weak limit of the predictive distributions and the empirical frequencies,
The model (1) is completed by choosing a prior distribution for . Inference consists in computing the conditional (posterior) distribution of given an observed sample , with most inferential conclusions depending on some average with respect to the posterior distribution; for example, under squared loss, for any measurable set , the best estimate of is the posterior mean, . In addition, the posterior mean can be utilized for predictive inference since
A different modeling strategy uses the Ionescu–Tulcea theorem to define the law of the process from the sequence of predictive distributions, . In that case, one can refer to Theorem 3.1 in [2] for necessary and sufficient conditions on to be consistent with exchangeability. The predictive approach to model building is deeply rooted in Bayesian statistics, where the parameter is assigned an auxiliary role and the focus is on observable “facts”, see [2,3,4,5,6]. Moreover, using the predictive distributions as primary objects allows one to make predictions instantly or helps ease computations. See [7] for a review on some well-known predictive constructions of priors for Bayesian inference.
In this work, we consider a class of predictive constructions based on measure-valued Pólya urn processes (MVPP). MVPPs have been introduced in the probabilistic literature [8,9] as an extension of k-color urn models, but their implications for (Bayesian) statistics have yet to be explored. A first aim of the paper is thus to show the potential use of MVPPs as predictive constructions in Bayesian inference. In fact, some popular models in Bayesian nonparametric inference can be framed in such a way, see Equation (8). A second aim of the paper is to suggest novel extensions of MVPPs that we believe can offer more flexibility in statistical applications.
MVPPs are essentially measure-valued Markov processes that have an additive structure, with the formal definition being postponed to Section 2.1 (Definition 1). Given an MVPP , we consider a sequence of random observations that are characterized by and, for ,
The random measure is not necessarily measurable with respect to , so the predictive construction (4) is more flexible than models based solely on the predictive distributions of ; for example, allows for the presence of latent variables or other sources of observable data (see also [10] for a covariate-based predictive construction). However, (4) can lead to an imbalanced design, which may break the symmetry imposed by exchangeability. Nevertheless, it is still possible that the sequence satisfies (2) for some , in which case Lemma 8.2 in [1] implies that is asymptotically exchangeable with directing random measure .
In Theorem 1, we show that, taking as primary, the sequence in (4) can be chosen such that
where is a measurable map from to the space of finite measures on . Models of the kind (4)–(5) are computationally efficient. Indeed, as new observations become available, predictions can be updated at a constant computational cost and with limited storage of information. If, in addition, is asymptotically exchangeable, then (4)–(5) can provide a computationally simple approximation of an exchangeable scheme for Bayesian inference, along the lines in [11].
The recursive formula (5) allows us to interpret the dynamics of MVPPs in terms of an urn sampling scheme, as the name suggests. Let be a non-random finite measure on . Suppose we have an urn whose contents are described by in the sense that denotes the total mass of balls with colors in . At time , a ball is extracted at random from the urn, and we denote its color by . The urn is then reinforced according to a replacement rule , so that the updated composition becomes . At any time , a ball of color is picked with probability distribution , and the contents of the urn are subsequently reinforced by . In the case the space of colors is finite, , the above procedure is better known as a generalized k-color Pólya urn [12].
We focus our analysis on MVPPs for which is concentrated on x; thus, after each draw, we reinforce only the color of the observed ball. More formally, we consider MVPPs that have a reinforcement measure of the kind , , where is some non-negative random variable. In that case, Equations (4) and (5) become
and
A notable example is Blackwell and MacQueen’s em Pólya sequence [13], which is a random process characterized by and, for ,
for some probability measure on and a constant . By [13], is exchangeable and corresponds to the model (1) with Dirichlet process prior with parameters . It is easily seen that (8) is related to the MVPP given by and, for ,
Therefore, we will call any MVPP a randomly reinforced Pólya process (RRPP) if it admits representation (6)–(7).
Existing studies on MVPPs look at models that have mostly a balanced design, i.e., , , and assume irreducibility-like conditions for , see [8,9,14,15] and Remark 4 in [16]. In contrast, RRPPs require that , and so are excluded from the analysis in those papers. In fact, this difference in reinforcement mechanisms mirrors the dichotomy within k-color urn models, where the replacement R is best described in terms of a matrix with random elements. There, the class of randomly reinforced urns [17] assumes an R with zero off-diagonal elements (i.e., we reinforce only the color of the observed ball), whereas the generalized Pólya urn models require the mean replacement matrix to be irreducible. Similarly to the k-color case, RRPPs need the use of different techniques, which yield completely different results than those in [8,9,14,15,16]. As an example, Theorem 1 in [16] and our Theorem 2 prove convergence of the kind (2), yet the limit probability measure in [16] is non-random.
The RRPP has been implicitly studied by [17,18,19,20,21,22,23], among others, with the focus being on the process . Those papers deal primarily with the k-color case (with the exception of [18,19,23]) and can be categorized on the basis of their assumptions on . For example, [18,19,21,22] assume that and are independent, in which case the process is conditionally identically distributed (c.i.d.) [21], that is, conditionally on current information, all future observations are identically distributed. It follows from [21] that c.i.d. processes preserve many of the properties of exchangeable sequences and, in particular, satisfy (2)–(3). In contrast, [17,20,23] assume that the reinforcement depends on the particular color , and prove a version of (2) where is concentrated on the set of dominant colors for which the expected reinforcement is maximum. In this work, we reconsider the above models in the framework of RRPPs. For the c.i.d. case, we prove results whose analogues have already been established by [23] for the model with dominant colors. In particular, we extend the convergence in (2) to be in total variation and give a unified central limit theorem. We also examine the number of distinct values that are generated by the sequence .
In some applications, the definition of an MVPP can be too restrictive as it assumes that the probability law of the reinforcement R is known. However, we can envisage situations where the law is itself random, so we extend the definition of an MVPP by introducing a random parameter V. The resulting generalized measure-valued Pólya urn process (GMVPP) turns out to be a mixture of Markov processes and admits representation (4)–(5), conditional on the parameter V. When the reinforcement measure is concentrated on x, we call a generalized randomly reinforced Pólya process (GRRPP). We give a characterization of GRRPPs with exchangeable weights and show that the process is partially conditionally identically distributed (partially c.i.d) [24], that is, conditionally on the past observations and the concurrent observation from the other sequence, the future observations are marginally identically distributed. We also extend some of the results for RRPPs to the generalized setting.
The paper is structured as follows. In Section 2.1, we recall the definition of a measure-valued Pólya urn process and prove representation (4)–(5) for a suitably selected sequence . Section 2.2 defines a particular subclass of MVPPs, called randomly reinforced Pólya processes (RRPP), which share with exchangeable Pólya sequences the property of reinforcing only the observed color. Section 3 is devoted to the study of the asymptotic properties of RRPPs. In Section 4, we give the definition of GMVPPs and GRRPPs, and obtain basic results.
2. Definitions and a Representation Result
Let be a complete separable metric space, endowed with its Borel -field . Denote by
the collections of measures on that are finite, finite and non-null, and probability measures, respectively. We regard , and as measurable spaces equipped with the -fields generated by , . We further let
be the collections of transition kernels K from to that are finite and probability kernels, respectively. Any non-null measure has a normalized version . If is measurable, then denotes the induced mapping of measures, , .
All random quantities are defined on a common probability space , which is assumed to be rich enough to support any required randomization. The symbol “⊥” will be used to denote independence between random objects, and “” equality in distribution.
2.1. Measure-Valued Pólya urn Processes
Let describe the contents of an urn, as in Section 1. Once a ball is picked at random from , the urn is reinforced according to a replacement rule, which is formally a kernel that maps colors to finite measures; thus,
represents the updated urn composition if a ball of color x has been observed. In general, R is random and there exists a probability kernel such that , . Then, the distribution of (9) prior to the sampling of the urn is given by
where is the measurable map from to . By Lemma 3.3 in [9], is a measurable map from to .
Definition 1
(Measure-Valued Pólya Urn Process [9]). A sequence of random finite measures on is called a measure-valued Pólya urn process (MVPP) with parameters and if it is a Markov process with transition kernel given by (10). If, in particular, for some , then is said to be a deterministic MVPP.
The representation theorem below formalizes the idea of MVPP as an urn scheme.
Theorem 1.
A sequence of random finite measures is an MVPP with parameters if and only if, for every ,
where is a sequence of -valued random variables such that and, for ,
and R is a random finite transition kernel on such that
Proof.
Conversely, suppose is a MVPP with parameters . As is a probability kernel from to and is Polish, then there exists by Lemma 4.22 in [25] a measurable function such that, for every ,
whenever U is a uniform random variable on , denoted .
Let us prove by induction that there exists a sequence such that , , , a.s., , and, for every ,
- (i)
- ;
- (ii)
- and ;
- (iii)
- a.s.;
- (iv)
- ;
- (v)
- .
Regarding the base case, let and be independent random variables such that and . It follows that, for any measurable set ,
thus, . By Theorem 8.17 in [25], there exist random variables and such that
and . Then, in particular, and , so
Regarding the induction step, assume that – hold true until some . Let and be such that , , and
It follows from that, for any measurable set ,
thus, . By Theorem 8.17 in [25], there exist random variables and such that
and . Then, in particular, , , and
Moreover,
therefore,
By Theorem 8.12 in [25], statement with is equivalent to and , . The latter follows from the induction hypothesis since, by , we have for every . □
The process in Theorem 1 corresponds to the sequence of observed colors from the implied urn sampling scheme. Furthermore, the replacement rule takes the form , where f is some measurable function, , and , from which it follows that
and
Thus, the sequence models the additional randomness in the reinforcement measure R. Janson [9] obtains a rather similar result; Theorem 1.3 in [9] states that any MVPP can be coupled with a deterministic MVPP on in the sense that
where is the Lebesgue measure on , and is the product measure on . In our case, the MVPP defined by and, for ,
has a non-random replacement rule and satisfies (16) on a set of probability one.
2.2. Randomly Reinforced Pólya Processes
It follows from (8) that any Pólya sequence generates a deterministic MVPP through
Here, we consider a randomly reinforced extension of Pólya sequences in the form of an MVPP with replacement rule , , where is a non-negative random variable.
Definition 2
(Randomly Reinforced Pólya Process). We call an MVPP with parameters a randomly reinforced Pólya process (RRPP) if there exists such that , , where is the map .
Observe that, for RRPPs, the reinforcement measure in (14)–(15) concentrates its mass on x; thus, we obtain the following variant of the representation result in Theorem 1.
Proposition 1.
Let be an RRPP with parameters . Then, there exist a measurable function and a sequence such that, using , we have for every that
where and, for , , , and
Moreover,
It follows from (19) that , , whenever . Then, the random measure
is such that , where appears in Definition 2.
3. Asymptotic Properties of RRPP
In this section, we study the asymptotic properties of RRPPs through the sequence in the representation (17). We show that the limit behavior of depends on the relationship between weights and observations. In particular, when in (20) is constant with respect to the color x, the process is conditionally identically distributed (c.i.d.) and, for every , the normalized sequence is a bounded martingale. We consider the c.i.d. case in Section 3.3. In contrast, if some colors x have a higher expected reinforcement, then they tend to dominate the observation process and, as n grows to infinity, the probability measure concentrates its mass on the subset of dominant colors, see Theorem 2.
3.1. Preliminaries
Our focus is on the convergence of the normalized sequence , which by Theorem 1 is a.s. equal to the predictive distributions (18). We also consider the sequence of empirical frequencies of , defined for by
We obtain conditions under which the convergence in (2) extends to convergence in total variation, where the total variation distance between any two probability measures is given by
To state some of the results, we recall the definition of support of a probability measure ,
Of particular interest is the conditional probability of observing a new color, given by
for , where . This would inform us on the number of distinct values in a sample of size n,
since .
The following modes of convergence are used when we investigate the rate of convergence of the distance between and .
Almost sure (a.s.) conditional convergence. Let be a filtration and . A sequence is said to converge to in the sense of a.s. conditional convergence w.r.t. if the conditional distribution of , given , converges weakly on a set of probability one to , that is, as ,
We refer to [22] for more details.
Stable convergence. Stable convergence is a strong form of convergence in distribution, albeit weaker than a.s. conditional convergence. A sequence is said to converge stably to if
for all continuous bounded functions f and any integrable random variable V. The main application of stable convergence is in central limit theorems that allow for mixing variables in the limit. See [26] for a complete reference on stable convergence.
In the sequel, the stable and a.s. conditional limits will be some Gaussian law, which we denote by for parameters , where .
3.2. RRPP with Dominant Colors
Using (20), let us define, for ,
We further let
be the set of dominant colors. The model (18) with has been studied by [23] under the assumption that is strictly greater than the next largest value of in the support of . Then, the probability of observing a non-dominant color, , vanishes, and the predictive and the empirical distributions converge in total variation to a common random probability measure, which is concentrated on . For completeness reasons, we report here the main results from [23].
Theorem 2
([23], Theorem 3.3). For any RRPP that satisfies
there exists a random probability measure on with a.s. such that
Under conditions (21), Theorem 3.3 in [23] implies . If is further diffuse, then a.s., and so by Theorem 1 in [27]; thus, by Theorem 1 in [27], Proposition 3.4 in [23] shows that the actual growth rate is that of a Pólya sequence,
In addition to the uniform convergence in Theorem 2, the authors in [23] obtain set-wise rates of convergence. To state their result, we introduce, for any ,
which exists a.s. under the assumptions of Theorem 2.
Theorem 3
Then,
and
where , is the filtration generated by .
3.3. RRPP with Independent Weights
Let be an RRPP with reinforcement distribution that does not depend on x. Using the notation of Section 3.2, we have
and, thus, . An equivalent formulation can be given in terms of the sequence of weights in Proposition 1, whereby
for some measurable function h, with and . Then, and , which implies that .
The model (18) with weights (24) has been studied by [18,19,22], among others, where the authors obtain central limit theorems and study the growth rate of when . Their results rely on the fact that is conditionally identically distributed (c.i.d.) with respect to the filtration generated by . By [21], an -valued random sequence that is adapted to a filtration is said to be c.i.d. with respect to if and only if is identically distributed and, for every ,
Proposition 2
([19], Lemma 6). For any RRPP with , the observation process is c.i.d. with respect to the filtration generated by .
C.i.d. processes preserve many of the properties of exchangeable sequences, see [21]. For example, if is c.i.d., then there exists a random probability measure such that (2)–(3) hold true with respect to the filtration used in the definition (25). It follows for the model in Proposition 2 that there exists such that, for every ,
In fact, by (25), the sequence is a bounded martingale. On the other hand, (23) implies that ; therefore, any RRPP with whose weights are bounded, , satisfies the assumptions of Theorem 2. In that case,
It follows from Theorem 4.2 in [23] that the boundedness condition in (21) is needed to show that ; and converge set-wise to , which is non-trivial in that setting. Here, is granted as is i.i.d., and has already been established; thus, we obtain the following result for RRPPs with independent weights.
Theorem 4.
For any RRPP with , there exists a random probability measure on such that
Proof.
Let be the joint observation process associated to by Proposition 1. As , Equation (19) implies that ; thus, by the strong law of large numbers,
Let us define, for ,
By Proposition 2, is c.i.d. with respect to , so there exists by Lemmas 2.1 and 2.4 in [21] a random probability measure on such that, for every ,
Moreover, a.s. for every bounded measurable . Fix . By a monotone class argument, we can show that, for every bounded measurable ,
thus, a.s., and so is a uniformly integrable martingale. It follows from martingale convergence that, as ,
Equation (26) implies that . If, in addition, , then a.s. and . In fact, as long as , the sequence grows at the same rate as (22).
Proposition 3
([18], Lemma 6). Let and be diffuse. If , then
If , then may approach zero fast enough that we stop seeing new observations as . For example, let us consider random reinforcement with a totally skewed stable distribution for and . If , then , and we show that is stochastically bounded, which implies that converges to a finite limit.
Proposition 4.
Let η be a distribution with stability parameter , and be diffuse. Then, and
Proof.
From the properties of stable distributions, we obtain for every and, as a consequence,
By Theorem 5.4.1 in [28], , and so a.s. It follows for every that , which can be made arbitrarily small by taking M large enough. Regarding the second assertion, as , we have
□
Proposition 4 can be extended for any fat tailed reinforcement distribution by means of a generalized central limit theorem (see, e.g., [28] (p. 62)).
The rate of convergence of (18) and has already been studied for the model with independent weights under different assumptions, see, e.g., [19] (p. 1363), Examples 4.2 and 4.5 in the technical report to [18], Corollary 4.1 in [22] for . In the next theorem, we combine ideas from [18,20] to give a fairly general result.
Theorem 5.
Let . If , then
If, in addition, , then, with respect to the filtration generated by ,
Proof.
Let us define, for ,
The assertions in Theorem 5 have already been established by [18] when . In that case, Examples 4.2 and 4.5 in the technical report to [18] show that (29) is a consequence of the fact that
where , and (30) follows from
Replicating the approach of Proposition 9 in [20], we avoid using the assumption by conditioning on the sets , . By (26), , so (29) follows from (31) with
whereas (30) is, ultimately, a result of
□
4. Generalized Measure-Valued Pólya Urn Processes
The definition of an MVPP assumes that the law of the reinforcement is fixed, yet, in some situations, can itself be random (e.g., RRPP with exchangeable weights, see Section 4.1). To avoid measurability issues, we assume a parametric model for , with the parameter taking values in a Polish space .
Definition 3
(Generalized Measure-Valued Pólya Urn Process). Let V be a -valued random variable. A sequence of random finite measure on is called a generalized measure-valued Pólya urn process (GMVPP) with uncertainty parameter V, initial state and replacement rule if , and, for every ,
where is the transition probability kernel from to given by
and is the map .
It follows from Definition 3 that any GMVPP is a mixture of Markov chains with initial state and transition kernel . A separate modeling approach, which we do not examine here, defines a measure-valued Markov chain with transition kernel
In fact, some of the predictive constructions in [11,29] can be framed in such a way.
Theorem 1 extends to GMVPPs, provided that we condition all quantities on the parameter V. As a consequence, there exists a measurable function f from to and a random sequence such that
where , , , and, for ,
and
The definition of a randomly reinforced Pólya process is similarly generalized to cover the case of a random reinforcement distribution .
Definition 4
(Generalized Randomly Reinforced Pólya Process). We call a GMVPP with parameters a generalized randomly reinforced Pólya process (GRRPP) if there exists such that , where is the map .
For GRRPPs, the function f in the representation (32)–(34) can be written as
where h is a measurable function from to such that for all and , whenever . Letting , we obtain
where
and
The weights in (36) allow us to incorporate additional information about the observations . As an example, consider the problem of computer-based classification, where the output usually includes confidence scores, which reflect the software’s confidence that the classifications are correct. In analyzing the number and dimension of the types already discovered, or the probability of detecting a new type, a typical procedure would take into account only those classifications whose confidence scores are above a certain threshold. Alternatively, we could adopt a Bayesian perspective and weigh each classification according to its confidence score. Denoting by the sequence of classifications and confidence scores, we would model the distribution of the next classification by (36).
4.1. GRRPP with Exchangeable Weights
Let be a GRRPP with reinforcement distribution that does not depend on x. Then,
for some measurable function . The next result shows that the sequence is exchangeable with directing random measure . Moreover, is completely parameterized by .
Theorem 6.
A sequence of random finite measures is a GRRPP with parameters for if and only if and, for every ,
where , , is an exchangeable process with directing random measure , and is a sequence of -valued random variables such that and, for ,
Proof.
Let be a GRRPP with parameters , and consider the representation (35)–(37). Put and . It follows from (37) that
thus, is exchangeable. Moreover, , , so (38) follows from (36).
Conversely, suppose , where the process is as described. It follows from (38) and Theorem 8.12 in [25] that
Since is exchangeable with directing random measure , we have
Furthermore, is measurable with respect to the tail -field of , so, by (39),
Then, and, for ,
□
It follows from the proof of Theorem 6 that and, for ,
As and are both symmetric with respect to , then (42) is a symmetric function of . This is a necessary but not sufficient condition for to be exchangeable, see Proposition 3.2 and Example 3.1 in [2]. In Proposition 5, we show that is exchangeable if and only if either is degenerate or the weights are a.s. identical. On the other hand, for every , the sequence satisfies
and
By [24], Equations (43) and (44) are defining a process that is partially conditionally identity distributed (partially c.i.d.). Analogously to the c.i.d. case, partially c.i.d. processes preserve many of the properties of partially exchangeable sequences, see [24].
Proposition 5.
Under the conditions of Theorem 6, is partially c.i.d. Moreover, is exchangeable if and only if either is degenerate or a.s., . In that case, is partially exchangeable.
Proof.
It follows that is partially c.i.d. if and only if , , and (44) is true for every with . By hypothesis, is exchangeable and , so . Moreover, applying (39) repeatedly, we obtain
On the other hand, by (38),
Analogously, , which completes the proof of the first part.
If is degenerate, then is trivially exchangeable. If a.s. instead, then one can show that satisfies condition of Proposition 3.2 in [2], which, together with the symmetry of (42), implies by Theorem 3.1 in [2] that is exchangeable.
Conversely, suppose that is exchangeable. As is partially c.i.d., the predictive distributions (42) converge to a product random measure [24]. It follows from de Finetti’s theorem that is partially exchangeable, so, in particular,
However, from (36), so . Thus, for every bounded measurable function , there exists a measurable function such that
Integrating with respect to (38) and rearranging the terms, we obtain
Assume that is non-degenerate. Then, there is an such that ; e.g., take for some such that . It follows that
therefore,
In other words, there exists a measurable function such that a.s., and so a.s. by partial exchangeability. It follows from that, for every ,
thus, a.s. and, from exchangeability, a.s., . □
4.2. Asymptotic Properties of GRRPP with Exchangeable Weights
It follows from (38) that the GRRPP with exchangeable weights is a mixture of RRPPs with independent weights, with the mixing distribution affecting only the sequence . Thus, we expect that the results in Section 3.3 carry over to this more general setting. In this section, we concentrate on the behavior of and the sequence .
Assume that . If , then a.s., and, by the law of large numbers for exchangeable random variables (see [1], Section 2),
Then, if is diffuse, and a.s., so Theorem 1 in [27] implies
If , then may converge to a finite limit, as . For example, let us consider a strictly stable reinforcement distribution as in Proposition 4.
Proposition 6.
Let be a GRRPP with parameters such that V is a strictly positive random variable with , is diffuse, and , is a distribution with stability parameter . Then, and
Proof.
It follows from how the weights in the representation (35) are chosen that we can take
where , , and is the inverse of the distribution function. Then,
for some such that . It follows for every that , which can be made arbitrarily small by taking M large enough. Regarding the second assertion, as and by Theorem 5.4.1 in [28], we have
□
Extensions of Proposition 6 can be obtained by exploiting the central limit theorems for exchangeable random variables, which are found in [30,31].
5. Discussion
In this paper, we study the extension of randomly reinforced urns [17] to an unbounded set of possible colors. The resulting measure-valued urn process provides a predictive characterization of the law of an asymptotically exchangeable sequence of random variables, which corresponds to the observation process of an implied urn sampling scheme. In fact, the model (6)–(7) fits into a line of recent research, which explores efficient predictive constructions for fast online prediction or approximately-Bayesian solutions, see [11,29,32] and references therein. To that end, one direction for future work is to generalize the functional relationship in (7) and/or, as one referee suggested, to consider finitely-additive measures, along the lines discussed in [33].
We investigate the asymptotic properties of the sequences of predictive distributions and empirical frequencies of the observation process, and prove their convergence in total variation distance to a common random limit. The rate of convergence of their difference is given set-wise; so, another possible direction for future research is to consider a stronger distance. As far as we know, the topic of merging of the predictive and empirical distributions is largely unexplored. Within the relevant literature, we mention the works of [4,34], where the authors study the rate of convergence of the Wasserstein or Prokhorov distances under exchangeability, and the papers by Berti et al. [21], Berti et al. [35], who consider the c.i.d. case and regard the difference between the predictive and empirical measures as a map in the space of real bounded functions.
Author Contributions
Formal analysis, S.F., S.P., H.S.; writing—original draft preparation, S.F., S.P., H.S.; writing—review and editing, S.F., S.P., H.S. All authors have read and agreed to the published version of the manuscript.
Funding
This project has received funding from the European Research Council (ERC) under the European Union’s Horizon 2020 research and innovation programme (grant agreement No. 817257). H.S. was partially supported by the Bulgarian Ministry of Education and Science under the National Research Programme “Young scientists and postdoctoral students” approved by DCM No. 577/17.08.2018.
Institutional Review Board Statement
Not applicable.
Informed Consent Statement
Not applicable.
Data Availability Statement
Not applicable.
Acknowledgments
We wish to express our sincere gratitude to Regazzini for his deeply inspiring ideas and for instilling in us his passion for research. We thank the four anonymous referees for the valuable comments.
Conflicts of Interest
The authors declare no conflict of interest.
References
- Aldous, D.J. Exchangeability and related topics. École D’Été De Probab. De St.-Flour XIII 1983 1985, 1117, 1–198. [Google Scholar]
- Fortini, S.; Ladelli, L.; Regazzini, E. Exchangeability, predictive distributions and parametric models. Sankhya Ser. A 2000, 62, 86–109. [Google Scholar]
- Cifarelli, D.M.; Regazzini, E. De Finetti’s contribution to probability and statistics. Statist. Sci. 1996, 11, 253–282. [Google Scholar] [CrossRef] [Scilit]
- Cifarelli, D.M.; Dolera, E.; Regazzini, E. Frequentistic approximations to Bayesian prevision of exchangeable random elements. Int. J. Approx. Reason. 2016, 78, 138–152. [Google Scholar] [CrossRef] [Scilit]
- Fortini, S.; Petrone, S. Predictive distribution (de Finetti’s view). In Wiley StatsRef: Statistics Reference Online; Wiley Online Library, 2014; pp. 1–9. Available online: https://onlinelibrary.wiley.com/doi/full/10.1002/9781118445112.stat07831 (accessed on 4 October 2021).
- Regazzini, E. Old and recent results on the relationship between predictive inference and statistical modeling either in nonparametric or parametric form. In Bayesian Statistics 6; Oxford University Press: Oxford, UK, 1999; pp. 571–588. [Google Scholar]
- Fortini, S.; Petrone, S. Predictive construction of priors in Bayesian nonparametrics. Braz. J. Probab. Stat. 2012, 26, 423–449. [Google Scholar] [CrossRef] [Scilit]
- Mailler, C.; Marckert, J.F. Measure-valued Pólya urn processes. Electron. Commun. Probab. 2017, 22, 33. [Google Scholar] [CrossRef] [Scilit]
- Janson, S. Random replacements in Pólya urns with infinitely many colours. Electron. Commun. Probab. 2019, 24, 11. [Google Scholar] [CrossRef] [Scilit]
- Aletti, G.; Ghiglietti, A.; Rosenberger, W.F. Nonparametric covariate-adjusted reponse-adaptive design based on a functional urn model. Ann. Stat. 2018, 46, 3838–3866. [Google Scholar] [CrossRef] [Scilit]
- Fortini, S.; Petrone, S. Quasi-Bayes properties of a procedure for sequential learning in mixture models. J. R. Stat. Soc. Ser. B 2020, 82, 1087–1114. [Google Scholar] [CrossRef] [Scilit]
- Zhang, L.X.; Hu, F.; Cheung, S.H.; Chan, W.S. Immigrated urn models—Theoretical properties and applications. Ann. Stat. 2011, 39, 643–671. [Google Scholar] [CrossRef] [Scilit]
- Blackwell, D.; MacQueen, J.B. Ferguson distributions via Pólya urn schemes. Ann. Stat. 1973, 1, 353–355. [Google Scholar] [CrossRef] [Scilit]
- Bandyopadhyay, A.; Thacker, D. Pólya urn schemes with infinitely many colors. Bernoulli 2017, 23, 3243–3267. [Google Scholar] [CrossRef] [Scilit]
- Janson, S. A.s. convergence for infinite colour Pólya urns associated with random walks. Ark. Mat. 2021, 59, 87–123. [Google Scholar] [CrossRef] [Scilit]
- Mailler, C.; Villemonais, D. Stochastic approximation on non-compact measure spaces and application to measure-valued Pólya processes. Ann. Appl. Probab. 2020, 30, 2393–2438. [Google Scholar] [CrossRef] [Scilit]
- Muliere, P.; Paganoni, A.M.; Secchi, P. A randomly reinforced urn. J. Stat. Plan. Inference 2006, 136, 1853–1874. [Google Scholar] [CrossRef] [Scilit]
- Bassetti, F.; Crimaldi, I.; Leisen, F. Conditionally identically distributed species sampling sequences. Adv. Appl. Probab. 2010, 42, 433–459. [Google Scholar] [CrossRef] [Scilit]
- Berti, P.; Crimaldi, I.; Pratelli, L.; Rigo, P. Rate of convergence of predictive distributions for dependent data. Bernoulli 2009, 15, 1351–1367. [Google Scholar] [CrossRef] [Scilit]
- Berti, P.; Crimaldi, I.; Pratelli, L.; Rigo, P. Central limit theorems for multicolor urns with dominated colors. Stoch. Process. Appl. 2010, 120, 1473–1491. [Google Scholar] [CrossRef] [Scilit]
- Berti, P.; Pratelli, L.; Rigo, P. Limit theorems for a class of identically distributed random variables. Ann. Probab. 2004, 32, 2029–2052. [Google Scholar] [CrossRef] [Scilit]
- Crimaldi, I. An almost sure conditional convergence result and an application to a generalized Pólya urn. Int. Math. Forum 2009, 4, 1139–1156. [Google Scholar]
- Sariev, H.; Fortini, S.; Petrone, S. Infinite-Color Randomly Reinforced Urns with Dominant Colors. 2021. Preprint. Available online: https://arxiv.org/abs/2106.04307 (accessed on 4 October 2021).
- Fortini, S.; Petrone, S.; Sporysheva, P. On a notion of partially conditionally identically distributed sequences. Stoch. Process. Appl. 2018, 128, 819–846. [Google Scholar] [CrossRef] [Scilit]
- Kallenberg, O. Foundations of Modern Probability, 3rd ed.; Springer: New York, NY, USA, 2021. [Google Scholar]
- Häusler, E.; Luschgy, H. Stable Convergence and Stable Limit Theorems; Springer: Cham, Switzerland, 2015. [Google Scholar]
- Dubins, L.; Freedman, D. A sharper form of the Borel-Cantelli lemma and the strong law. Ann. Math. Stat. 1965, 36, 800–807. [Google Scholar] [CrossRef] [Scilit]
- Uchaikin, V.V.; Zolotarev, V.M. Chance and Stability: Stable Distributions and Their Applications; Walter de Gruyter: Berlin, Germany, 2011. [Google Scholar]
- Fong, E.; Holmes, C.; Walker, S. Martingale Posterior Distributions. 2021. Preprint. Available online: https://arxiv.org/abs/2103.15671 (accessed on 4 October 2021).
- Fortini, S.; Ladelli, L.; Regazzini, E. A central limit problem for partially exchangeable random variables. Theory Probab. Appl. 1997, 41, 224–246. [Google Scholar] [CrossRef] [Scilit]
- Fortini, S.; Ladelli, L.; Regazzini, E. Central limit theorem with exchangeable summands and mixtures of stable laws as limits. Boll. Unione Mat. Ital. 2012, 5, 515–542. [Google Scholar]
- Berti, P.; Dreassi, E.; Pratelli, L.; Rigo, P. A class of models for Bayesian predictive inference. Bernoulli 2021, 27, 702–726. [Google Scholar] [CrossRef] [Scilit]
- de Cooman, G.; Bock, J.D.; Diniz, M.A. Coherent predictive inference under exchangeability with imprecise probabilities. J. Artif. Intell. Res. 2015, 52, 1–95. [Google Scholar] [CrossRef] [Scilit]
- Dolera, E.; Regazzini, E. Uniform rates of the Glivenko-Cantelli convergence and their use in approximating Bayesian inferences. Bernoulli 2019, 25, 2982–3015. [Google Scholar] [CrossRef] [Scilit]
- Berti, P.; Pratelli, L.; Rigo, P. Limit theorems for empirical processes based on dependent data. Electron. J. Probab. 2012, 17, 1–18. [Google Scholar] [CrossRef] [Scilit]
Publisher’s Note: MDPI stays neutral with regard to jurisdictional claims in published maps and institutional affiliations. |
© 2021 by the authors. Licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license (https://creativecommons.org/licenses/by/4.0/).