Literature DB >> 35465466

A possibilistic analogue to Bayes estimation with fuzzy data and its application in machine learning.

Mohsen Arefi1, Reinhard Viertl2, S Mahmoud Taheri3.   

Abstract

A Bayesian approach in a possibilistic context, when the available data for the underlying statistical model are fuzzy, is developed. The problem of point estimation with fuzzy data is studied in the possibilistic Bayesian approach introduced. For calculating the point estimation, we introduce a method without considering a loss function, and one considering a loss function. For the point estimation with a loss function, we first define a risk function based on a possibilistic posterior distribution, and then the unknown parameter is estimated based on such a risk function. Briefly, the present work extended the previous works in two directions: First the underlying model is assumed to be probabilistic rather than possibilistic, and second is that the problem of Bayes estimation is developed based on two cases of without and with considering loss function. Then, the applicability of the proposed approach to concept learning is investigated. Particularly, a naive possibility Bayes classifier is introduced and applied to some real-world concept learning problems.
© The Author(s), under exclusive licence to Springer-Verlag GmbH Germany, part of Springer Nature 2022.

Entities:  

Keywords:  Lifetime data; Maximum possibilistic posterior estimator; Point estimation; Possibilistic Bayes approach; Possibilistic posterior distribution; Risk function

Year:  2022        PMID: 35465466      PMCID: PMC9019817          DOI: 10.1007/s00500-022-07021-y

Source DB:  PubMed          Journal:  Soft comput        ISSN: 1432-7643            Impact factor:   3.732


Introduction

Bayesian inference for parametric statistical models is based on two assumptions: (1) The parameter of interest, , of the underlying model is of stochastic nature and has a probabilistic prior distribution, . (2) The available data related to the random variable X are precise. The first item is the most important point in the Bayesian paradigm. But, considering as a random variable with a probabilistic prior distribution is a matter of challenge between frequentist and Bayesian statisticians. On the other hand, concerning item 2 above, in many situations, the available data are vague (non-precise) rather than crisp (precise) (see Zimmermann 2000). In real-world problems, we often face such situations, situations in which the interesting parameter has a possibilistic nature and/or the data available are fuzzy. For instance, consider a medical study about the rate of people affected by virus COVID-19 in a certain population. Suppose that, based on the experiences, this rate is determined to be “about 0.15” (expressed by a fuzzy number with certain spreads as ambiguity). Moreover, it is completely reasonable that some medical tests about presence or absence of the virus yield uncertain results; for example, “with high possibility” she/he has the virus, the possibility is “about 0.8”s to have virus, etc. In such a study, which is very usual in practical problems, we need to use a possibilistic version of the Bayes formula based on fuzzy information. In this study, we propose a possibilistic version of the Bayesian approach, in which the underlying model is probabilistic, but the prior information about is formulated as a possibility distribution (called possibilistic prior distribution) rather than a probabilistic one. In addition, we consider that the data available for the random variable X are presented as fuzzy numbers rather than as crisp numbers. To this end, the likelihood function based on such a fuzzy-valued random sample is defined, and then, the extended likelihood function is combined with the possibilistic prior distribution, to obtain a possibilistic version of the posterior distribution. Such a possibilistic posterior distribution is based on a T-norm, so that it is flexible with respect to different situations. In this regard, the problem of point estimation without and with loss function is presented, too. Also, we present a possibilistic predictive distribution for predicting future fuzzy values of the underlying variable. In fact, we introduce a new Bayes classifier which is based on the possibility prior (rather than a probability one) and fuzzy data (rather than crisp ones). Therefore, the novelty of the current work compared to previous works is: The application of the proposed approach can be used in machine learning in which the Bayesian procedures present some methods to combine the prior knowledge and available data (see, e.g., Cui et al. 2013; He et al. 2014; Jiang et al. 2014; Subrahmanya and Shin 2013; Wang et al. 2018; You et al. 2019). Considering the underlying model to be probabilistic rather than possibilistic. Studying the problem of Bayes estimation based on two cases of without and with considering loss function, in a decision-theoretic framework. Introducing the possibilistic predictive distribution function, in a probabilistic–possibilistic context. Introducing the naive possibilistic Bayes classifier for crisp and fuzzy data. The topic of Bayesian inference in vague (non-precise) environments and/or in expert systems has been studied by some authors. Let us review some recent works in this topic. Lapointe and Bobée (2000) studied a possibilistic posterior distribution based on a possibilistic prior distribution and a possibilistic statistical model. Arefi and Taheri (2016) extended Lapointe and Bobée’s approach when the available data are fuzzy. They also studied the applicability of their approach in the field of concept learning. Taheri and Behboodian (2001, 2006) extended the Bayes approach to testing fuzzy hypotheses for both crisp (exact) and fuzzy (non-exact) data (also, see Torabi and Behboodian (2007). Some approaches of point estimations with fuzzy random samples are investigated by Akbari and Khanjari Sadegh (2012). Osoba et al. (2011) considered the problem of Bayesian inference using some certain fuzzy priors. Bacani and Barros (2017), Bardakhchyan (2017), Hareter and Viertl (2004), Mandal and Ranadive (2019), Osoba et al. (2012), Viertl (2011), Viertl and Hareter (2004), and Zhang and Chi (2008) investigated some aspects of the Bayes approaches in imprecise (fuzzy/vague) environments. This paper is organized as follows: In Sect. 2, we recall some basic concepts of possibility theory. Two new concepts, the possibilistic prior distribution and possibilistic posterior distribution, are introduced and investigated in Sect. 3. In Sect. 4, we study the problem of parameter estimation, in the possibilistic Bayes paradigm, without considering a loss function. In Sect. 5, using a loss function, the posterior risk function is initially defined and then, the problem of parameter estimation is developed based on this function. The concept of possibilistic predictive distributions for future values of fuzzy data is introduced in Sect. 6. Some applications of the proposed model in machine learning are explained in Sect.  7. In Sect. 8, the proposed approach is compared with some other approaches. A brief conclusion is provided in Sect. 9.

Possibility measure and possibility function

The Bayesian approach in statistics is fundamentally based on considering the parameter of interest (related to the statistical model ) as a random variable with a prior probabilistic distribution . However, in many problems, we have an imprecise (not necessarily stochastic) information on . In these cases, it is reasonable to consider as a possibilistic variable with a vague (fuzzy) prior information. Below, we recall two basic definitions of “Possibility Theory,” which we will need in the present article. The reader is referred to Dubois and Prade (1988), Klir and Folger (1988), Krtschmer (2004), and Zadeh (1968) for more details.

Definition 1

A possibility measure on a measurable space is defined to be a functionthat satisfies the following axioms Also, is said to be a possibility space. , for all , , (for every sequence ).

Definition 2

Suppose that is a possibility space. The function is a possibility function related to if it satisfiesIn the special case, .

Remark 1

The possibility measure and the possibility function are comparable to the probability measure and the probability density function, respectively. Note that if is a probability density function on , then we have

Remark 2

Note that the possibility measure and the probability measure are set functions, but the possibility function and the probability function are real-valued functions defined on .

Possibilistic posterior distribution with fuzzy data

In this section, we extend the concept of likelihood function to fuzzy data. Moreover, we define the possibilistic posterior distribution based on a possibilistic prior distribution, when the observations of the underlying model are fuzzy. In the following, we assume that is a measurable space, in which is the sample space. Also, is a probability space, where P is a probability measure on .

Definition 3

Let be a probability space. Let be a fuzzy-valued random sample of size n of X, associated with the probability density function (PDF) (or a probability mass function) , i.e., a sequence of fuzzy numbers as fuzzy realizations of the original random variable X. Then, the likelihood function based on such a fuzzy-valued random sample is defined bywhere is the Radon–Nikodym derivative of P with respect to (a -finite measure). The measure usually is “counting measure” or “Lebesgue measure,” and is the support of the random variable X.

Remark 3

It should be mentioned that when the available data are crisp numbers , then the above definition reduces to the ordinary definition of the likelihood function, i.e., . Note that the above extension is according to Zadeh’s definition (Zadeh 1968) for the probability of fuzzy events.

Definition 4

Let be a statistical model with unknown parameter . Suppose that the information about is formulated as a possibility function . This possibility function is called possibilistic prior distribution for .

Definition 5

Consider the fuzzy-valued random sample with the likelihood function . Suppose that the parameter has a possibilistic prior distribution . The possibilistic posterior distribution, under T-norm T(., .), is defined bywhere is called the marginal function (for obtaining a normal posterior distribution, i.e., ). The flowchart of procedure for calculating the possibilistic posterior distribution is given in Fig. 1.
Fig. 1

Flowchart of procedure for calculating the possibilistic posterior distribution

Flowchart of procedure for calculating the possibilistic posterior distribution

Remark 4

It is remarkable that in the above discussion, we have two kinds of uncertainty: the probabilistic one which is due to the random variable and the possibilistic one which is arisen from the possibilistic prior. The difficulty is to combine these two different kinds of uncertainties to update the information about the unknown parameter . Note that although there have been some attempts to define the conditional possibilities (e.g., Coletti and Vantaggi (2009), De Baets et al. (1999), Ferracuti and Vantaggi (2006), Hisdal (1978), Kramosil (1998), and Nguyen (1978)), to the best knowledge of the authors, there has not been any work in the problem of updating a possibility information using probabilistic fuzzy data.

Remark 5

The above definition is, in some sense, consistent to some definitions for conditional possibility. First, note that the marginal function is analogous to a corresponding relation for probabilities in probability theory for which the operators summation and product are used instead of and T(., .), respectively (see Eq. (33), in Nguyen (1978), Eq. (2.8) in Hisdal (1978), and Eq. (6) in Kramosil (1998)). A second item is related to use a T-norm in the numerator in Eq. (2). In this regard, see the discussion in Sect. 3 in De Baets et al. (1999) and Sects. 3 and 4 in Coletti and Vantaggi (2009).

Example 1

The data in Table 1 (centers of the fuzzy numbers) show the lifetimes (in 1000 km) of front disk brake pads on a randomly selected set of 40 cars (same model) that were monitored by a dealer network (see, Lawless, 2003, p. 337). Suppose that the lifetime of the front disk brake pad has an exponential distribution with density functionwhere is the mean lifetime of the front disk brake pad. An expert believes that the value of the variable lies in the interval [40,50] with a possibility of one. Moreover, he/she believes it is possible that is smaller than 40, but never below 30, and bigger than 50, but never above 60. We use a trapezoidal fuzzy number to model the possibilistic prior distribution based on the expert opinion as follows:In practice, measuring the lifetime of a disk may not yield an exact result. A disk may work perfectly over a certain period but be braking for some time and finally be unusable at a certain time. So, such data may be reported as imprecise quantities. Assume that the lifetimes of front disk brake pads are reported as fuzzy numbers in Table 1. In fact, imprecision is formulated by fuzzy numbers , with , , as follows:
Table 1

Lifetimes of front disk brake pads of cars in Example 1

\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$${\tiny No.}$$\end{document}No.\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$${\widetilde{X}}_i$$\end{document}X~i\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$${\tiny No.}$$\end{document}No.\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$${\widetilde{X}}_i$$\end{document}X~i\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$${\tiny No.}$$\end{document}No.\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$${\widetilde{X}}_i$$\end{document}X~i\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$${\tiny No.}$$\end{document}No.\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$${\widetilde{X}}_i$$\end{document}X~i
1\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(86.2,4.3)_R$$\end{document}(86.2,4.3)R11\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(36.7,1.8)_R$$\end{document}(36.7,1.8)R21\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(61.5,3.1)_R$$\end{document}(61.5,3.1)R31\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(45.9,2.3)_R$$\end{document}(45.9,2.3)R
2\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(38.4,1.9)_R$$\end{document}(38.4,1.9)R12\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(22.6,1.1)_R$$\end{document}(22.6,1.1)R22\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(42.7,2.1)_R$$\end{document}(42.7,2.1)R32\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(50.6,2.5)_R$$\end{document}(50.6,2.5)R
3\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(45.5,2.3)_R$$\end{document}(45.5,2.3)R13\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(81.7,4.1)_R$$\end{document}(81.7,4.1)R23\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(46.9,2.3)_R$$\end{document}(46.9,2.3)R33\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(59.0,3.0)_R$$\end{document}(59.0,3.0)R
4\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(22.7,1.1)_R$$\end{document}(22.7,1.1)R14\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(102.5,5.1)_R$$\end{document}(102.5,5.1)R24\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(33.9,1.7)_R$$\end{document}(33.9,1.7)R34\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(62.4,3.1)_R$$\end{document}(62.4,3.1)R
5\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(48.8,2.4)_R$$\end{document}(48.8,2.4)R15\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(28.4,1.4)_R$$\end{document}(28.4,1.4)R25\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(54.2,2.7)_R$$\end{document}(54.2,2.7)R35\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(34.4,1.7)_R$$\end{document}(34.4,1.7)R
6\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(42.8,2.1)_R$$\end{document}(42.8,2.1)R16\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(31.7,1.6)_R$$\end{document}(31.7,1.6)R26\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(81.3,4.1)_R$$\end{document}(81.3,4.1)R36\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(50.2,2.5)_R$$\end{document}(50.2,2.5)R
7\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(73.1,3.7)_R$$\end{document}(73.1,3.7)R17\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(52.1,2.6)_R$$\end{document}(52.1,2.6)R27\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(51.6,2.6)_R$$\end{document}(51.6,2.6)R37\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(50.7,2.5)_R$$\end{document}(50.7,2.5)R
8\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(59.8,3.0)_R$$\end{document}(59.8,3.0)R18\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(56.4,2.8)_R$$\end{document}(56.4,2.8)R28\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(38.8,1.9)_R$$\end{document}(38.8,1.9)R38\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(64.5,3.2)_R$$\end{document}(64.5,3.2)R
9\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(45.1,2.3)_R$$\end{document}(45.1,2.3)R19\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(42.2,2.1)_R$$\end{document}(42.2,2.1)R29\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(53.6,2.7)_R$$\end{document}(53.6,2.7)R39\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(33.8,1.7)_R$$\end{document}(33.8,1.7)R
10\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(41.0,2.1)_R$$\end{document}(41.0,2.1)R20\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(40.0,2.0)_R$$\end{document}(40.0,2.0)R30\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(80.6,4.0)_R$$\end{document}(80.6,4.0)R40\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$(56.7,2.8)_R$$\end{document}(56.7,2.8)R
Lifetimes of front disk brake pads of cars in Example 1 The likelihood function based on such fuzzy data is calculated asHence, the possibilistic posterior distribution based on the product T-norm, , is obtained as follows (see Fig. 2)where
Fig. 2

The possibilistic posterior distribution in Example 1

The possibilistic posterior distribution in Example 1 The possibilistic prior distribution in Example 2

Example 2

Consider a study for estimating the proportion of a kind of tree in a forest, which is infected with a plague. We take a sample of 3 trees and examine each tree separately for the presence of plague. Suppose that we have no precise mechanisms for an exact distinction between the presence or absence of plague, but we can identify the information with a fuzzy set on . The usual model for this problem starts from the Bernoulli experiment X associated with the presence of plague (). Suppose that, based on some prior information, we know that is approximately 0.2. In this case, a suitable possibilistic distribution for modeling such prior information is the following fuzzy number (see Fig. 3)
Fig. 3

The possibilistic prior distribution in Example 2

Suppose that, based on a random sample of size , we observe the following fuzzy dataThe likelihood function based on these fuzzy data is ()Based on the product T-norm, the possibilistic posterior distribution would be (see Fig. 4)which is maximized in . Also, the marginal function for the observed data is
Fig. 4

The possibilistic posterior distribution in Example 2

The possibilistic posterior distribution in Example 2

Point estimation without loss functions

Definition 6

Consider a possibilistic posterior distribution based on an observed fuzzy random sample. Then, , as a function on the fuzzy-valued random sample to R, is called the maximum possibilistic posterior estimation of () if The above definition is similar to the definition of the maximum Bayesian likelihood estimator, defined as the posterior mode, in the probabilistic approach (see, Robert (Robert 2001), p. 166). In spite of Definition 6, we can use defuzzification methods (see, i.e., Ross (1995)) to obtain an estimation for based on the possibilistic posterior distribution.

Definition 7

Let be a defuzzification operator. The point estimation based on the possibilistic posterior distribution is defined as i) Consider the “Center of Gravity” defuzzification method. Based on this method, the point estimation is obtained asii) Consider the “Center of Area” defuzzification method. Based on this method, the point estimation is obtained asoriii) Consider the “Supremum of Center” defuzzification method. Based on this method, the point estimation is obtained as

Example 3

Consider the possibilistic posterior distribution in Example 1. The point estimations of , based on the above different methods, are calculated as follows: The maximum possibilistic posterior estimation: The center of gravity estimation: The center of area estimation: hence The supremum of center estimation:

Example 4

Consider the possibilistic posterior distribution in Example 2. The point estimations of , based on the above different methods, are calculated as follows: The maximum possibilistic posterior estimation: The center of gravity estimation: The center of area estimation: hence The supremum of center estimation:

Point estimation based on a loss functions

In a decision-theoretic analysis, the quality of a decision function (here, a point estimation) is quantified by the risk function of the decision function. In this section, we first define a risk function based on the possibilistic posterior distribution , and then, the point estimation (decision function) is obtained based on this risk function. Let be the parameter space. Any function is called a loss function, where D is the space of possible decisions (here, the space of all point estimations of ).

Definition 8

The posterior risk function with fuzzy data for the estimation (decision function) under the probability density function (or probability mass function) and with the possibilistic prior distribution , based on a loss function , is defined as

Definition 9

The estimation based on the loss function and the possibilistic posterior distribution based on the fuzzy data are called a possibilistic Bayes estimation ifwhere D is the set of all estimations for .

Example 5

Consider the possibilistic posterior distribution in Example 1. The point estimation based on a loss function is calculated as follows: (i) Based on the quadratic loss function , the posterior risk function is as in Fig. 5. Here, the possibilistic Bayes estimation of , for which the posterior risk function is minimized, is .
Fig. 5

The posterior risk function in Example 5

(ii) Based on the loss function , the posterior risk function in terms of d is as in Fig. 5. Here, the possibilistic Bayes estimation of , for which the posterior risk function is minimized, is The posterior risk function in Example 5

Example 6

Consider the possibilistic posterior distribution in Example 2. The point estimation based on a loss function is calculated as follows: i) Based on the quadratic loss function , the posterior risk function in terms of d for the random sample of size is shown in Fig. 6. Here, the possibilistic Bayes estimation of , for which the posterior risk function is minimized, is .
Fig. 6

The posterior risk functions in Example 6

ii) Based on the loss function , the posterior risk function for the random sample of size is shown in Figure 6. Here, the possibilistic Bayes estimation of , for which the posterior risk function is minimized, is . The posterior risk functions in Example 6

Possibilistic predictive distribution

The future values of fuzzy data are provided by a possibility function, called “possibilistic predictive distribution” (ppd).

Definition 10

Let be the possibilistic posterior distribution based on the fuzzy data . Also, suppose that the random variable Y has the pdf . The possibilistic predictive distribution is defined aswhere .

Remark 6

In the Bayesian approach, the predictive distribution for future values is introduced aswhere is the probabilistic posterior distribution based on the precise data and is the probability density function for the random variable Y (see, Robert, 2001, p. 22). We list some differences of the possibilistic predictive distribution and the predictive distribution as follows: 1- The predictive distribution is based on precise data, but the possibilistic predictive distribution is based on fuzzy data. 2- The predictive distribution is a probability distribution, but the possibilistic predictive distribution is a possibility distribution. 3- The possibilistic predictive distribution is flexible under different T-norms.

Example 7

Consider the possibilistic posterior distribution based on the product T-norm, in Example 1. The possibilistic predictive distribution based on the product T-norm, , is obtained as follows (see Fig. 7)The possibilistic predictive distribution can be expressed by “approximately 50.” Table 2 shows some values of the possibilistic predictive distribution. The new front disk brake pad works 50 (in 1000 km) with possibility 1, and it works less/more than 50 with less possibility. (For example, with possibility 90%, the lifetime of the disk is equal to 30.)
Fig. 7

The possibilistic predictive distribution in Example 7

Table 2

Some values of possibilistic predictive distribution in Example 7

t 1020304050
ppd0.44510.72880.89510.97711
Some values of possibilistic predictive distribution in Example 7 The possibilistic predictive distribution in Example 7

Example 8

Consider Example 2. The possibilistic predictive distribution based on some T-norms is obtained as follows: i) Based on the product T-norm, , the possibilistic posterior distribution is obtained as follows (see Example 2) Hence, the possibilistic predictive distribution isii) Based on the minimum T-norm, , the possibilistic posterior distribution is obtained as follows (see Fig. 8)
Fig. 8

Possibilistic posterior distribution based on the minimum T-norm in Example 8

Possibilistic posterior distribution based on the minimum T-norm in Example 8 Hence, the possibilistic predictive distribution is given by (see Fig. 9)
Fig. 9

Possibilistic predictive distribution based on the minimum T-norm in Example 8

Possibilistic predictive distribution based on the minimum T-norm in Example 8

Application in machine learning

A common problem in machine learning is the concept learning problem, for which the Bayes classifier is a well-known method, (see, e.g., He et al. (2014), Mitchell (1997), and Wang et al. (2014)). Any system that classifies new instances according to the following equation is called a Bayes optimal classifier or Bayes optimal learner (see Mitchell (1997))where , are the attribute values, and is the probabilistic posterior function. Using Bayes rule, we can rewrite the above expression in the following way,Particularly, the naive Bayes classifier assumes that the attribute values are conditionally independent given the target value. In this case, the so-called naive Bayes classifier is defined asNow, based on the results in Sect. 3, we investigate and develop the Bayes classifier in the possibility environment. Particularly, we investigate naive possibilistic Bayes classifier both for crisp information and for fuzzy information.

Naive possibilistic Bayes classifier for crisp data

The possibilistic Bayesian approach to classify the new crisp instance is to assign the most target value of maximum a possibilistic posteriori () given the crisp attribute values as follows,where and is the possibilistic prior distribution. If the attribute values are conditionally independent given the target value, then the joint conditional probability is obtained as follows,The following example is a possibilistic version of the example in Mitchell (1997), p. 157, about a medical diagnosis.

Example 9

In this example, we investigate a learning task in a possibilistic context rather than probabilistic one. Consider a target attribute Medical Diagnosis with two values: (1) The patient has a particular form of cancer, and (2) the patient does not. The available data are from a laboratory test with two possible outcomes: (the result of test is positive) and (the result of test is negative.) Formally, the set of target values is , where and . Based on the prior information, a patient has cancer and nocancer with the possibilities and , respectively. Here, the amount reflects the consistency of the symptoms of cancer in this patient (note that these values are not related to the randomness or relative frequency, which are usual in the probabilistic context). The test returns a correct positive result in only 98% of the cases in which the disease is actually present and a correct negative result in only 97% of the cases in which the disease is not present. In other cases, the test returns the opposite result. The conditional probabilities are summarized asNow, suppose that the laboratory test of a new patient returns a positive result. Our task is to diagnose the target value (the patient has cancer or not) based on the possibilistic posterior distribution. Here, the possibilistic posterior distribution is obtained as follows,The results based on different T-norms are given in Table 3. Based on this information, . Hence, if the result of test in a new patient is positive, then our best estimate is that he/she has cancer.
Table 3

The possibilistic posterior distribution in Example 9

T(ab)\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\theta =\theta _1$$\end{document}θ=θ1\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\theta =\theta _2$$\end{document}θ=θ2
\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\min (a,b)$$\end{document}min(a,b)10.30
a.b10.29
\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\max (0,a+b-1)$$\end{document}max(0,a+b-1)10
\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\frac{a.b}{1+(1-a)(1-b)}$$\end{document}a.b1+(1-a)(1-b)10.28
\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\frac{a.b}{a+b-ab}$$\end{document}a.ba+b-ab10.30
The possibilistic posterior distribution in Example 9

Naive possibilistic Bayes classifier for fuzzy data

In this subsection, we extend the method in the previous subsection to the case when the observed data are fuzzy rather than crisp. We use the possibilistic posterior distribution to classify the new fuzzy instance. Generally, suppose that the new fuzzy instance is available and we want to classify it in the target space . We first calculate the possibilistic posterior distribution based on Definition 5, and then, the most target value of the possibilistic posteriori () is obtained as follows,The following example is a possibilistic version of the example in Mitchell (1997), p. 178, for the target concept “PlayTennis” (see also Arefi and Taheri (2016)).

Example 10

Consider the learning task represented by the training examples of Table 4. Here, the target attribute PlayTennis, which can have values yes or no for different Saturday mornings, is to be predicted based on other attributes of the morning in question. Formally, the set of target values is , where and . Based on the prior information, the player selects and with the possibilities and , respectively. Let each day be described by the following attribute values,Based on the training examples given in Table 4, the conditional probabilities are obtained as in Table 5.
Table 4

Training examples for the target concept PlayTennis in Example 10

DayOutlookTemperatureHumidityWindPlayTennis
DlSunnyHotHighWeakNo
D2SunnyHotHighStrongNo
D3OvercastHotHighWeakYes
D4RainMildHighWeakYes
D5RainCoolNormalWeakYes
D6RainCoolNormalStrongNo
D7OvercastCoolNormalStrongYes
D8SunnyMildHighWeakNo
D9SunnyCoolNormalWeakYes
Dl0RainMildNormalWeakYes
Dl1SunnyMildNormalStrongYes
Dl2OvercastMildHighStrongYes
Dl3OvercastHotNormalWeakYes
Dl4RainMildHighStrongNo
Table 5

Conditional probabilities in Example 10

\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Outlook=Sunny|Yes)=\frac{2}{9}$$\end{document}P(Outlook=Sunny|Yes)=29\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Temp.=Hot|Yes)=\frac{2}{9}$$\end{document}P(Temp.=Hot|Yes)=29
\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Outlook=Overcast|Yes)=\frac{4}{9}$$\end{document}P(Outlook=Overcast|Yes)=49\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Temp.=Mild|Yes)=\frac{4}{9}$$\end{document}P(Temp.=Mild|Yes)=49
\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Outlook=Rain|Yes)=\frac{3}{9}$$\end{document}P(Outlook=Rain|Yes)=39\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Temp.=Cool|Yes)=\frac{3}{9}$$\end{document}P(Temp.=Cool|Yes)=39
\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Outlook=Sunny|No)=\frac{3}{5}$$\end{document}P(Outlook=Sunny|No)=35\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Temp.=Hot|No)=\frac{2}{5}$$\end{document}P(Temp.=Hot|No)=25
\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Outlook=Overcast|No)=0$$\end{document}P(Outlook=Overcast|No)=0\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Temp.=Mild|No)=\frac{2}{5}$$\end{document}P(Temp.=Mild|No)=25
\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Outlook=Rain|No)=\frac{2}{5}$$\end{document}P(Outlook=Rain|No)=25\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Temp.=Cool|No)=\frac{1}{5}$$\end{document}P(Temp.=Cool|No)=15
\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Humidity= High|Yes)=\frac{3}{9}$$\end{document}P(Humidity=High|Yes)=39\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Wind=Weak|Yes)=\frac{6}{9}$$\end{document}P(Wind=Weak|Yes)=69
\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Humidity= Normal|Yes)=\frac{6}{9}$$\end{document}P(Humidity=Normal|Yes)=69\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Wind=Strong|Yes)=\frac{3}{9}$$\end{document}P(Wind=Strong|Yes)=39
\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Humidity= High|No)=\frac{4}{5}$$\end{document}P(Humidity=High|No)=45\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Wind=Weak|No)=\frac{2}{5}$$\end{document}P(Wind=Weak|No)=25
\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Humidity= Normal|No)=\frac{1}{5}$$\end{document}P(Humidity=Normal|No)=15\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$P(Wind=Strong|No)=\frac{3}{5}$$\end{document}P(Wind=Strong|No)=35
Training examples for the target concept PlayTennis in Example 10 Conditional probabilities in Example 10 Suppose that we observe the following fuzzy information for a new instance,whereBased on Definition 3, the likelihood function under is calculated aswhere Finally, the possibilistic posterior distribution is obtained as and where . Here,Therefore, the naive possibilistic Bayes classifier assigns the target value to the new instance.

Comparison with other works

Comparison with the Gil et al.’s work

Gil et al. (1985) studied the point estimation problem with fuzzy information. They defined a probabilistic posterior distribution as followswhere is defined in Definition 3, and is a probabilistic prior distribution. Also, the posterior risk function based on the loss function for the estimated parameter is defined as followsIn contrast with the Gil et al.’s method, it should be noted that: In our method, the parameter has a possibilistic nature, but it has a probabilistic nature in Gil et al.’s. Hence, in our method, the posterior distribution will also be of a possibilistic–probabilistic nature, whereas it will be of a probabilistic nature in Gil et al.’s. We define the posterior distribution based on a T-norm. Hence, the posterior distribution is flexible under the different T-norms.

Comparison with Viertl’s work

Viertl (2011) proposed a Bayesian inference method in the situation of fuzzy prior information and fuzzy data. He supposed that a fuzzy sample consists of fuzzy numbers with related membership functions . First, he combined the membership functions of the fuzzy data to obtain a membership function of the combined fuzzy sample . Then, based on the extension principle on the likelihood function , the characterizing (membership) function of the fuzzy value , based on fuzzy data , is given byIn order to obtain the generalized Bayes’ theorem, he used the -level functions and of and and of . Hence, a fuzzy posterior density is calculated by its -level functions and as followsBoth the methods proposed in this article and in Viertl’s work are based on the likelihood function, but our method has the following advantages: In our method, the parameter has a possibilistic nature, but it has a probabilistic nature in Viertl’s. Hence, the posterior distribution based on our method is possibilistic–probabilistic in nature, whereas it is fuzzy probabilistic in the Viertl method. The posterior distribution proposed in the present paper is based on T-norms. Hence, the posterior distribution is flexible using different T-norms. In our method, we investigate the case in which a loss function has been considered in estimating the unknown parameter.

Comparison with Arefi and Taheri’s work

Arefi and Taheri (2016) studied a Bayesian approach, when the available data are fuzzy. This approach is an extended version of Lapointe and Bobée’s approach (2000) with fuzzy data. In this approach, based on a possibilistic model of fuzzy sample and a prior possibility distribution , the possibilistic posterior distribution is defined, when the available data are fuzzy. In the proposed approach, the possibility of a fuzzy sample is first defined aswhere T(., .) is a T-norm. Then, based on a possibilistic prior distribution , the possibilistic posterior distribution is defined bywhere , , and is a residuated implication operator (see also Lapointe and Bobée (2000)). Some cases are the following: The approaches introduced in Lapointe and Bobée (2000) (with crisp data) and Arefi and Taheri (2016) (with fuzzy data) are full possibilistic nature, i.e., the model of data and the prior distribution have possibilistic distributions, but our method is a combined version of probability and possibility, i.e., the model of data is probabilistic and the prior distribution is possibilistic. In this paper, we investigate the case in which a loss function has been considered in estimating the unknown parameter.

Conclusions

In the Bayesian method for estimation of an unknown parameter, it is often difficult to assuming the prior information in stochastic terms. In addition, sometimes the observed data are non-precise (fuzzy) rather than precise (crisp). In this paper, by introducing the concept of likelihood function for fuzzy data of a probabilistic model, a possibilistic approach was described for dealing with such situations. The present approach employed the possibility distribution for modeling the prior information. Then, based on the possibilistic posterior distribution, some methods were proposed to estimate the unknown parameter of interest with/without a loss function. The proposed approach was applied to a well-known problem in machine learning called the concept learning problem. Particularly, the naive possibility Bayes classifier is introduced and applied to some real-world concept learning problems. In summary, we can express the contributions of the paper as follows: It should be mentioned that, according to the proposed approach, choosing (or constructing) a suitable prior distribution is an important issue (see, e.g., Hill and Spall (1994), Chen et al. (2000), and Dubois (2006)), which is a potential subject of future research. Introducing a new approach to combination of probabilistic model and possibilistic information to get a possibilistic posterior distribution function. Considering the observed data of the underlying model as fuzzy rather than crisp. Employing a decision-theoretic approach in which the problem of estimation was studied based on two cases of without and with considering loss function. In addition, the maximum possibilistic posterior estimation (as a duality of the concept of maximum Bayesian likelihood estimation) is introduced. Introducing the possibilistic predictive distribution function, based on probabilistic model, fuzzy observations, and possibilistic prior information. Investigating some applications of the proposed Bayesian approach in the problem of concept learning, by introducing the naive possibilistic Bayes classifier for crisp data and fuzzy data.
  2 in total

1.  Non-Naive Bayesian Classifiers for Classification Problems With Continuous Attributes.

Authors:  Xi-Zhao Wang; Yu-Lin He; Debby D Wang
Journal:  IEEE Trans Cybern       Date:  2013-02-26       Impact factor: 11.448

2.  Bayesian inference with adaptive fuzzy priors and likelihoods.

Authors:  Osonde Osoba; Sanya Mitaim; Bart Kosko
Journal:  IEEE Trans Syst Man Cybern B Cybern       Date:  2011-04-07
  2 in total

北京卡尤迪生物科技股份有限公司 © 2022-2023.