Home
  About Conference
  General Information
  About Barcelona
  Venue
  Topics
  Call for papers
  Instructions for Authors
  Special-Invited Session
  Plenary Sessions
  Tutorials. Entrance free
  Deadlines
  Registration
  Program
  Student grants
  Committees
  Hotel reservation
  Sponsors
 
 

Contact us
 

Intranet

   Printable version

EUSFLAT-LFA 2005
Joint 4th EUSFLAT & 11th LFA Conference
7-9 September, 2005 - Barcelona, Spain

Plenary Sessions

Some Applications of Fuzzy Logic in Data Mining and Information Retrieval

Wednesday September 7, 2005 – 9:00 to 9:50

Bernadette Bouchon-Meunier
Computer Science Laboratory LIP6
Université Paris 6-CNRS
Bernadette.Bouchon-Meunier@lip6.fr

Abstract: Data mining and information retrieval are two domains difficult to cope with for various reasons. First, most of the databases are complex, large, and contain heterogeneous, imprecise, vague, uncertain, incomplete data. Furthermore, the queries may be imprecise or subjective in the case of information retrieval, the mining results must be easily understandable by a user in the case of data mining or knowledge discovery. Fuzzy logic provides an interesting tool for such tasks, mainly because of its capability to represent imperfect information, for instance by means of imprecise categories, measures of resemblance or aggregation methods. We will focus our study on two main paradigms, underlying most of the real world problems we have been facing. The first one is fuzzy inductive learning, based on fuzzy decision trees and specific measures of discrimination or entropy, providing an efficient way to extract relevant information from training sets of examples described by means of numerical, symbolic, approximate or linguistic values of attributes. The second one is a general framework for measures of comparison, compatible with Tversky's contrast model, providing tools to identify similar or dissimilar descriptions of objects, and leading to the construction of fuzzy prototypes of classes, for instance in a case-based reasoning or a classification approach.

We present some cases where these paradigms have been exploited among others to manage various types of data such as medical images, multimedia information, sensorial criteria, risk factors.

About the speaker: Bernadette Bouchon-Meunier is a director of research at the National Center for Scientific Research (CNRS), head of the department of Machine learning in the Computer Science Laboratory of the University Paris 6.

She graduated from the Ecole Normale Supérieure at Cachan and she received the degrees of B.S. in Mathematics and in Computer Science, Ph.D. in Applied Mathematics and D.Sc. in Computer Science (1978) from the University of Paris.

Editor-in-chief of the International Journal of Uncertainty, Fuzziness and Knowledge-based systems (published by World Scientific) since 1993, she is also the editor or co-editor of eighteen books (published by Springer Verlag, Physica Verlag, Hermès, Lavoisier, Elsevier, World Scientific) and the author or co-author of four books in French on Fuzzy Logic and Uncertainty Management in Artificial Intelligence.

B. Bouchon-Meunier is the co-founder and co-chairperson of the International Conference on Information Processing and Management of Uncertainty in Knowledge-based systems (IPMU) held every other year since 1986.

She is an IEEE senior member and an IFSA fellow. She is in charge of the Women in Computational Intelligence Committee of the IEEE Computational Intelligence Society.

Fuzzy Multivalued Logics and T-Norms

Wednesday September 7, 2005 – 16:50 to 17:40

Francesc Esteva
Artificial Intelligence Research Institute (IIIA)
Universitat Autonoma de Barcelona
esteva@iiia.csic.es

Abstract: Fuzzy logic in narrow sense: A summary on t-norm-based logics
In the seventies and eighties operations over [0,1] corresponding to connectives of a multi-valued logic has been studied and used to define Fuzzy Set operations. From then t-norms, t-conorms, negation and implication functions have been widely studied and applied. The nineties were the start point of the study of the kernel of fuzzy logic (fuzzy logics in narrow sense in Zadeh nomenclature), a multivalued residuated logic whose semantic is defined by the structure defined over [0,1] by a t-norm (modelling the «and») its residuum (modelling the implication) and the negation defined as «imply zero». The definition of BL (for basic fuzzy logic) by P.Hŕjek (proved to be the logic of continuous t-norm and their residua) and of MTL (for monoidal t-norm-based logic) by Esteva-Godo (proved to be the logic of left-continuous t-norm and their residua) are the basis for any further study of these family of logics. In the conference a summary of the recent result in these logics and their relation with t-norms will be given.

About the speaker: Francesc Esteva received his B.Sc. in Mathematics in 1969 and his Ph.D. also in Mathematics in 1974, both from the University of Barcelona. He was Full Professor at the Technical University of Catalunya and he is currently the Director of the Artificial Intelligence Research Institute (IIIA) of the Spanish Research Council (CSIC).

He has published about 100 papers in Journals and Conferences mainly in the area of approximate reasoning and its applications to different areas like knowledge-based systems and case-based reasoning. Multi-valued Logics, Modal and Multi-modal Systems and mainly Fuzzy Logic in narrow sense are the topics he has been working in.

He has been President of the Spanish Association for Fuzzy Logic and Technology and the first President of EUSFLAT. He is member of the editorial board of the Handbook of Fuzzy Computation published by IOS Press, area editor of IJUFKBS and member of the editorial board of some other journals.

Three Visual Cluster Validity Methods for Object and Relational Data

Thursday September 8, 2005 – 10:40 to 11:30

Jim Bezdek
Computer Science Department
University of West Florida
jbezdek@uwf.edu

Abstract: This talk is about three visual cluster validity methods developed by the authors that can be used for both object and relational data sets. The original method VAT (visual assessment of clustering tendency) works nicely for dissimilarity data up to about n = 5000 objects, but VAT quickly bumps up against storage and resolution limits, and is of limited utility for large data sets. The second method in the family is sVAT (scalable VAT). This algorithm generates a sample taken from the rows (or columns) of very large, square, relational data, and builds a VAT image on the sample. sVAT is shown to be exact when the data contain compact, separated clusters in the sense of Dunn, and sVAT can be used with arbitrarily large data sets. The third method is called coVAT. This method can be used to estimate the number of clusters in objects represented by the rows and columns of a rectangular dissimilarity matrix, as well as the number of mixed and unmixed clusters in the union of the row and column objects. We show examples of using coVAT to detect the number of clusters in the four canonical clustering problems that are present in rectangular relational data.

About the speaker: Jim Bezdek received the BSCE from the U. of Nevada (Reno) in 1969, and the Ph.D. in Applied Math from Cornell in 1973. He is currently the Nystul Professor of Computer Science at the University of West Florida. His previous experience includes the directorship of Boeing's HTC Inf. Proc. Lab, and a term as head of Computer Science at the University of South Carolina.

Jim's interests include woodworking, optimization, motorcycles, pattern recognition, fishing, vision and image processing, skiing, computational neural networks, blues music, medical applications and cigars. Jim is the founding editor of the Int'l. Jo. Approximate Reasoning and the IEEE Transactions on Fuzzy Systems. He has been a distinguished lecturer for the IEEE and ACM, is an IEEE fellow, and was president of the IEEE Neural Networks Council in 1997-1998.

Current research topics: multiple prototype classifier designs, mixed fuzzy-possibilistic c-means clustering models, rule extraction with clustering, generalized nearest prototype classifier networks, fusion of heterogeneous fuzzy data, target recognition with LADAR data, mammographic image analysis, topics in cluster validity, robotic control models, fuzzy learning vector quantization, clustering with genetic algorithms, and acceleration of image processing algorithms.

Fuzzy Logic as a Basis for Theory of Precisiation of Meaning (TPM)

Thursday September 8, 2005 – 16:30 to 17:30

Lotfi A. Zadeh
Professor in the Graduate School, Computer Science Division
Department of Electrical Engineering and Computer Sciences
University of California
Berkeley, CA 94720 -1776
Director, Berkeley Initiative in Soft Computing (BISC)
zadeh@eecs.berkeley.edu

Abstract: The concept of precision is ubiquitous. It has a position of centrality not only in science but, more generally, in almost all domains of human reasoning and discourse. And yet, there are some fundamental issues which relate to the concept of precision that have received very little, if any, attention. One such issue is that of precisiation of natural languages. Another basic issue relates to definition of concepts. The Theory of Precisiation of Meaning (TPM) may be viewed as an attempt to construct a conceptual framework for dealing with these and related issues.

Much of human knowledge is expressed in a natural language. Close relationship between human knowledge and natural language is one of the principal reasons why mechanization of natural language understanding has long been one of the important objectives of AI.

Over the years, impressive progress has been made toward achievement of this objective. But what is widely unrecognized is that there is a fundamental limitation to what can be achieved through the use of commonly-employed methods of meaning representation. The aim of this paper is, first, to highlight this limitation and, second, to suggest ways of removing it.

To understand the nature of the limitation, two facts have to be considered. First, a natural language, NL, is basically a system for describing perceptions; and second, perceptions are intrinsically imprecise, reflecting the bounded ability of human sensory organs, and ultimately the brain, to resolve detail and store information. More specifically, perceptions are f-granular in the sense that (a) the boundaries of perceived classes are unsharp (fuzzy); and (b) the values of perceived attributes are granular, with a granule being a clump of values drawn together by indistinguishability, similarity, proximity or functionality.

Imprecision of perceptions is passed on to natural languages. What this implies is that imprecision of semantics of natural languages is rooted in imprecision of perceptions. Semantic imprecision of natural languages is not a problem for humans, but it is a major problem for machines.

To clarify the issue, let p be a proposition, concept, question or command. For p to be understood by a machine, it must be precisiated, that is, expressed in a mathematically well-defined language. A precisiated form of p, Pre(p), will be referred to as a precisiand of p and will be denoted as p*.

To precisiate p we can employ a number of meaning-representation languages, e.g., Prolog, predicate logic, semantic networks, conceptual graphs, LISP, SQL, etc. The commonly-used meaning-representation languages are bivalent, i.e., are based on bivalent logic. Are we moving in the right direction when we employ such languages for mechanization of natural language understanding? The answer is: No. The reason relates to an important issue which we have not addressed: cointension of p*, with intension used in its logical sense as attribute-based meaning. More specifically, cointension is a measure of the goodness of fit of the intension of a precisiand, p*, to the intended intension of p. Thus, cointension is a desideratum of precisiation. What this implies is that mechanization of natural language understanding requires more than precisiation—it requires cointensive precisiation. Note that definition is a form of precisiation. In plain words, a definition is cointensive if its meaning is a good fit to the intended meaning of the definiendum.

Here is where the fundamental limitation which was alluded to earlier comes into view. In a natural language, NL, most p’s are fuzzy, that is, in one way or another, are a matter of degree. Simple examples: propositions «most Swedes are tall;» and «overeating causes obesity;» concepts «mountain;» and «honest;» question «is Albert honest?» and command «take a few steps.»

Employment of commonly-used meaning-representation languages to precisiate a fuzzy p leads to a bivalent (crisp) precisiand p*. The problem is that, in general, a bivalent p* is not cointensive. As a simple illustration, consider the concept of recession. The standard definition of recession is: A period of general economic decline; specifically, a decline in GDP for two or more consecutive quarters. Similarly, a definition of bear market is: We classify a bear market as a 30 percent decline after 50 days, or a 13 percent decline after 145 days. (Robert Shuster, Ned Davis Research.) Clearly, neither definition is cointensive. Another example is the classical definition of stability. Consider a ball of diameter D which is placed on an open bottle whose mouth is of diameter d. If D is somewhat larger than d, the configuration is stable: Obviously, as D increases, the configuration becomes less and less stable. But, according to Lyapounov’s bivalent definition of stability, the configuration is stable for all values of D greater than d. This contradiction is characteristic of crisp definitions of fuzzy concepts—a well-known example of which is the Greek sorites (heap) paradox. The magnitude of the problem becomes apparent where we consider that many concepts in scientific theories are fuzzy, but are defined and treated as if they are crisp. This is particularly true in fields in which the concepts which are defined are descriptions of perceptions.

To remove the fundamental limitation, bivalence must be abandoned. Furthermore, new concepts, ideas and tools must be developed and deployed to deal with the issues of cointensive precisiation, definability and deduction. The principal tools are Precisiated Natural Language (PNL); Protoform Theory (PFT); and the Generalized Theory of Uncertainty (GTU). These tools form the core of what may be called the Computational Theory of Precisiation of Meaning (TPM). The centerpiece of TPM is the concept of a generalized constraint.

A generalized constraint, GC, is an expression of the form X is R, where X is the constrained variable, R is a constraining relation and r is an indexing variable whose value defines the modality of the constraint, that is, its semantics. The principal constraints are possibilistic (r=blank); veristic (r=v); probabilistic (r=p); usuality (r=u); random set (r=rs); fuzzy graph (r=fg); bimodal (r=bm); and group (r=g). The primary constraints are probabilistic, possibilistic and veristic. The standard constraints, which are the constraints that underlie standard logic and probability theory, are bivalent possibilistic, bivalent veristic and probabilistic. The semantics of a generalized constraint, GC, is defined by its test-score function, ts(u), in which u is an object to which the constraint applies and ts(u) is the degree to which u satisfies GC. The set of all generalized constraints together with the rules governing combination, qualification and constraint propagation constitutes the Generalized Constraint Language (GCL). Similarly, the set of all standard constraints together with the rules governing combination, probability qualification and constraint propagation constitutes the Standard Constraint Language (SCL).

The concept of a generalized constraint plays a key role in TPM by providing a basis for precisiation of meaning. More specifically, if p is a proposition or a concept, its precisiand, Pre(p), is represented as a generalized constraint, GC. Thus, Pre(p)=GC. In this sense, the concept of a generalized constraint may be viewed on a bridge from natural languages to mathematics.

Representing precisiands of p as elements of GCL is the pivotal idea in TPM. Each precisiand is associated with the degree to which it is cointensive with p. Given p, the problem is that of finding those precisiands which are cointensive, that is, have a high degree of cointension. If p is a fuzzy proposition or concept then in general there are no cointensive precisiands in SCL.

In TPM, a refinement of the concept of precisiation is needed. First, a differentiation is made between v-precision (precision in value) and m-precision (precision in meaning). For example, proposition p: X is 5, is both v-precise and m-precise; p: X is between 5 and 7, is v-imprecise and m-precise; and p: X is small, is both v-imprecise and m-imprecise; however, p can be m-precisiated by defining small as a fuzzy set or a probability distribution. A perception is v-imprecise and its description is m-imprecise. PNL makes it possible to m-precisiate descriptions of perceptions.

Granulation of a variable, e.g., representing the values of age as young, middle-aged and old, may be viewed as a form of v-imprecisiation. Granulation plays an important role in human cognition by serving as a means of (a) exploiting tolerance for imprecision through omission of irrelevant information. (b) lowering precision and thereby lowering cost; and (c) facilitating understanding and articulation. In fuzzy logic, granulation is m-precisiated through the use of the concept of a linguistic variable.

Further refinement of the concept of precisiation relates to two modalities of m-precisiation: (a) human-oriented, denoted as mh-precisiation; and (b) machine-oriented, denoted as mm-precisiation. Unless stated to the contrary, in TPM, precisiation should be understood as mm-precisiation.

In a bimodal dictionary or lexicon, the first entry, p, is a concept or proposition; the second entry, p*, is mh-precisiand of p; and the third entry is mm-precisiand of p. To illustrate, the entries for recession might read: mh-precisiand—a period of general economic decline; and mm-precisiand—a decline in GDP for two or more consecutive quarters.

Viewed in a broader perspective, what should be noted is that precisiation of meaning is not the ultimate goal—it is an intermediate goal. Once precisiation of meaning is achieved, the next goal is that of deduction of decision-relevant information. The ultimate goal is decision.

In TPM, a concept which plays a key role in deduction is that of a protoform—an abbreviation for prototypical form. Briefly, a protoform of p, PF(p), is an abstracted summary of p. For example, for p: Monika is young, PF(p) is A(B) is C, where A is abstraction of age, B is abstraction of Monika and C is abstraction of young. Basically, a protoform of p serves to place in evidence the deep semantic structure of p.

In TPM, a deduction rule has two components: symbolic and computational. For the most part, a deduction rule is a rule which governs generalized constraint propagation from premises to conclusion. Protoformal deduction in TPM may be viewed as a computationally-extended version of symbolic reasoning.

There is a simple analogy which helps to understand the meaning of cointensive precisiation. Specifically, a proposition, p, is analogous to a system, S; precisiation is analogous to modelization; a precisiand, expressed as a generalized constraint, GC(p), is analogous to a model, M(S), of S; test-score function is analogous to input-output relation; cointensive precisiand is analogous to well-fitting model; GCL is analogous to the class of all fuzzy-logic-based systems; and SCL is analogous to the subclass of all bivalent-logic-based systems. To say that, in general, a cointensive definition of a fuzzy concept cannot be formulated within the conceptual structure of bivalent logic and probability theory, is similar to saying that, in general, a linear system cannot be a well-fitting model of a nonlinear system.

Ramifications of the concept of cointensive precisiation extend well beyond mechanization of natural language understanding. A broader basic issue is validity of definitions in scientific theories, especially in the realms of human-oriented fields such as law, economics, medicine, psychology and linguistics. More specifically, the concept of cointensive precisiation calls into question the validity of many of the existing definitions of basic concepts—among them the concepts of causality, relevance, independence, stability, complexity and optimality.

About the speaker: Lotfi A. Zadeh is a Professor in the Graduate School, Computer Science Division, Department of EECS, University of California, Berkeley. In addition, he is serving as the Director of BISC (Berkeley Initiative in Soft Computing).

Lotfi Zadeh is an alumnus of the University of Tehran, MIT and Columbia University. He held visiting appointments at the Institute for Advanced Study, Princeton, NJ; MIT; IBM Research Laboratory, San Jose, CA; SRI International, Menlo Park, CA; and the Center for the Study of Language and Information, Stanford University. His earlier work was concerned in the main with systems analysis, decision analysis and information systems. His current research is focused on fuzzy logic, computing with words and soft computing, which is a coalition of fuzzy logic, neurocomputing, evolutionary computing, probabilistic computing and parts of machine learning.

Lotfi Zadeh is a Fellow of the IEEE, AAAS, ACM, AAAI, and IFSA. He is a member of the National Academy of Engineering and a Foreign Member of the Russian Academy of Natural Sciences and the Finnish Academy of Sciences. He is a recipient of the IEEE Education Medal, the IEEE Richard W. Hamming Medal, the IEEE Medal of Honor, the ASME Rufus Oldenburger Medal, the B. Bolzano Medal of the Czech Academy of Sciences, the Kampe de Feriet Medal, the AACC Richard E. Bellman Control Heritage Award, the Grigore Moisil Prize, the Honda Prize, the Okawa Prize, the AIM Information Science Award, the IEEE-SMC J. P. Wohl Career Achievement Award, the SOFT Scientific Contribution Memorial Award of the Japan Society for Fuzzy Theory, the IEEE Millennium Medal, the ACM 2001 Allen Newell Award, the Norbert Wiener Award of the Systems, Man and Cybernetics Society, Civitate Honoris Causa by Budapest Tech (BT) Polytechnical Institution, Budapest, Hungary, the V. Kaufmann Prize, International Association for Fuzzy-Set Management and Economy (SIGEF), other awards and twenty-three honorary doctorates. He has published extensively on a wide variety of subjects relating to the conception, design and analysis of information/intelligent systems, and is serving on the editorial boards of over fifty journals.

Fuzzy Logic and Natural Language

Friday September 9, 2005 – 14:40 to 15:30

Vilem Novak
University of Ostrava
Institute for Research and Applications of Fuzzy Modeling
30. dubna 22, 701 03 Ostrava 1, Czech Republic
Vilem.Novak@osu.cz

Abstract: One of the most important features of fuzzy set theory and fuzzy logic is their power to model semantics of a certain part of natural language. The results are further elaborated in topics such as computing with words which encompasses computing with fuzzy numbers, the theory of evaluating linguistic expressions as well as that of fuzzy IF-THEN rules and approximate reasoning.

In the talk, we will precisely characterize a small part of natural language expressions called evaluating linguistic expressions. Recall that these are expressions such "very large, extremely deep, roughly one thousand, more or less hot", etc., i.e. the expressions considered in many applications of fuzzy logic. Furthermore, we will extend this theory to fuzzy IF-THEN rules, namely, we will show that they can be taken as genuine conditional expressions of natural language. Sets of them form linguistic descriptions of some decision or classification situation, or control strategy. A logical deduction based on linguistic descriptions has a power to mimic human way of reasoning. The discussed methods can be taken as a possible concretization of the general paradigm of Precisiated Natural Language as proposed by L. A. Zadeh.

We will also mention, how evaluating expressions can be used in applications. First, we will demonstrate an application in geology where the goal was to mimic the way how geologist determines rock sequences using which movement of the ancient sea level can be estimated. The source of information was vague geologist's description of the method in natural language.

The second is application in decision-making when a system of linguistic descriptions can help in finding optimal decision on the basis of subjective knowledge of the problem characterized in natural language only. The difference in importance of various criteria is in the language naturally expressed without necessity to use sophisticated and somewhat artificial methods for weights assignment.

About the speaker: Education: Mining University, Ostrava, Dipl. Ing., 1969-1975. Charles University, Prague, PhD in Mathematical logic, 1982-1988. Institute for System Science of the Polish Academy of Sciences, DSc (doctor of sciences), 1995. Palacky University, Olomouc, Associate Professor, 1995. Masaryk University, Brno, Professor, 2001.

Vilem Novak is Author or co-author of more than 100 publications from various fields in fuzzy set theory, fuzzy logic, fuzzy control and computer science. He is also author and/or editor of several books.

His main topics of interest are: Fuzzy Logic, Approximate Reasoning, Modeling of Natural Language Semantics and Fuzzy Control.

Problems? Comments? Please, contact to webadmin.esaii@upc.es
Last update: September 1, 2005