Showing 1–2 of 2 results for author: Escriva, E
-
L'explicabilité au service de l'extraction de connaissances : application à des données médicales
Authors:
Robin Cugny,
Emmanuel Doumard,
Elodie Escriva,
Haomiao Wang
Abstract:
The use of machine learning has increased dramatically in the last decade. The lack of transparency is now a limiting factor, which the field of explainability wants to address. Furthermore, one of the challenges of data mining is to present the statistical relationships of a dataset when they can be highly non-linear. One of the strengths of supervised learning is its ability to find complex stat…
▽ More
The use of machine learning has increased dramatically in the last decade. The lack of transparency is now a limiting factor, which the field of explainability wants to address. Furthermore, one of the challenges of data mining is to present the statistical relationships of a dataset when they can be highly non-linear. One of the strengths of supervised learning is its ability to find complex statistical relationships that explainability allows to represent in an intelligible way. This paper shows that explanations can be used to extract knowledge from data and shows how feature selection, data subgroup analysis and selection of highly informative instances benefit from explanations. We then present a complete data processing pipeline using these methods on medical data. -- --
L'utilisation de l'apprentissage automatique a connu un bond cette dernière décennie. Le manque de transparence est aujourd'hui un frein, que le domaine de l'explicabilité veut résoudre. Par ailleurs, un des défis de l'exploration de données est de présenter les relations statistiques d'un jeu de données alors que celles-ci peuvent être hautement non-linéaires. Une des forces de l'apprentissage supervisé est sa capacité à trouver des relations statistiques complexes que l'explicabilité permet de représenter de manière intelligible. Ce papier montre que les explications permettent de faire de l'extraction de connaissance sur des données et comment la sélection de variables, l'analyse de sous-groupes de données et la sélection d'instances avec un fort pouvoir informatif bénéficient des explications. Nous présentons alors un pipeline complet de traitement des données utilisant ces méthodes pour l'exploration de données médicales.
△ Less
Submitted 27 February, 2023; v1 submitted 6 February, 2023;
originally announced February 2023.
-
Coalitional strategies for efficient individual prediction explanation
Authors:
Gabriel Ferrettini,
Elodie Escriva,
Julien Aligon,
Jean-Baptiste Excoffier,
Chantal Soulé-Dupuy
Abstract:
As Machine Learning (ML) is now widely applied in many domains, in both research and industry, an understanding of what is happening inside the black box is becoming a growing demand, especially by non-experts of these models. Several approaches had thus been developed to provide clear insights of a model prediction for a particular observation but at the cost of long computation time or restricti…
▽ More
As Machine Learning (ML) is now widely applied in many domains, in both research and industry, an understanding of what is happening inside the black box is becoming a growing demand, especially by non-experts of these models. Several approaches had thus been developed to provide clear insights of a model prediction for a particular observation but at the cost of long computation time or restrictive hypothesis that does not fully take into account interaction between attributes. This paper provides methods based on the detection of relevant groups of attributes -- named coalitions -- influencing a prediction and compares them with the literature. Our results show that these coalitional methods are more efficient than existing ones such as SHapley Additive exPlanation (SHAP). Computation time is shortened while preserving an acceptable accuracy of individual prediction explanations. Therefore, this enables wider practical use of explanation methods to increase trust between developed ML models, end-users, and whoever impacted by any decision where these models played a role.
△ Less
Submitted 1 April, 2021;
originally announced April 2021.