Digitala Vetenskapliga Arkivet

Ändra sökning
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf
Trustworthy explanations: Improved decision support through well-calibrated uncertainty quantification
Jönköping University, Internationella Handelshögskolan, IHH, Informatik.ORCID-id: 0000-0001-9633-0423
2023 (Engelska)Doktorsavhandling, sammanläggning (Övrigt vetenskapligt)
Abstract [en]

The use of Artificial Intelligence (AI) has transformed fields like disease diagnosis and defence. Utilising sophisticated Machine Learning (ML) models, AI predicts future events based on historical data, introducing complexity that challenges understanding and decision-making. Previous research emphasizes users’ difficulty discerning when to trust predictions due to model complexity, underscoring addressing model complexity and providing transparent explanations as pivotal for facilitating high-quality decisions.

Many ML models offer probability estimates for predictions, commonly used in methods providing explanations to guide users on prediction confidence. However, these probabilities often do not accurately reflect the actual distribution in the data, leading to potential user misinterpretation of prediction trustworthiness. Additionally, most explanation methods fail to convey whether the model’s probability is linked to any uncertainty, further diminishing the reliability of the explanations.

Evaluating the quality of explanations for decision support is challenging, and although highlighted as essential in research, there are no benchmark criteria for comparative evaluations.

This thesis introduces an innovative explanation method that generates reliable explanations, incorporating uncertainty information supporting users in determining when to trust the model’s predictions. The thesis also outlines strategies for evaluating explanation quality and facilitating comparative evaluations. Through empirical evaluations and user studies, the thesis provides practical insights to support decision-making utilising complex ML models.

Abstract [sv]

Användningen av Artificiell intelligens (AI) har förändrat områden som diagnosticering av sjukdomar och försvar. Genom att använda sofistikerade maskininlärningsmodeller predicerar AI framtida händelser baserat på historisk data. Modellernas komplexitet resulterar samtidigt i utmanande beslutsprocesser när orsakerna till prediktionerna är svårbegripliga. Tidigare forskning pekar på användares problem att avgöra prediktioners tillförlitlighet på grund av modellkomplexitet och belyser vikten av att tillhandahålla transparenta förklaringar för att underlätta högkvalitativa beslut.

Många maskininlärningsmodeller erbjuder sannolikhetsuppskattningar för prediktionerna, vilket vanligtvis används i metoder som ger förklaringar för att vägleda användare om prediktionernas tillförlitlighet. Dessa sannolikheter återspeglar dock ofta inte de faktiska fördelningarna i datat, vilket kan leda till att användare felaktigt tolkar prediktioner som tillförlitliga. Därutöver förmedlar de flesta förklaringsmetoder inte om prediktionernas sannolikheter är kopplade till någon osäkerhet, vilket minskar tillförlitligheten hos förklaringarna.

Att utvärdera kvaliteten på förklaringar för beslutsstöd är utmanande, och även om det har betonats som avgörande i forskning finns det inga benchmark-kriterier för jämförande utvärderingar.

Denna avhandling introducerar en innovativ förklaringsmetod som genererar tillförlitliga förklaringar vilka inkluderar osäkerhetsinformation för att stödja användare att avgöra när man kan lita på modellens prediktioner. Avhandlingen ger också förslag på strategier för att utvärdera kvaliteten på förklaringar och underlätta jämförande utvärderingar. Genom empiriska utvärderingar och användarstudier ger avhandlingen praktiska insikter för att stödja beslutsfattande användande komplexa maskininlärningsmodeller.

Ort, förlag, år, upplaga, sidor
Jönköping: Jönköping University, Jönköping International Business School , 2023. , s. 72
Serie
JIBS Dissertation Series, ISSN 1403-0470 ; 159
Nyckelord [en]
Explainable Artificial Intelligence, Interpretable Machine Learning, Decision Support Systems, Uncertainty Estimation, Explanation Methods
Nationell ämneskategori
Systemvetenskap, informationssystem och informatik med samhällsvetenskaplig inriktning Datavetenskap (datalogi)
Identifikatorer
URN: urn:nbn:se:hj:diva-62865ISBN: 978-91-7914-031-1 (tryckt)ISBN: 978-91-7914-032-8 (digital)OAI: oai:DiVA.org:hj-62865DiVA, id: diva2:1810440
Disputation
2023-12-12, B1014, Jönköping International Business School, Jönköping, 13:15 (Engelska)
Opponent
Handledare
Tillgänglig från: 2023-11-08 Skapad: 2023-11-08 Senast uppdaterad: 2025-12-02Bibliografiskt granskad
Delarbeten
1. Interpretable instance-based text classification for social science research projects
Öppna denna publikation i ny flik eller fönster >>Interpretable instance-based text classification for social science research projects
2018 (Engelska)Ingår i: Archives of Data Science, Series A, ISSN 2363-9881, Vol. 5, nr 1Artikel i tidskrift (Refereegranskat) Published
Abstract [en]

In this study, two groups of respondents have evaluated explanations generated from an instance-based explanation method called WITE (Weighted Instance-based Text Explanations). One group consisted of 24 non-experts who answered a web survey about the words characterising the concepts of the classes and the other group consisted of three senior researchers and three respondents from a media house in Sweden who answered a questionnaire with open questions. The data used originates from one of the researchers’ project on media consumption in Sweden. The results from the non-experts indicate that WITE identified many words that corresponded to the human understanding but also included some insignificant or contrary words as important. In the results from the expert evaluation, there were indications that there is a risk that the explanations could persuade the users of the correctness of a prediction, even if it is incorrect. Consequently, the study indicates that an explanation method could be seen as a new actor which is able to persuade and interact with the humans and cause a change in the results of the classification of a text.

Ort, förlag, år, upplaga, sidor
KIT – Die Forschungsuniversität in der Helmholtz-Gemeinschaft, 2018
Nationell ämneskategori
Data- och informationsvetenskap
Identifikatorer
urn:nbn:se:hj:diva-49118 (URN)10.5445/KSP/1000087327/15 (DOI)
Tillgänglig från: 2020-06-10 Skapad: 2020-06-10 Senast uppdaterad: 2025-10-13Bibliografiskt granskad
2. A meta survey of quality evaluation criteria in explanation methods
Öppna denna publikation i ny flik eller fönster >>A meta survey of quality evaluation criteria in explanation methods
2022 (Engelska)Ingår i: Intelligent Information Systems: CAiSE Forum 2022, Leuven, Belgium, June 6–10, 2022, Proceedings / [ed] J. De Weerdt, Jochen & A. Polyvyanyy, Cham: Springer, 2022, s. 55-63Konferensbidrag, Publicerat paper (Refereegranskat)
Abstract [en]

The evaluation of explanation methods has become a significant issue in explainable artificial intelligence (XAI) due to the recent surge of opaque AI models in decision support systems (DSS). Explanations are essential for bias detection and control of uncertainty since most accurate AI models are opaque with low transparency and comprehensibility. There are numerous criteria to choose from when evaluating explanation method quality. However, since existing criteria focus on evaluating single explanation methods, it is not obvious how to compare the quality of different methods.

Ort, förlag, år, upplaga, sidor
Cham: Springer, 2022
Serie
Lecture Notes in Business Information Processing, ISSN 1865-1348, E-ISSN 1865-1356 ; 452
Nyckelord
Explanation method, Evaluation metric, Explainable artificial intelligence, Evaluation of explainability, Comparative evaluations
Nationell ämneskategori
Datavetenskap (datalogi)
Identifikatorer
urn:nbn:se:hj:diva-57114 (URN)10.1007/978-3-031-07481-3_7 (DOI)000871754800007 ()2-s2.0-85131293203 (Scopus ID)978-3-031-07480-6 (ISBN)978-3-031-07481-3 (ISBN)
Konferens
CAiSE Forum 2022, Leuven, Belgium, June 6–10, 2022
Forskningsfinansiär
KK-stiftelsen
Tillgänglig från: 2022-06-13 Skapad: 2022-06-13 Senast uppdaterad: 2026-01-16Bibliografiskt granskad
3. On the Definition of Appropriate Trust and the Tools that Come with it
Öppna denna publikation i ny flik eller fönster >>On the Definition of Appropriate Trust and the Tools that Come with it
2023 (Engelska)Ingår i: 2023 Congress in Computer Science, Computer Engineering, & Applied Computing (CSCE), Institute of Electrical and Electronics Engineers (IEEE), 2023, s. 1555-1562Konferensbidrag, Publicerat paper (Övrigt vetenskapligt)
Abstract [en]

Evaluating the efficiency of human-AI interactions is challenging, including subjective and objective quality aspects. With the focus on the human experience of the explanations, evaluations of explanation methods have become mostly subjective, making comparative evaluations almost impossible and highly linked to the individual user. However, it is commonly agreed that one aspect of explanation quality is how effectively the user can detect if the predictions are trustworthy and correct, i.e., if the explanations can increase the user's appropriate trust in the model. This paper starts with the definitions of appropriate trust from the literature. It compares the definitions with model performance evaluation, showing the strong similarities between appropriate trust and model performance evaluation. The paper's main contribution is a novel approach to evaluating appropriate trust by taking advantage of the likenesses between definitions. The paper offers several straightforward evaluation methods for different aspects of user performance, including suggesting a method for measuring uncertainty and appropriate trust in regression.

Ort, förlag, år, upplaga, sidor
Institute of Electrical and Electronics Engineers (IEEE), 2023
Nyckelord
Appropriate Trust, Calibrated Trust, Comparative Evaluations, Evaluation of Explanations, Explanation Methods, Metrics, XAI
Nationell ämneskategori
Människa-datorinteraktion (interaktionsdesign)
Identifikatorer
urn:nbn:se:hj:diva-64149 (URN)10.1109/CSCE60160.2023.00256 (DOI)2-s2.0-85191166512 (Scopus ID)979-8-3503-2759-5 (ISBN)
Konferens
2023 Congress in Computer Science, Computer Engineering, and Applied Computing, CSCE 2023 Las Vegas 24 July 2023 through 27 July 2023
Forskningsfinansiär
KK-stiftelsen, 20160035
Tillgänglig från: 2024-05-07 Skapad: 2024-05-07 Senast uppdaterad: 2025-12-02Bibliografiskt granskad
4. Investigating the impact of calibration on the quality of explanations
Öppna denna publikation i ny flik eller fönster >>Investigating the impact of calibration on the quality of explanations
2023 (Engelska)Ingår i: Annals of Mathematics and Artificial Intelligence, ISSN 1012-2443, E-ISSN 1573-7470Artikel i tidskrift (Refereegranskat) Epub ahead of print
Abstract [en]

Predictive models used in Decision Support Systems (DSS) are often requested to explain the reasoning to users. Explanations of instances consist of two parts; the predicted label with an associated certainty and a set of weights, one per feature, describing how each feature contributes to the prediction for the particular instance. In techniques like Local Interpretable Model-agnostic Explanations (LIME), the probability estimate from the underlying model is used as a measurement of certainty; consequently, the feature weights represent how each feature contributes to the probability estimate. It is, however, well-known that probability estimates from classifiers are often poorly calibrated, i.e., the probability estimates do not correspond to the actual probabilities of being correct. With this in mind, explanations from techniques like LIME risk becoming misleading since the feature weights will only describe how each feature contributes to the possibly inaccurate probability estimate. This paper investigates the impact of calibrating predictive models before applying LIME. The study includes 25 benchmark data sets, using Random forest and Extreme Gradient Boosting (xGBoost) as learners and Venn-Abers and Platt scaling as calibration methods. Results from the study show that explanations of better calibrated models are themselves better calibrated, with ECE and log loss for the explanations after calibration becoming more conformed to the model ECE and log loss. The conclusion is that calibration makes the models and the explanations better by accurately representing reality.

Ort, förlag, år, upplaga, sidor
Springer, 2023
Nyckelord
Calibration, Decision support systems, Explainable artificial intelligence, Predicting with confidence, Uncertainty in explanations, Venn Abers
Nationell ämneskategori
Data- och informationsvetenskap
Identifikatorer
urn:nbn:se:hj:diva-60033 (URN)10.1007/s10472-023-09837-2 (DOI)000948763400001 ()2-s2.0-85149810932 (Scopus ID)HOA;;870772 (Lokalt ID)HOA;;870772 (Arkivnummer)HOA;;870772 (OAI)
Forskningsfinansiär
KK-stiftelsen
Tillgänglig från: 2023-03-27 Skapad: 2023-03-27 Senast uppdaterad: 2025-10-13
5. Calibrated explanations: With uncertainty information and counterfactuals
Öppna denna publikation i ny flik eller fönster >>Calibrated explanations: With uncertainty information and counterfactuals
2024 (Engelska)Ingår i: Expert systems with applications, ISSN 0957-4174, E-ISSN 1873-6793, Vol. 246, artikel-id 123154Artikel i tidskrift (Refereegranskat) Published
Abstract [en]

While local explanations for AI models can offer insights into individual predictions, such as feature importance, they are plagued by issues like instability. The unreliability of feature weights, often skewed due to poorly calibrated ML models, deepens these challenges. Moreover, the critical aspect of feature importance uncertainty remains mostly unaddressed in Explainable AI (XAI). The novel feature importance explanation method presented in this paper, called Calibrated Explanations (CE), is designed to tackle these issues head-on. Built on the foundation of Venn-Abers, CE not only calibrates the underlying model but also delivers reliable feature importance explanations with an exact definition of the feature weights. CE goes beyond conventional solutions by addressing output uncertainty. It accomplishes this by providing uncertainty quantification for both feature weights and the model’s probability estimates. Additionally, CE is model-agnostic, featuring easily comprehensible conditional rules and the ability to generate counterfactual explanations with embedded uncertainty quantification. Results from an evaluation with 25 benchmark datasets underscore the efficacy of CE, making it stand as a fast, reliable, stable, and robust solution.

Ort, förlag, år, upplaga, sidor
Elsevier, 2024
Nyckelord
Explainable AI, Feature Importance, Calibrated Explanations, Venn-Abers, Uncertainty Quantification, Counterfactual Explanations
Nationell ämneskategori
Systemvetenskap, informationssystem och informatik
Identifikatorer
urn:nbn:se:hj:diva-62864 (URN)10.1016/j.eswa.2024.123154 (DOI)001164089000001 ()2-s2.0-85182588063 (Scopus ID)HOA;;1810433 (Lokalt ID)HOA;;1810433 (Arkivnummer)HOA;;1810433 (OAI)
Forskningsfinansiär
KK-stiftelsen, 20160035
Anmärkning

Included in doctoral thesis in manuscript form.

Tillgänglig från: 2023-11-08 Skapad: 2023-11-08 Senast uppdaterad: 2025-10-13Bibliografiskt granskad

Open Access i DiVA

Kappa(4387 kB)765 nedladdningar
Filinformation
Filnamn FULLTEXT01.pdfFilstorlek 4387 kBChecksumma SHA-512
a19e72154ebf280b3ec89abf383995b5b3202d00d92dd6f58024cc52b30a117dcc58224ce1439b5624698720294562a2288aa92d317f2c08270bd4ac2fca601e
Typ fulltextMimetyp application/pdf

Sök vidare i DiVA

Av författaren/redaktören
Löfström, Helena
Av organisationen
IHH, Informatik
Systemvetenskap, informationssystem och informatik med samhällsvetenskaplig inriktningDatavetenskap (datalogi)

Sök vidare utanför DiVA

GoogleGoogle Scholar
Totalt: 766 nedladdningar
Antalet nedladdningar är summan av nedladdningar för alla fulltexter. Det kan inkludera t.ex tidigare versioner som nu inte längre är tillgängliga.

isbn
urn-nbn

Altmetricpoäng

isbn
urn-nbn
Totalt: 5233 träffar
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf