Bridging trust and performance in intelligent systems: Hybrid explainable AI approaches for interpreting large language models

Large Language Models (LLMs) achieve state-of-the-art performance across natural language processing tasks but remain opaque, limiting adoption in high-stakes domains that demand accountability and transparency. This paper introduces a hybrid explainability framework that integrates saliency-based attribution, causal reasoning, and user-centered visualization into a unified, efficiency-aware pipeline. Unlike prior single-method approaches such as LIME, SHAP, or attention visualization, the framework provides explanations that are both technically faithful and accessible to human evaluators. The framework was systematically evaluated on benchmark datasets (GLUE, SQuAD, IMDB, and domain-specific corpora) and tested on representative architectures (BERT, T5, GPT, and LLaMA). Results show up to a 15–20% improvement in fidelity compared to attention-based methods. Fidelity was measured using standardized insertion and deletion metrics across all benchmark datasets using a consistent evaluation protocol, ensuring objective and comparable assessment of explanation faithfulness across different LLM architectures. The proposed framework also achieved higher clarity and trust ratings in user studies while introducing less than 25% additional computational overhead. Case studies in sentiment analysis and question answering further demonstrate that hybrid explanations produce precise, intuitive reasoning paths that outperform existing baselines. The main contributions are: (1) a multi-method pipeline that reconciles the trade-off between faithfulness and interpretability; (2) a human-centered evaluation showing hybrid explanations are more trustworthy than single techniques; and (3) an efficiency-aware design indicating the potential suitability of the proposed framework for practical applications in domains such as healthcare, finance, and law. By aligning methodological rigor with societal and regulatory demands, this study advances both the practice and theory of explainable AI, positioning hybrid XAI as a pathway toward responsible LLM adoption.

Authors

Institutions

Publication Details

Journal
PLoS ONE
Published
2026-09-15
DOI
https://doi.org/10.1371/journal.pone.0343472
Primary Topic
Explainable Artificial Intelligence (XAI)
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

Bridging trust and performance in intelligent systems: Hybrid explainable AI approaches for interpreting large language models

Tamije Selvy P., Arul Selvam P.
PLoS ONE
Explainable Artificial Intelligence (XAI)
article

Bridging trust and performance in intelligent systems: Hybrid explainable AI approaches for interpreting large language models

Tamije Selvy P., Arul Selvam P.
article en

Abstract

Large Language Models (LLMs) achieve state-of-the-art performance across natural language processing tasks but remain opaque, limiting adoption in high-stakes domains that demand accountability and transparency. This paper introduces a hybrid explainability framework that integrates saliency-based attribution, causal reasoning, and user-centered visualization into a unified, efficiency-aware pipeline. Unlike prior single-method approaches such as LIME, SHAP, or attention visualization, the framework provides explanations that are both technically faithful and accessible to human evaluators. The framework was systematically evaluated on benchmark datasets (GLUE, SQuAD, IMDB, and domain-specific corpora) and tested on representative architectures (BERT, T5, GPT, and LLaMA). Results show up to a 15–20% improvement in fidelity compared to attention-based methods. Fidelity was measured using standardized insertion and deletion metrics across all benchmark datasets using a consistent evaluation protocol, ensuring objective and comparable assessment of explanation faithfulness across different LLM architectures. The proposed framework also achieved higher clarity and trust ratings in user studies while introducing less than 25% additional computational overhead. Case studies in sentiment analysis and question answering further demonstrate that hybrid explanations produce precise, intuitive reasoning paths that outperform existing baselines. The main contributions are: (1) a multi-method pipeline that reconciles the trade-off between faithfulness and interpretability; (2) a human-centered evaluation showing hybrid explanations are more trustworthy than single techniques; and (3) an efficiency-aware design indicating the potential suitability of the proposed framework for practical applications in domains such as healthcare, finance, and law. By aligning methodological rigor with societal and regulatory demands, this study advances both the practice and theory of explainable AI, positioning hybrid XAI as a pathway toward responsible LLM adoption.

PLoS ONEVol. 21(9)
Hindustan Institute of Technology and Science (IN)
Openalex Percentile: Top 8%
Explainable Artificial Intelligence (XAI)
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.

Bridging trust and performance in intelligent systems: Hybrid explainable AI approaches for interpreting large language models — Tamije Selvy P., Arul Selvam P. · PLoS ONE (2026) | TGRS Research Map | TGRS