arXiv ScienceSearch

arXiv · 2505.20305

Making Sense of the Unsensible: Reflection, Survey, and Challenges for XAI in Large Language Models Toward Human-Centered AI

Abstract

As large language models (LLMs) are increasingly deployed in sensitive domains such as healthcare, law, and education, the demand for transparent, interpretable, and accountable AI systems becomes more urgent. Explainable AI (XAI) acts as a crucial interface between the opaque reasoning of LLMs and the diverse stakeholders who rely on their outputs in high-risk decisions. This paper presents a comprehensive reflection and survey of XAI for LLMs, framed around three guiding questions: Why is explainability essential? What technical and ethical dimensions does it entail? And how can it fulfill its role in real-world deployment? We highlight four core dimensions central to explainability in LLMs, faithfulness, truthfulness, plausibility, and contrastivity, which together expose key design tensions and guide the development of explanation strategies that are both technically sound and contextually appropriate. The paper discusses how XAI can support epistemic clarity, regulatory compliance, and audience-specific intelligibility across stakeholder roles and decision settings. We further examine how explainability is evaluated, alongside emerging developments in audience-sensitive XAI, mechanistic interpretability, causal reasoning, and adaptive explanation systems. Emphasizing the shift from surface-level transparency to governance-ready design, we identify critical challenges and future research directions for ensuring the responsible use of LLMs in complex societal contexts. We argue that explainability must evolve into a civic infrastructure fostering trust, enabling contestability, and aligning AI systems with institutional accountability and human-centered decision-making.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Francisco Herrera. 2025-05-18. Making Sense of the Unsensible: Reflection, Survey, and Challenges for XAI in Large Language Models Toward Human-Centered AI. https://arxiv.org/abs/2505.20305

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Algorithmic Shortlisting in Participatory Budgeting

Participatory budgeting is a democratic innovation that allows citizens to propose and vote on public investment projects. To help organizers manage large volumes of submissions, we design and test privacy-preserving methods for algorithmic shortlisting. These algorithms predict which projects are likely to be funded using only project features and anonymous historical voting data. We demonstrate the limitations of a naive approach that uses a large language model to rank projects based on past success and propose a vote-based pipeline that enables state-of-the-art LLMs to perform on par with classical machine learning. Our findings indicate that user preferences in participatory budgeting are stable enough to allow algorithmic shortlisting to approximate an initial selection of projects effectively.

cs.CY

Human Resilience in the AI Era -- What Machines Can't Replace

AI is changing work and decision making faster than many institutions can adapt their operating practices. We argue that this adaptation gap makes human resilience a core capability for the AI era. We define resilience as the capacity to absorb disruption while preserving effective action and human agency around core purposes. The framework operates at three interacting levels. Psychological resilience keeps a person goal-directed under stress. Social resilience makes trusted support and correction available across a group. Organizational resilience turns detected problems into learning and recovery. We connect established resilience and technostress research with direct AI-in-the-loop experiments. General resilience is trainable, while AI-specific causal evidence is still emerging. Direct AI studies show that assistance can raise productivity and spread expertise. Other experiments show improved expressed empathy and more calibrated reliance. We translate these findings into a practical agenda for AI education, workplace design, governance, and evaluation. The central proposal is socio-technical: structural safeguards define the operating boundary, while resilient people and institutions provide adaptive capacity when conditions change.

cs.CY

Toward a Time-Aware Assessment Framework for the Carbon Cost of AI-Enabled Decarbonization

AI is increasingly used to support decarbonization decisions across the built environment, yet the development, training, and use of AI consume energy and induce CO2e emissions. However, existing assessments often report physical-system savings while omitting AI-side emissions. Moreover, they rarely account for the mismatch between when AI costs occur and when decarbonization benefits materialize, which may be substantial for infrastructure-scale projects. To address these issues, we present a time-aware assessment framework that models avoided emissions and AI-induced emissions as discrete-time streams over a finite time horizon. In demonstrating this process, we seek to show that time-aware assessment can support temporal decision-making, identify cases in which accounting for time value of carbon can change preferred rankings relative to time-invariant totals, and explore how decisions may vary with slightly different governance priorities. Using four representative interventions with intentionally different temporal profiles (multi-project low-carbon concrete design support, AI-assisted construction logistics, agentic HVAC control, and predictive maintenance), we demonstrate how discounting can change preferred rankings relative to time-invariant totals and supports ranking sensitivity analysis, discounted payback screening, and break-even discount-rate analysis. We also provide decision guidelines that support go/no-go screening, timing decisions, and minimum "bang-for-your-buck" thresholds. Ultimately, this work contributes a lightweight framework for deciding whether and when to deploy AI-enabled interventions for decarbonization under explicit time preference.

cs.CY