Artificial IntelligencearXiv — cs.LGTue, May 19, 2026, 4:00 AMNeutral

Could Large Language Models work as Post-hoc Explainability Tools in Credit Risk Models?

A recent study evaluated the potential of large language models (LLMs) as post-hoc explainability tools for credit risk models, specifically analyzing their effectiveness in generating human-readable narratives from model-based explanations. The research utilized a LendingClub dataset and compared outputs from LLMs like GPT-4-turbo, Claude-Sonnet-4.5, and Gemini-2.5-Flash against traditional attribution methods.

WPN Brief

  • What Happened

    A recent study evaluated the potential of large language models (LLMs) as post-hoc explainability tools for credit risk models, specifically analyzing their effectiveness in generating human-readable narratives from model-based explanations. The research utilized a LendingClub dataset and compared outputs from LLMs like GPT-4-turbo, Claude-Sonnet-4.5, and Gemini-2.5-Flash against traditional attribution methods.

  • Why It Matters

    The findings indicate that while LLMs can replicate feature-importance rankings under controlled conditions, their autonomous explanations lack alignment with established methods, suggesting a limited role in formal credit risk governance.

  • The Bigger Picture

    This development highlights ongoing discussions about the capabilities and limitations of LLMs in various applications, including the need for human-centered approaches that prioritize user values and the challenges faced by LLMs in maintaining coherence and reliability across different contexts.

Ask WPN AI

Related Reports

More coverage on this story

10 reports across the wire

arXiv — cs.CL
May 18

Large Language Models Could Be Rote Learners

A recent study highlights that Large Language Models (LLMs) may exhibit rote learning behaviors, particularly when evaluated through benchmark tests that are susceptible to contamination. This research indicates that LLMs can achieve inflated performance on memorized benchmarks, complicating the assessment of their genuine capabilities.

Artificial Intelligenceneutral
arXiv — cs.CL
May 12

Decomposing and Steering Functional Metacognition in Large Language Models

Recent research has proposed that large language models (LLMs) possess a decomposable space of functional metacognitive states, which include factors like evaluation awareness and self-assessed capability. This study utilizes residual stream analysis to demonstrate that these states can be decoded from internal activations, revealing distinct layer-wise profiles.

Artificial Intelligenceneutral
arXiv — cs.LG
May 11

Beyond Confidence: Rethinking Self-Assessments for Performance Prediction in LLMs

A recent study has proposed a multidimensional framework for self-assessment in Large Language Models (LLMs), moving beyond traditional confidence metrics to include various appraisal dimensions such as effort and ability. This approach aims to enhance the reliability of performance predictions across multiple tasks and models.

Artificial Intelligenceneutral
arXiv — cs.CL
May 11

Reflections and New Directions for Human-Centered Large Language Models

Recent advancements in Large Language Models (LLMs) emphasize the need for a human-centered approach in their development, as outlined in a new framework that integrates insights from Natural Language Processing, Human-Computer Interaction, and responsible AI. This framework advocates for addressing human priorities throughout the entire model development process, rather than just post-training.

Artificial Intelligencepositive
arXiv — cs.CL
May 12

Large Language Models as Students Who Think Aloud: Overly Coherent, Verbose, and Confident

A recent study evaluated large language models (LLMs), particularly GPT-4.1, as novice learners in AI-based tutoring systems. The research analyzed 630 think-aloud utterances from students tackling multi-step chemistry problems, revealing that while LLMs generate fluent responses, their reasoning tends to be overly coherent, verbose, and less variable compared to human learners.

Artificial Intelligenceneutral
arXiv — cs.CL
May 11

Why do Large Language Models Fail in Low-resource Translation? Unraveling the Token Dynamics of Large Language Models for Machine Translation

Recent research has systematically analyzed the failure modes of Large Language Models (LLMs) in machine translation, revealing that non-English-centric language pairs consistently yield lower COMET scores compared to English-centric ones. The study introduces the Token Activation Rate (TAR) as a metric to assess how effectively models utilize language-specific tokens during translation.

Artificial Intelligenceneutral
arXiv — cs.LG
May 11

Evaluating Large Language Models in Scientific Discovery

arXiv:2512.15567v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly applied to scientific research, yet prevailing science benchmarks probe decontextualized knowledge and overlook the iterative reasoning, hypothesis generation, and observation interpretation that drive scientific discovery. We introduce a scenario-grounded benchmark that evaluates LLMs across biology, chemistry, materials, and physics, where domain experts define research projects of genuine interest and decompose them into modular research scenarios from which vetted questions are sampled. The framework assesses models at two levels: (i) question-level accuracy on scenario-tied items and (ii) project-level performance, where models must propose testable hypotheses, design simulations or experiments, and interpret results. Applying this two-phase scientific discovery evaluation (SDE) framework to state-of-the-art LLMs reveals a consistent performance gap relative to general science benchmarks, diminishing return of scaling up model sizes and reasoning, and systematic weaknesses shared across top-tier models from different providers. Large performance variation in research scenarios leads to changing choices of the best performing model on scientific discovery projects evaluated, suggesting all current LLMs are distant to general scientific "superintelligence". Nevertheless, LLMs already demonstrate promise in a great variety of scientific discovery projects, including cases where constituent scenario scores are low, highlighting the role of guided exploration and serendipity in discovery. This SDE framework offers a reproducible benchmark for discovery-relevant evaluation of LLMs and charts practical paths to advance their development toward scientific discovery.

Artificial Intelligence
arXiv — cs.CL
May 11

Are LLM Agents Behaviorally Coherent? Latent Profiles for Social Simulation

A recent study investigates the behavioral coherence of Large Language Models (LLMs) as potential substitutes for human participants in research, focusing on their consistency across different experimental settings. The research aims to reveal latent profiles of these agents and assess their conversational behaviors in alignment with expected human responses.

Artificial Intelligenceneutral
arXiv — stat.ML
May 14

Uncovering Symmetry Transfer in Large Language Models via Layer-Peeled Optimization

A recent study published on arXiv investigates the geometric structures induced in large language models (LLMs) through a constrained layer-peeled optimization approach. This method treats the output projection matrix and last-layer context embeddings as optimization variables, revealing that symmetries in target distributions are transferred to the model's global minimizers.

Artificial Intelligenceneutral
arXiv — cs.LG
May 15

Rethinking Layer Relevance in Large Language Models Beyond Cosine Similarity

Recent research has highlighted the limitations of cosine similarity as a measure of layer relevance in large language models (LLMs), suggesting that it often fails to accurately reflect the impact of layer removal on model performance. The study proposes a new metric that could provide a more reliable assessment of layer importance, which is crucial for enhancing the interpretability and optimization of LLM architectures.

Artificial Intelligenceneutral

Apps

Useful picks

Explore all apps

Articles

Continue Reading

arXiv — cs.CVArtificial Intelligenceyesterday

Generation Models Know Space: Unleashing Implicit 3D Priors for Scene Understanding

Recent advancements in artificial intelligence have led to the introduction of VEGA-3D, a framework that repurposes pre-trained video diffusion models to enhance scene understanding by leveraging implicit 3D priors. This development addresses the limitations of existing multimodal large language models (MLLMs) that struggle with spatial reasoning and geometric dynamics.

arXiv — cs.CLArtificial Intelligenceyesterday

Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability

Recent research has explored the emergence of moral biases, specifically the Knobe effect, in finetuned large language models (LLMs). The study revealed that these biases are not only learned during the finetuning process but can also be localized to specific layers within the model, allowing for targeted interventions to mitigate their effects without the need for retraining.

arXiv — cs.CLArtificial Intelligenceyesterday

Code-MUE: Measuring Code LLMs' Uncertainty through Execution-based Semantic Interaction Graphs

A new framework named Code-MUE has been introduced to measure the uncertainty of Code Large Language Models (LLMs) through execution-based Semantic Interaction Graphs. This approach addresses the limitations of existing uncertainty estimation methods, which struggle with closed-source models and the unique fragility of code.

arXiv — cs.CLArtificial Intelligenceyesterday

Digital Pantheon: Simulating and Auditing Coalition Formation with LLM Agents

The recent study titled 'Digital Pantheon: Simulating and Auditing Coalition Formation with LLM Agents' explores the complexities of political coalition formation, utilizing Large Language Models (LLMs) to simulate negotiations based on ideological alignment and policy objectives. The framework combines Supervised Fine-Tuning, Direct Preference Optimization, and Retrieval-Augmented Generation to create partisan agents that adhere to political manifestos, operationalized through the 2019 Flemish election case study.

arXiv — cs.LGArtificial Intelligenceyesterday

Bifocal Attention: Harmonizing Geometric and Spectral Positional Embeddings for Algorithmic Generalization

A recent study introduces Bifocal Attention, a new architectural paradigm designed to enhance the capabilities of Large Language Models (LLMs) by harmonizing geometric and spectral positional embeddings. This approach addresses the limitations of Rotary Positional Embeddings (RoPE), which struggle with long-range recursive logic due to a fixed geometric decay.

arXiv — cs.CLArtificial Intelligenceyesterday

An MLIR-Based Compilation Method for Large Language Models

A new compilation method based on MLIR (Multi-Level Intermediate Representation) has been introduced for deploying Large Language Models (LLMs) on specialized hardware, addressing challenges in model importation and memory-efficient scheduling during inference. The method utilizes two operator dialects, TopOp for high-level graph representation and TpuOp for hardware-specific optimizations.

arXiv — cs.CLArtificial Intelligenceyesterday

Frontier AI performance across the business disciplines: a case-grounded benchmark of knowledge work and analytical reasoning

A new benchmark called BusinessCaseBench has been introduced to evaluate the performance of large language models (LLMs) in analytical reasoning and knowledge work, addressing a significant gap in existing AI benchmarks that focus primarily on factual recall and narrow question answering. This benchmark aims to assess the ability of AI systems to synthesize complex information and exercise judgment in uncertain environments, reflecting the skills required in white-collar professions.

arXiv — cs.CVArtificial Intelligenceyesterday

Instance-Enriched Semantic Maps for Visual Language Navigation

A new framework called Instance-Enriched Semantic Maps has been proposed to enhance Visual Language Navigation (VLN), enabling agents to navigate complex environments by following natural language instructions. This framework addresses limitations in existing systems by incorporating instance-level object details and robust query processing, which are crucial for reliable navigation in intricate indoor settings.