explainability

Tag

Cards List
#explainability

Mathematical Principles and Experimental Discoveries of the Emergence of Symbolic Patterns in Artificial Neural Networks

arXiv cs.LG ↗ · 2026-08-10 Cached

This paper proves that across a broad class of ANNs, inference logic can be reformulated as sparse symbolic interactions, supported by mathematical criteria and extensive experiments, offering novel insights into explainability and generalization.

0 favorites 0 likes
#explainability

Beyond Attention: Signed Integrated Gradients Attribution in a BiomeGPT-Style Microbiome Transformer

arXiv cs.LG ↗ · 2026-08-10 Cached

This paper proposes using signed, fusion-aware Integrated Gradients for attributing predictions in feature-tokenized transformers like BiomeGPT, overcoming limitations of CLS attention weights and revealing disease-supporting versus protective microbial signals.

0 favorites 0 likes
#explainability

@UnTalNixon_exe: The biggest problem with AI agents isn't that they hallucinate. It's that no one can explain why they made a decision. …

X AI KOLs Timeline ↗ · 2026-08-09 Cached

Semantica is an open-source tool that turns any data into a Context Graph, logging every AI agent decision with its full causal chain and W3C PROV-O provenance, enabling deterministic explainability for finance, healthcare, and government.

0 favorites 0 likes
#explainability

THBKG: A Temporal Biomedical Knowledge Graph for Decision-Aligned Clinical Advancement Prediction

arXiv cs.LG ↗ · 2026-08-07 Cached

This paper introduces THBKG, a temporal heterogeneous biomedical knowledge graph covering 110k entities and 11.1M edges with yearly evidence timestamps, designed to predict whether target-disease pairs entering Phase II trials advance to Phase III using only evidence available at the decision time. The graph-based approach outperforms direct-evidence baselines, especially for pairs lacking direct evidence, and the authors release it as a continually updated resource.

0 favorites 0 likes
#explainability

Beyond Feature Importance: A Comparative Analysis of Pattern Detection Methods in Cluster Interpretation

arXiv cs.LG ↗ · 2026-08-07 Cached

This paper presents a comparative evaluation of post-hoc analysis methods (Random Forest surrogate, LIME, PCA) for detecting structured patterns in clustering results, using synthetic datasets with injected patterns. It finds that none of the methods consistently detects all pattern types, highlighting a gap in existing explainability tools.

0 favorites 0 likes
#explainability

An Explainable LLM Agent Layer for Open-World Anomaly Detection in Oil Wells

arXiv cs.LG ↗ · 2026-08-06 Cached

This paper proposes an explainable LLM agent layer placed downstream of an open-world learning pipeline for oil well anomaly detection, using the Qwen3.5-397B-A17B model to provide natural-language justifications and novelty naming on the 3W dataset.

0 favorites 0 likes
#explainability

AI agents have never been so explainable until now, with GraphARC!

Reddit r/AI_Agents ↗ · 2026-08-02

GraphArc is an open-source tool that visualizes AI agent workflows as interactive, real-time graphs, enabling users to inspect, debug, and approve agent actions before execution to make agentic AI more explainable and controllable.

0 favorites 0 likes
#explainability

FADEx: Feature Attribution and Distortion-based Explanation of Dimensionality Reduction

arXiv cs.LG ↗ · 2026-07-31 Cached

FADEx introduces a novel local per-instance feature attribution method for explaining dimensionality reduction techniques, using Taylor expansion and Singular Value Decomposition to provide model-agnostic explanations and distortion analysis.

0 favorites 0 likes
#explainability

Position, Not Provenance: Separating Reasoning Mediation from Sycophancy in Medical Vision-Language Models

arXiv cs.LG ↗ · 2026-07-31 Cached

This paper introduces CoT-Mediate, a behavioral framework to test whether chain-of-thought reasoning in medical vision-language models actually drives predictions or merely decorates them. Auditing LLaVA-Med and MedGemma on VQA-RAD, it finds that how reasoning is injected (prefix-forcing vs re-prompting) and the attributed source (self vs expert) significantly affect model faithfulness and sycophancy.

0 favorites 0 likes
#explainability

TraceCoder: Explainable and Auditable Code Generation with Position-Key Snippet Versioning

arXiv cs.AI ↗ · 2026-07-31 Cached

This paper presents TraceCoder, a code generation system that records and visualizes the repair history of AI-generated code at snippet granularity, enabling explainable and auditable auditing of LLM-based coding agents.

0 favorites 0 likes
#explainability

Automorphism-Induced Non-Canonicity in Top-k Explanations of Graph Neural Networks

arXiv cs.LG ↗ · 2026-07-30 Cached

This paper identifies a fundamental issue in top-k explanations for graph neural networks: automorphisms in input graphs cause non-unique explanations, as the model cannot distinguish symmetric elements. The authors provide a criterion to detect such arbitrariness and verify it using automated reasoning in Lean 4, showing the problem is widespread in molecular datasets.

0 favorites 0 likes
#explainability

Open-source tabular model validation toolkit TanML needs feedback [D]

Reddit r/MachineLearning ↗ · 2026-07-29

TanML is an MIT-licensed automated model-validation toolkit for tabular machine-learning models, designed for regulated environments. The developers seek feedback on its features and reports.

0 favorites 0 likes
#explainability

The Failures of Marginal Influence-Based Attribution Methods for Global Time Series Explanations

arXiv cs.LG ↗ · 2026-07-21 Cached

This paper proves that existing marginal influence-based attribution methods fundamentally fail to capture the conditional dependency structure of time series models, and proposes DAG-faithfulness as a new criterion for faithful explanations.

0 favorites 0 likes
#explainability

Robust Explanations for User Trust in Enterprise NLP Systems

arXiv cs.CL ↗ · 2026-07-20 Cached

This paper proposes a unified black-box robustness evaluation framework for token-level explanations in enterprise NLP, comparing encoder (BERT, RoBERTa) and decoder (Qwen, Llama) models. It finds decoder LLMs produce substantially more stable explanations, with stability improving with scale, and provides a cost-robustness tradeoff curve for pre-deployment model selection.

0 favorites 0 likes
#explainability

From Plausible to Actionable: A Position on LLM Self-Explanations

arXiv cs.CL ↗ · 2026-07-20 Cached

This position paper argues that LLM self-explanations can be plausible, questionably faithful, but highly actionable, and proposes evaluation guidelines beyond traditional metrics.

0 favorites 0 likes
#explainability

Improving Molecular Property Prediction in Small Language Models Using Graph-based Tools

arXiv cs.AI ↗ · 2026-07-16 Cached

This paper proposes a Context-Augmented Prompting framework that uses a GNN expert model to provide predictive hints and explanatory subgraphs to improve molecular property prediction in small language models. Experiments on MUTAG and Tox21 show accuracy gains of up to 74% over SMILES-only baselines.

0 favorites 0 likes
#explainability

Evidence-Backed Video Question Answering

Hugging Face Daily Papers ↗ · 2026-07-13 Cached

This paper introduces Evidence-Backed Video Question Answering (E-VQA), a new task requiring models to output both semantic answers and precise spatio-temporal evidence like tracked object segmentation masklets. The authors create a human-verified benchmark and a scalable training dataset, showing significant improvements over baselines.

0 favorites 0 likes
#explainability

Towards the Explainability of Temporal Graph Networks via Memory Backtracking and Topological Attribution

arXiv cs.LG ↗ · 2026-07-10 Cached

This paper introduces MemExplainer, a method to explain predictions of Temporal Graph Networks (TGNs) by attributing contributions through topology attribution trees and memory backtracking trees, using Layer-wise Relevance Propagation (LRP) for faithful explanations.

0 favorites 0 likes
#explainability

@snowboat84: https://x.com/snowboat84/status/2075374060637503560

X AI KOLs Timeline ↗ · 2026-07-10 Cached

This article provides a systematic and comprehensive overview of AI explainability, covering its needs (debugging, compliance, safety), classic methods, and cutting-edge challenges, emphasizing that faithful explanations are more important than plausible ones.

0 favorites 0 likes
#explainability

Large Language Models (LLMs) and Generative AI in Cybersecurity and Privacy: A Survey of Dual-Use Risks, AI-Generated Malware, Explainability, and Defensive Strategies

arXiv cs.CL ↗ · 2026-07-09 Cached

A comprehensive survey examining the dual-use risks and benefits of LLMs and generative AI in cybersecurity, covering AI-generated malware, defensive strategies, and explainability, with case studies from major platforms.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback