What does it mathematically mean for an AI-generated claim to be "true", "justified", and "trustworthy"?
Summary
A researcher describes a project to mathematically formalize truth, justification, and trustworthiness of AI-generated claims, seeking input on formal methods, logic, and probability theory for building a 'Trust Engine'.
Similar Articles
What would actually make you trust an AI? Not "it sounds right," but trust it the way you trust a person or an institution?
A discussion exploring what specific conditions (transparency, verifiable track record, persistent identity, accountability) would make people trust AI systems as they trust humans or institutions, rather than just accepting them as tools.
Engineering Trustworthy Agentic AI for Critical Systems
This survey proposes a trustworthiness model for agentic AI in critical engineering systems, covering safety, robustness, transparency, accountability, and security across domains like power systems and autonomous vehicles.
Truth for Believable AI: Expressed Doubt, Provenance, and Belief Revision as an Engineerable Stance
The paper introduces and evaluates a behavior layer for conversational agents that enables expressed doubt, provenance-aware assertions, and belief revision to improve truthfulness in AI systems.
Epistemic Trustworthiness in Generative AI: A Normative Framework for Warranted Reliance in High-Stakes Workflows
This paper proposes a normative framework for when reliance on generative AI outputs is epistemically warranted, based on three conditions: epistemic humility, epistemic access, and resistance to epistemic injustice. It analyzes real-world cases in legal, medical, and hiring contexts.
Verifiable AI inference
The article discusses the concept of verifiable AI inference, exploring methods like trusted attestation and cryptographic proofs to ensure the authenticity and provenance of AI-generated outputs without rerunning the model.