Everyone verifies the agent. Almost no one verifies the claim. I built a trust layer that grades every transmitter in a multi-agent chain

Reddit r/AI_Agents Papers

Summary

作者发布了ISNAD框架,借鉴约1400年历史的伊斯兰学术验证方法论,为多智能体AI链中的每个传递者打分,以验证AI生成言论的真实性和独立佐证。

Built this myself and just put it on arXiv — sharing here because this sub is exactly the people who'll poke the right holes in it. Here's the thing about agent chains: one answer moves through a scraper, an extractor, a few models, a synthesizer. Some links are reliable, some aren't. When they fail, they fail silently — a confident, fluent answer that's quietly wrong. We're all racing to verify the agent's identity, permissions, access. Almost no one's verifying the claim: whether what it said is true and independently corroborated. So I took a ~1,400-year-old methodology built for precisely this. Islamic scholars verifying transmitted statements graded every claim by its chain of transmitters (isnād), scored each transmitter on integrity and precision (rijāl), treated a chain as only as strong as its weakest link, let independent chains raise confidence, and judged the message separately from its chain. I rebuilt it as a claim-level trust layer for multi-agent AI. It's called ISNAD. The failures are in the paper too — validated mechanisms and not-yet-validated ones, spelled out. Honesty is kind of the whole point of a trust framework. Would love the disagreement as much as the agreement.
Original Article

Similar Articles